Dev Tools
29 tools.
Agent Zero-Trust
securityZero-trust pre-flight repo scanner for AI coding assistants: scans a codebase's instruction surface for prompt injection and supply-chain attacks before Claude Code, Cursor, Codex, or Gemini reads it. Ships with a false-negative ledger.
BentoML
framework · servingAn open-source framework and platform designed for packaging, deploying, and scaling AI/ML models and custom inference pipelines in production.
Braintrust
eval-observabilityAn AI observability and evaluation platform designed to monitor, trace, and test AI products and agents at scale.
Chroma
ragAn open-source, serverless, and scalable search infrastructure for AI products supporting vector, full-text, regex, and metadata search.
Confident AI
eval-observability · securityAn enterprise AI evaluation, red teaming, and observability platform that helps teams monitor, trace, and stress-test LLM applications.
CrewAI
frameworkOpen-source framework that orchestrates multiple autonomous AI agents through role-playing collaboration, using intuitive crew/agent/task abstractions. CrewAI Cloud and Enterprise add hosting and operational tooling.
Dex
data-engAgent-native analytics engineering toolkit: point it at your warehouse and dbt project and it learns the structure, writes transformations, and pinpoints exactly what to fix when schemas drift. Built for analytics and data engineers who want coding agents to do more of the data work.
Dify
framework · securityAn open-source LLM application development platform for building, testing, and deploying agentic workflows, RAG pipelines, and AI assistants.
Dynamo-Triton
servingAn open-source inference server designed to deploy AI models across major frameworks (TensorRT, PyTorch, ONNX, OpenVINO, Python, and RAPIDS FIL) on GPUs, CPUs, and accelerators.
Flowise
framework · eval-observabilityAn open-source, visual developer platform for building and orchestrating AI agents, chatbots, and retrieval-augmented generation workflows using a drag-and-drop interface.
Galileo AI
eval-observabilityAn enterprise-grade AI observability and evaluation platform for LLM applications, enabling developers to run offline evaluations, auto-tune metrics, and deploy optimized low-latency guardrails using compact Luna models.
Guardrails AI
framework · securityAn AI reliability platform and framework designed for building, governing, and scaling production GenAI applications. It provides runtime guardrails to detect policy violations, hallucinations, and data leakage, alongside capabilities for simulating realistic datasets and generating evaluation datasets.
Helicone
eval-observabilityAn AI gateway and LLM observability platform designed to route, debug, and analyze LLM applications.
Kastor
frameworkDeclarative language and toolchain for AI agents: define agents, tools, and prompts in HCL, then compile to a framework or manage them with plan/apply semantics on a hosted platform. Billed as "Terraform for AI agents".
Lakera
security · freeAn AI-native security platform that provides runtime protection, data leakage prevention, and red teaming for GenAI applications, agents, and workforce AI usage.
Langfuse
eval-observabilityAn open-source AI engineering platform designed for LLM observability, tracing, prompt management, evaluation, and experiments.
LangGraph
frameworkLow-level orchestration framework for stateful, long-running agents, built by LangChain. Provides state management, persistence, and controllable graph execution; used in production by Klarna, Replit, and Elastic. LangGraph Platform adds managed deployment.
LlamaIndex
framework · ragLlamaIndex provides document parsing, indexing, and orchestration workflows for AI agents and RAG pipelines. Its suite includes LlamaParse, a VLM-powered cloud parser for complex layouts, and LiteParse, a local open-source alternative for parsing PDFs, Office docs, and images.
Ollama
servingA tool that enables running open large language models locally and scaling to the cloud.
Pinecone
ragA fully managed vector database designed for building AI applications, agent memory, semantic search, and recommendation systems.
Promptfoo
securityAn automated security testing and evaluation tool designed to find and fix vulnerabilities, jailbreaks, data leaks, and model risks in LLM applications, agents, and RAG pipelines.
Qdrant
rag · data-engQdrant is a high-performance, open-source vector search engine and database written in Rust. It supports native hybrid search (dense and sparse vectors), built-in multivector support, advanced metadata filtering, and real-time indexing for AI applications like RAG and recommendation systems.
Ragas
framework · eval-observabilityAn open-source framework designed to evaluate and monitor Retrieval-Augmented Generation (RAG) systems by providing automated metrics, synthetic test data generation, and online quality tracking.
Semgrep
securityAn application security platform that combines AI reasoning with rule-based static analysis to detect, triage, and remediate vulnerabilities, dependencies, and hardcoded secrets.
Snyk
securityAn AI-powered security platform that validates AI-generated code, governs development agents, and secures AI-native applications.
TruLens
eval-observability · frameworkAn open-source evaluation and tracing framework for AI agents and LLM applications, allowing developers to measure execution flows and application quality using metrics like groundedness and context relevance.
Unstructured
data-eng · ragAn ingestion and pre-processing platform designed for GenAI that extracts, parses, chunks, and embeds unstructured data from over 64 file types into structured formats ready for AI models and analysis.
vLLM
serving · freeA high-throughput and memory-efficient LLM inference and serving engine designed for easy, fast, and cost-efficient model deployment on multiple hardware platforms.
Weaviate
rag · data-engWeaviate is an open-source vector database and AI development platform that enables developers to store, index, and search high-dimensional vectors. It includes built-in embeddings generation, a natural language Query Agent, and features for building Retrieval-Augmented Generation (RAG) and personalized AI experiences.