Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.
Updated 2026-10-11 21:24:49 +00:00
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
Updated 2026-10-11 21:11:56 +00:00
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.
Updated 2026-10-11 21:10:33 +00:00
🪢 Open source agent evals & observability: Trace, evaluate, and improve LLM applications with one open platform.
Updated 2026-10-11 02:54:33 +00:00
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
Updated 2026-10-05 17:17:20 +00:00
Supercharge Your LLM Application Evaluations 🚀
Updated 2026-02-24 07:47:18 +00:00