All 227 Dependencies
Every package agenta depends on, ranked by repo health score.
Agenta is a comprehensive open-source LLMOps platform built for engineering and product teams who need to move from ad-hoc prompt experimentation to systematic, production-grade LLM application development. It brings together prompt management, automated evaluation, human annotation, and full-stack observability in a single platform, eliminating the scattered tooling that slows down AI teams.
At its core, Agenta provides an interactive playground where teams can compare prompts side-by-side against test cases, version configurations with branching and environment promotion, and collaborate with subject matter experts without requiring code changes. The platform supports 50+ LLM providers out of the box including OpenAI, Anthropic, and custom self-hosted models.
For evaluation, Agenta offers flexible test set management sourced from production logs, playground experiments, or uploaded CSVs — paired with over 20 pre-built evaluators, LLM-as-a-judge scoring, and a UI and API for both technical and non-technical reviewers. The human annotation module supports structured feedback workflows that close the loop between production observations and iterative improvement.
On the observability side, Agenta ingests OpenTelemetry traces natively and provides pre-built integrations for popular LLM frameworks and providers. Teams get full trace trees, cost and latency breakdowns, and the ability to save any production span as a test case — creating a tight feedback loop between what happens in production and what gets evaluated in development.