Ragas
Evaluation metrics for RAG pipelines and agents.
Library of reference-free metrics — faithfulness, answer relevance, context precision/recall, agent goal accuracy — plus synthetic test-set generation.
- Vendor
- Exploding Gradients
- Category
- Observability & evals
- Pricing
- Open source
- Open source
- Yes
- License
- Apache 2.0
- Platforms
- Python
- Launched
- 2023
- Website
- ragas.io
Features
- RAG metrics
- Test-set generation
- Framework integrations
Best for
- Research
- Agents
More observability & evals
- LangSmith — LangChain. Tracing, evals and monitoring for LLM apps.
- Langfuse — Langfuse. Open-source LLM engineering platform.
- Braintrust — Braintrust. The evals platform for AI products.
- promptfoo — promptfoo. Test and red-team your LLM apps.
- Arize Phoenix — Arize AI. Open-source AI observability and evaluation.
- W&B Weave — Weights & Biases. Track and evaluate LLM applications.
- Helicone — Helicone. LLM observability via a one-line proxy.
- DeepEval — Confident AI. Pytest-style unit testing for LLM outputs.