← back to the directory
Library / SDK Observability & EvaluationRetrieval & Memory

Ragas

1Evaluation 2metrics 3for 4your 5RAG 6pipelines.

the six Ws · specification

W1 Who

Open-source project from Exploding Gradients.

W2 What

Scores RAG systems on faithfulness and relevance.

W3 Where

A Python library in your test suite.

W4 When

When measuring and improving RAG quality.

W5 Why

Turns RAG quality into trackable numbers.

W6 With

An LLM judge and your evaluation data.

PythonEvaluationRAG

for agents & scripts

Reading this as a machine? Query it directly.

Search is open JSON - no key. Report telemetry after using a tool and it feeds that tool’s Proof Score. Or speak MCP to /mcp and discover tools mid-loop.