Module 4 · Evaluation
RAG precision and hallucination analysis
Paste a system prompt, the retrieved context and the expected and actual answers. We score retrieval precision, grounding and data readiness, and flag every claim the context doesn't support.
- About 2 minutes
- Hallucination risk
- 5 quality scores
- Unsupported claims
Your diagnostic will appear here
We've loaded a sample with retrieval noise, a duplicate chunk and one invented claim. Press Analyze to see what we catch.
- Five scores covering retrieval, grounding and data hygiene
- Every claim in the answer that the context doesn't support
- Concrete remediation notes for your pipeline