Test RAG apps before launch
Test retrieval grounding, source drift, and instruction conflicts in doc-backed assistants.
Your sources are not enough if the AI follows the wrong instructions.
What can go wrong
- Answers overreach what retrieved docs support.
- Hostile content in sources hijacks behavior.
- Stale knowledge produces confident wrong answers.
Why happy-path testing misses this
- Indexing checks do not test adversarial questions or poisoned sources.
What Agnostics tests
- Retrieval drift
- Prompt injection via retrieved content
- Hallucination pressure
Useful finding example
Retrieval drift finding: answer cited wrong section for warranty terms.
Release Gate
Monitor or Fix when grounding breaks on launch-critical knowledge.
Questions
Does Agnostics replace evals on my corpus?
No. It pressure-tests live behavior, including how retrieval and instructions interact.