DeepYardDeepYard

promptfoo vs RAG Failure Diagnostics Clinic

Side-by-side comparison with live GitHub signals. Last updated August 17, 2026.

p

promptfoo

Test and evaluate LLM prompts and agents — 11K+ stars

OSSFree
24.3Ktoday317
R

RAG Failure Diagnostics Clinic

Diagnose and fix common RAG pipeline failure modes

OSSFree
132.9Ktoday95
MetricpromptfooRAG Failure Diagnostics Clinic
GitHub Stars24.3K132.9K
Contributors31795
Last CommitAug 17, 2026Aug 17, 2026
Open Issues50914
Licenseopen-sourceopen-source
Pricingopen-sourceopen-source
Free TierYesYes
Categorydev-toolsdev-tools
TrendingNoNo

Shared Tags

evaluation

Only in promptfoo

testingred-teamingsecurityci-cdopen-source

Only in RAG Failure Diagnostics Clinic

ragdebuggingdiagnosticspython

About promptfoo

promptfoo is an open-source tool for testing, evaluating, and red-teaming LLM applications. Run automated evaluations across multiple models and prompts, compare outputs side-by-side, detect regressions, and test for security vulnerabilities. Supports custom assertions, CI/CD integration, and model-graded evaluations.

View full listing

About RAG Failure Diagnostics Clinic

A diagnostic tool that identifies why RAG pipelines produce poor results. It tests for common failure modes: irrelevant retrieval, missing context, hallucination over context, chunking issues, and embedding quality problems. Provides a structured report with specific fix recommendations for each detected issue. Essential for debugging production RAG systems. Part of the awesome-llm-apps collection.

View full listing