Local Python tool for evaluating LLM-generated QA findings with manual validation and false-positive analysis
-
Updated
May 16, 2026 - Python
Local Python tool for evaluating LLM-generated QA findings with manual validation and false-positive analysis
Practical AI response evaluation portfolio with evidence-based scoring, worked reviews, and controlled response comparisons
Model-agnostic AI response evaluation handbook with evidence-based rubrics, calibration guides, case studies, and reusable templates
AI evaluation, prompt engineering, technical QA, and AI-assisted software testing portfolio
Practical, model-agnostic prompt engineering knowledge base for AI evaluation, technical QA, troubleshooting, and software testing
To associate your repository with the technical-qa topic, visit your repo's landing page and select "manage topics."