📄️ AI & LLM Testing Quick Reference
Copy-paste reference for testing AI and LLM features — what to test, assertion types, pytest and promptfoo snippets, OWASP LLM Top 10, RAG and agent checklists, metrics and thresholds, attack prompts, judge rubrics and tools.
📄️ AI & LLM Testing Best Practices
Habits that make AI feature testing trustworthy — evaluate not assert, build golden and attack sets with experts, test meaning not text, pin versions and run regression, calibrate judges, enforce agent limits in code, watch cost and data handling, keep humans accountable, and a pre-release checklist.