Skip to main content

One doc tagged with "llm"

View all tags

AI & LLM Testing Cheat Sheet

A beginner-to-advanced reference for testing AI and LLM features — why they differ, golden datasets, hallucination, prompt injection and the OWASP LLM Top 10, structured output, RAG and agents, LLM-as-judge, an evaluation harness in pytest and promptfoo, regression and AI-augmented QA.