Learn LLM Path / pillar 9 of 10
stop guess-and-tweak. Evaluation-driven development is the single biggest predictor of agent-building success (per Andrew Ng)
- Sign in to track this item
Objective vs LLM-judge; error analysis; traces KEY
Agentic AI - Module 4 (Andrew Ng) ↗articlein course
- Sign in to track this item
RAG evals (RAGAS): faithfulness, precision, recall
- Sign in to track this item
RAGAS in code (run in CI)
RAGAS Evaluation Tutorial (local, no API) - write-up ↗articleread
- Sign in to track this item
Tracing/observability in practice
LangSmith docs ↗docsdocs