Secure your spot in AI Evals in Practice. Enrollment closes in just 3 days! If you’ve been thinking about joining, now’s the time.
ByteByteGo has teamed up with Manjeet Singh, Senior Director at Salesforce, to bring you this live, hands-on course on building reliable evaluation systems for production AI agents.
You’ll learn how to:
Design evals for quality, safety, reliability, cost, and latency
Build and validate LLM-as-a-Judge systems
Red-team agents for prompt injection and jailbreaks
Create meaningful eval datasets from real and synthetic data
Evaluate tool use, RAG, multi-step execution, and multi-agent handoffs
Run evals in CI/CD and production to catch regressions and drift
Turn failures into permanent regression tests


