StudyToCert

All certifications / Azure AI Fundamentals / Lessons

Microsoft Certified: Azure AI Fundamentals AI-900 · Domain 5: Generative AI workloads on Azure

Evaluating generative AI apps: groundedness, relevance, fluency and safety evaluations, and red teaming

▶ Watch the overview video

Last reviewed September 25, 2026 · Leer en español

Traditional ML models are evaluated with metrics like accuracy or RMSE against known labels. Generative AI output is open-ended text, so there is rarely one right answer to compare with. Evaluation therefore uses a mix of quality metrics, safety metrics, human review and adversarial testing, and it continues after release.

Free account

Keep reading for free

Create a free StudyToCert account to read the rest of this lesson: 5 more sections, 4 key terms, a real-world example, an exam tip and self-check questions. Every lesson, lab and practice test is free with an account.

Sign up free Log in

Study Azure AI Fundamentals for free
A week-by-week plan with every lesson, quizzes, checkpoint tests, a practice exam and hands-on labs.
Open the Azure AI Fundamentals study plan