Lyra
Learn
AI Learning Platform
π Midnight
π Comfort
π₯ Ember
π Paper
β Contrast
Exams
Sign in to track progress
β Back to the lesson
Module 13 Β· Quiz
Evaluation Sets and Regression Testing
1. What is the main purpose of an evaluation set in the context of AI systems?
To track user complaints
To determine if changes will affect quality before user exposure
To measure performance metrics after deployment
To generate new training data for the model
2. What role do golden questions play in an evaluation set?
They represent unimportant questions to avoid
They help define model architecture
They guide the expected behavior of a system in response to specific queries
They should be removed from the evaluation process
3. What happens when a regression test fails during continuous integration?
The test is ignored
The production release proceeds anyway
The build fails and the changes are not deployed to production
The model receives automatic retraining
4. How does the evaluation record enhance the evaluation sets?
By storing outdated golden questions
By providing new metrics that do not relate to production performance
By reflecting real-world failures and refusals to update the evaluation set
By disconnecting tests from production data
Submit answers
Continue: The Azure AI Landscape β
Review this lesson
Retake quiz