Your question is Evaluate Model Bias Rigorously. Take a moment with it on the right.
Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).
Your team wants a clear way to assess whether a model treats different groups unfairly. You need to separate real bias signals from noise, and explain how you would test for disparities in model behavior across groups.
How would you evaluate whether an AI model is biased, and what metrics or tests would you use?