Your question is Monitor a Production ML Pipeline. Take a moment with it on the right.
Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).
You're working on a production ML pipeline and want clear visibility into where it slows down or becomes unstable. You need a telemetry strategy that helps you spot issues across ingestion, processing, model execution, and downstream delivery.
Describe how you would implement a robust telemetry and monitoring system to detect performance bottlenecks and sources of instability in a production ML pipeline.