Your question is Troubleshooting Kubernetes Connectivity. Take a moment with it on the right.
Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).
You are given a Kubernetes cluster with connectivity and scheduling problems. How would you figure out what is wrong and validate the fix?
Explain a practical investigation using Kubernetes APIs, node and pod events, scheduler and kubelet logs, service networking checks, and Prometheus or Grafana metrics. Cover how you would distinguish control-plane, node, CNI, DNS, resource, taint, affinity, and admission failures. Describe the evidence you would collect, the least risky remediation sequence, rollback criteria, and how you would validate recovery without masking intermittent failures.