Your question is Handle Late-Arriving Data in ETL. Take a moment with it on the right.
Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).
How would you handle late-arriving data in a daily batch ETL job?
Explain how you would detect records received after their expected processing window, update affected partitions or aggregates, and preserve idempotency. Address watermark selection, reprocessing and backfilling, deduplication, dependent downstream models, data quality validation, orchestration, and monitoring. Provide a practical design using specific technologies and include representative SQL or Python code.