Welcome to your interview.
The question is on your right: Design an ETL Pipeline for Large Datasets. Take a moment with it first.
Talk your thinking through with me if you like - when you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes). Discussion and graded submissions share your five interviewer interactions, so spend them well.
DataCorp, a financial services provider, aggregates large datasets from various internal and external sources (transaction logs, market feeds, user activity). The current batch ETL process, running nightly, struggles with data quality issues and delays in reporting, impacting business decisions. The goal is to design a robust ETL pipeline that can handle 10TB of data daily while ensuring data integrity and quality.