Your question is Data Cleaning in ETL Pipelines. Take a moment with it on the right.
Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).
You're working on a data pipeline and need to prepare raw data before it can be used for analysis or reporting. The focus is on how you clean messy inputs and structure them into something reliable for downstream use.
What methods do you use for data cleaning and preparation?