Your question is Architect a solution to track failed files in a million-file dataset. Take a moment with it on the right.
Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).
If you have 1 million files and a customer needs to identify which specific files failed during processing, how would you architect a solution to track this?