Your question is Spark Memory and OOM Prevention. Take a moment with it on the right.
Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).
How does Spark manage memory, and what strategies do you use to prevent Out-Of-Memory (OOM) errors during large-scale joins?