Dataford
Interview QuestionsInterview GuidesExperiencesMock InterviewsPricing
Get started
Dataford
Popular roles
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Browse
Browse All RolesEvery role hub, from analyst to MLBrowse All CompaniesCompany-specific interview loopsAll Interview GuidesThe full guide library
Top questions by role
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Top questions by skill
SQLPythonStatisticsMachine LearningA/B TestingSystem DesignGenerative AIProduct SenseMetricsBehavioral
Browse all questions →Try a mock interview
Experiences
Practice
Mock InterviewsTimed interview simulations with feedbackSuccess PathYour 6-week structured planModulesCurated lessons by topicWebinarsTalks from ex-Big Tech data leadsPlaygroundA free-form scratch editor
Learn
BlogInterview strategy and career adviceTech Job Market ReportHiring trends across data and AI rolesFor UniversitiesDataford for career centersAbout DatafordWho we are and how we build
Pricing
Build my plan
SQL And Spark Querying
00:00
5 left

SQL And Spark Querying

MediumSQL · PostgreSQL

Problem

Context

Exusia's Data Fabric team needs a repeatable way to identify the highest-paid employees while retaining department information when available. The result must handle salary ties and employees whose department reference does not match a department record.

Task

Write a PostgreSQL query that returns every employee earning the maximum salary in the employees table. Also describe how to produce the same result with a DataFrame and an RDD.

Requirements

  1. Use a CTE to calculate the overall maximum non-NULL salary.
  2. Return every employee tied at that salary, including the employee name and salary.
  3. Use a LEFT JOIN to include the employee when its department reference has no matching department.
  4. Order the result by employee name.

Schema

employees
ColumnTypeDescription
employee_idPKINTUnique employee identifier
employee_nameVARCHAR(100)Employee full name
salaryNUMERIC(12,2)Annual salary
department_idINTReferenced department identifier
job_titleVARCHAR(100)Current job title
departments
ColumnTypeDescription
department_idPKINTUnique department identifier
department_nameVARCHAR(100)Department name
Tablesemployeesdepartments
Interviewer

Your question is SQL And Spark Querying. Start with the requirements and the two tables in the Question tab.

Run and submit as often as you like. When you're ready, talk me through your approach or go straight to the code.

You need to log in / sign up to run or submit.
CodePostgreSQL
Sign up free to run your codeLog inLn 1
Run your query to see results here.