Dataford
Interview QuestionsInterview GuidesExperiencesMock InterviewsPricing
Get started
Dataford
Popular roles
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Browse
Browse All RolesEvery role hub, from analyst to MLBrowse All CompaniesCompany-specific interview loopsAll Interview GuidesThe full guide library
Top questions by role
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Top questions by skill
SQLPythonStatisticsMachine LearningA/B TestingSystem DesignGenerative AIProduct SenseMetricsBehavioral
Browse all questions →Try a mock interview
Experiences
Practice
Mock InterviewsTimed interview simulations with feedbackSuccess PathYour 6-week structured planModulesCurated lessons by topicWebinarsTalks from ex-Big Tech data leadsPlaygroundA free-form scratch editor
Learn
BlogInterview strategy and career adviceTech Job Market ReportHiring trends across data and AI rolesFor UniversitiesDataford for career centersAbout DatafordWho we are and how we build
Pricing
Build my plan
SQL: Duplicate Rows and Top Percent
00:00
5 left

SQL: Duplicate Rows and Top Percent

MediumSQL · PostgreSQL

Problem

Altimetrik's Workforce Directory team needs a validation query for employee extracts. Using the tables below, write PostgreSQL SQL that produces both a deliberately duplicated employee row and the first 50% of eligible workforce records.

Requirements

  1. Join employees to departments and retain only active employees with a matching department.
  2. Return the row for employee_id = 104 exactly twice using UNION ALL.
  3. Return the first half of eligible employees, ordered by ascending employee_id. Use NTILE(2) so the query remains data-driven.
  4. Add an operation column with duplicated or first_half, and sort the final result by operation and employee ID.

Schema

employees
ColumnTypeDescription
employee_idPKINTUnique employee identifier
employee_nameVARCHAR(100)Employee full name
department_idINTReferenced department identifier
employment_statusVARCHAR(20)Current employment status
hire_dateDATEDate the employee joined
emailVARCHAR(150)Work email address
departments
ColumnTypeDescription
department_idPKINTUnique department identifier
department_nameVARCHAR(100)Department display name
Tablesemployeesdepartments
Interviewer

Your question is SQL: Duplicate Rows and Top Percent. Start with the requirements and the two tables in the Question tab.

Run and submit as often as you like. When you're ready, talk me through your approach or go straight to the code.

You need to log in / sign up to run or submit.
CodePostgreSQL
Sign up free to run your codeLog inLn 1
Run your query to see results here.