Dataford
Interview QuestionsInterview GuidesExperiencesMock InterviewsPricing
Get started
Dataford
Popular roles
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Browse
Browse All RolesEvery role hub, from analyst to MLBrowse All CompaniesCompany-specific interview loopsAll Interview GuidesThe full guide library
Top questions by role
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Top questions by skill
SQLPythonStatisticsMachine LearningA/B TestingSystem DesignGenerative AIProduct SenseMetricsBehavioral
Browse all questions →Try a mock interview
Experiences
Practice
Mock InterviewsTimed interview simulations with feedbackSuccess PathYour 6-week structured planModulesCurated lessons by topicWebinarsTalks from ex-Big Tech data leadsPlaygroundA free-form scratch editor
Learn
BlogInterview strategy and career adviceTech Job Market ReportHiring trends across data and AI rolesFor UniversitiesDataford for career centersAbout DatafordWho we are and how we build
Pricing
Build my plan
Optimize Slow SQL on Large Data
00:00
5 left

Optimize Slow SQL on Large Data

MediumSQL · PostgreSQL

Problem

Optimize a provided SQL query that is currently performing poorly on a large dataset for HackerRank.

Rewrite the query so it returns the same business result with less repeated work. Include only submissions from January 2025 whose user and challenge records exist.

Output

  1. One row per challenge with submissions in the date range.
  2. Columns: difficulty, challenge_id, title, total_submissions, accepted_submissions, distinct_submitters, acceptance_rate, and performance_rank.
  3. Rank challenges within difficulty by acceptance rate descending, then total submissions descending, then challenge ID ascending. Order by difficulty, rank, and challenge ID.

Schema

users
ColumnTypeDescription
user_idPKINTUnique HackerRank user identifier
usernameVARCHAR(100)Public username
countryVARCHAR(80)User country
challenges
ColumnTypeDescription
challenge_idPKINTUnique challenge identifier
titleVARCHAR(200)Challenge title
difficultyVARCHAR(30)Challenge difficulty level
submissions
ColumnTypeDescription
submission_idPKINTUnique submission identifier
user_idINTSubmitting user identifier
challenge_idINTSubmitted challenge identifier
submitted_atTIMESTAMPSubmission timestamp
statusVARCHAR(20)Submission outcome
scoreDECIMAL(5,2)Submission score
Tablesuserschallengessubmissions
Interviewer

Your question is Optimize Slow SQL on Large Data. Start with the requirements and the three tables in the Question tab.

Run and submit as often as you like. When you're ready, talk me through your approach or go straight to the code.

You need to log in / sign up to run or submit.
CodePostgreSQL
Sign up free to run your codeLog inLn 1
Run your query to see results here.