Dataford
Interview QuestionsInterview GuidesExperiencesMock InterviewsPricing
Get started
Dataford
Popular roles
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Browse
Browse All RolesEvery role hub, from analyst to MLBrowse All CompaniesCompany-specific interview loopsAll Interview GuidesThe full guide library
Top questions by role
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Top questions by skill
SQLPythonStatisticsMachine LearningA/B TestingSystem DesignGenerative AIProduct SenseMetricsBehavioral
Browse all questions →Try a mock interview
Experiences
Practice
Mock InterviewsTimed interview simulations with feedbackSuccess PathYour 6-week structured planModulesCurated lessons by topicWebinarsTalks from ex-Big Tech data leadsPlaygroundA free-form scratch editor
Learn
BlogInterview strategy and career adviceTech Job Market ReportHiring trends across data and AI rolesFor UniversitiesDataford for career centersAbout DatafordWho we are and how we build
Pricing
Build my plan
Pandas Data Manipulation Approach
00:00
5 left

Pandas Data Manipulation Approach

MediumSQL · PostgreSQL

Problem

Can you explain how you would approach a data manipulation task using Pandas?

For this SQL version, use the supplied instrument and observation data to produce a concise summary for active instruments. Include only observations from January 2025 and preserve active instruments that have no matching observations.

Output

  1. One row per active instrument, with instrument_code, instrument_name, valid_observations, average_observation, latest_observation, and observation_band.
  2. Classify averages of at least 100 as High, lower averages as Standard, and instruments without valid observations as No data.
  3. Order by instrument_code ascending.

Schema

instruments
ColumnTypeDescription
instrument_idPKINTUnique instrument identifier
instrument_codeVARCHAR(20)Instrument display code
instrument_nameVARCHAR(100)Instrument name
is_activeBOOLEANWhether the instrument is currently active
observations
ColumnTypeDescription
observation_idPKINTUnique observation identifier
instrument_idINTReferenced instrument
observed_atDATEObservation date
observed_valueDECIMAL(10,2)Measured observation value
Tablesinstrumentsobservations
Interviewer

Your question is Pandas Data Manipulation Approach. Start with the requirements and the two tables in the Question tab.

Run and submit as often as you like. When you're ready, talk me through your approach or go straight to the code.

You need to log in / sign up to run or submit.
CodePostgreSQL
Sign up free to run your codeLog inLn 1
Run your query to see results here.