Dataford
Interview QuestionsInterview GuidesExperiencesMock InterviewsPricing
Get started
Pandas Data Manipulation Approach
00:00
5 left

Pandas Data Manipulation Approach

MediumSQL · PostgreSQL

Problem

Can you explain how you would approach a data manipulation task using Pandas?

For this SQL version, use the supplied instrument and observation data to produce a concise summary for active instruments. Include only observations from January 2025 and preserve active instruments that have no matching observations.

Output

  1. One row per active instrument, with instrument_code, instrument_name, valid_observations, average_observation, latest_observation, and observation_band.
  2. Classify averages of at least 100 as High, lower averages as Standard, and instruments without valid observations as No data.
  3. Order by instrument_code ascending.

Schema

instruments
ColumnTypeDescription
instrument_idPKINTUnique instrument identifier
instrument_codeVARCHAR(20)Instrument display code
instrument_nameVARCHAR(100)Instrument name
is_activeBOOLEANWhether the instrument is currently active
observations
ColumnTypeDescription
observation_idPKINTUnique observation identifier
instrument_idINTReferenced instrument
observed_atDATEObservation date
observed_valueDECIMAL(10,2)Measured observation value
Tablesinstrumentsobservations
Interviewer

Your question is Pandas Data Manipulation Approach. Start with the requirements and the two tables in the Question tab.

Run and submit as often as you like. When you're ready, talk me through your approach or go straight to the code.

You need to log in / sign up to run or submit.
CodePostgreSQL
You need to log in / sign up to run or submit.Ln 1
Run your query to see results here.