Dataford
Interview QuestionsInterview GuidesExperiencesMock InterviewsPricing
Get started
Dataford
Popular roles
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Browse
Browse All RolesEvery role hub, from analyst to MLBrowse All CompaniesCompany-specific interview loopsAll Interview GuidesThe full guide library
Top questions by role
Software EngineerData AnalystData ScientistData EngineerBusiness AnalystAI EngineerMachine Learning EngineerProduct Manager
Top questions by skill
SQLPythonStatisticsMachine LearningA/B TestingSystem DesignGenerative AIProduct SenseMetricsBehavioral
Browse all questions →Try a mock interview
Experiences
Practice
Mock InterviewsTimed interview simulations with feedbackSuccess PathYour 6-week structured planModulesCurated lessons by topicWebinarsTalks from ex-Big Tech data leadsPlaygroundA free-form scratch editor
Learn
BlogInterview strategy and career adviceTech Job Market ReportHiring trends across data and AI rolesFor UniversitiesDataford for career centersAbout DatafordWho we are and how we build
Pricing
Build my plan

Design Real-Time Fraud Risk Scoring

HardSystem Design00:00
Practice interviewer
In session
5 left
00:00

Your question is Design Real-Time Fraud Risk Scoring. Take a moment with it on the right.

Talk me through your thinking if you like. When you're confident, submit your answer and I'll grade it like a real screen (7/10 or better passes).

You need to log in / sign up to chat or submit.

Problem

Scenario

You are designing a real-time ML system for a digital payments platform that scores every card transaction for fraud before authorization. The score is used to approve, decline, or step up transactions, so the model directly affects both fraud losses and customer conversion. Fraud patterns shift quickly, labels are delayed by chargebacks and investigations, and the business wants decisions to incorporate the latest user and merchant behavior. The system must support low-latency predictions globally while remaining robust to drift, outages, and feature inconsistencies.

Scale

SignalValue
Daily active cardholders18M
Peak transaction scoring QPS45K
Average transaction scoring QPS18K
Distinct merchants9M
User + merchant feature lookups per request40-80
End-to-end decision latency budget (p99)120ms
Fraud label delay7-45 days

Question

How would you design this end-to-end system so it can make accurate real-time predictions at this scale while handling delayed labels, feature freshness, training-serving skew, and operational failures?