Data scientist interview questions for startups: 20 questions

data-scientist-interview-questions-for-startups:-20-questions

Updated October 2026

A startup data scientist interview works best in four stages: a problem-framing screen, a technical round on the flavour you are hiring for, a work sample presented to the founder, and a final round with references. They test, in turn, judgement, technical depth, communication of uncertainty and stage fit.

Choose the flavour first: analytics, decision science or machine learning. Then cut the questions that don't apply. A loop that tests all three rewards generalists who interview well and punishes the specialist you need.

The running example is an illustrative Seed-stage language-learning app hiring a decision science profile to work out why users drop off in week two and whether a new paywall helped.

The data scientist interview plan

StageWhoLengthWhat it tests
1. Framing screenFounder30 minFlavour, decisions their work changed
2. Flavour deep diveSenior data person or advisor60 minDepth in your chosen flavour
3. Work sample readoutFounder plus product lead60 minAnalysis and explaining uncertainty
4. Final and referencesFounder45 minStage fit; a reference from a stakeholder, not a peer
Which question groups to use
Group
Analytics
Decision science
ML
Framing
Use
Use
Use
Analytics
Use
Light
Skip
Experiments
Light
Use
Light
Production
Skip
Skip
Use
Uncertainty
Use
Use
Use

Framing the problem (every flavour)

  1. Tell me about work of yours that changed a decision. What was the decision, and who made it?
  2. A founder asks why revenue dipped last month. What do you ask before opening a notebook?
  3. When did you last tell a stakeholder the data could not answer their question?
  4. What's the simplest approach you'd try before reaching for a model?

Strong answer: Names a decision and a decision-maker, asks about timing, segments and tracking changes first, and starts with a rule or a chart.

Red flags: Lists models built with no decision attached, or never says no to a question.

Analytics

  1. Weekly active users rose but revenue fell. Give me three explanations and how you would check each.
  2. How would you build a retention cohort table from our raw events?
  3. Which metric in our pitch deck would you challenge, and why?

Strong answer: Several competing explanations, each with a quick check, and a willingness to question your own numbers.

Red flags: One explanation defended to the end, or cohorts built without checking how events are logged.

Experimentation

  1. Walk me through testing a new paywall, from hypothesis to readout.
  2. How do you fix the run time before the test starts?
  3. Halfway through, the product lead wants to stop because it looks like a win. What do you say?
  4. We lack the traffic for a clean A/B test. What would you do instead?
  5. A test shows no difference. What can and can't we conclude?

Strong answer: Sets a primary metric and sample size in advance, resists peeking, offers alternatives such as before-and-after with a holdout, and knows no difference is not proof of no effect.

Red flags: "Run it until it is significant", or no plan for low traffic.

"Did the new paywall work?" Two readouts from the illustrative language app
Scores 1 on communicating uncertainty

"Yes, it was statistically significant, so we should roll it out to everyone."

Scores 4

"Probably, for monthly plans. For annual plans the range still includes no effect. Roll out monthly now and run annual for two more weeks."

Getting models into production

  1. Tell me about a model that never reached production. Why not?
  2. What baseline would you compare a first recommendation model against?
  3. How do you know a live model has quietly got worse?
  4. Who owned deployment in your last team, and what did you hand over?
  5. When would you recommend not building a model at all?

Strong answer: An honest story of a shelved model and what they learned, a simple baseline such as most popular items, and monitoring of inputs as well as accuracy.

Red flags: Every model shipped, no baseline, or "engineering handled that" with no idea what happened next.

Communicating uncertainty

  1. Explain a confidence interval to a founder who must decide by Friday.
  2. Tell me about a recommendation of yours that turned out wrong. What did you say afterwards?
  3. The founder wants one number for the board. Your analysis gives a range. What goes on the slide?

Strong answer: Plain language, a recommendation despite the range, and ownership of past mistakes.

Red flags: Jargon, false precision, or refusing to recommend anything until the data is perfect.

Work sample: a paywall readout

Share synthetic results from a paywall test with a borderline overall result, one segment that moved and one tracking gap. Ask for a five-slide readout for the founder with a recommendation. Cap it at 3 hours. For the ML flavour, ask instead for a baseline and a monitoring plan, not a tuned model.

Look for three things: they question the data before the result, they state a range, and they still make a call. Present it in stage three to the founder, who is the real audience.

Scoring rubric

Outcome1234
Decisions changedNone namedVague influenceClear decisionSeveral, with decision-makers named
Experiment qualityPeeks until significantBasic designSized in advanceHandles low traffic well
Communicating uncertaintyHides or ignores rangeJargonClear rangeRange plus a call
Data foundationsTrusts all dataSpots issues if toldChecks tracking firstFound the planted gap
Models in production (ML only)Never shippedShipped with helpShipped and monitoredKnows when not to build
Suggested weighting for a decision science hire
Communicating uncertainty
30%
Experiment quality
30%
Decisions changed
25%
Data foundations
15%
Suggested weighting, not market data. For the ML flavour, give models in production a quarter of the weight.

Where Funded.club fits

We run data and engineering searches on a fixed fee agreed upfront, with one dedicated recruiter and first screened candidates within 7 days. See pricing.

Frequently asked questions

How do I interview a data scientist if I'm not technical?

Borrow an advisor for stage two and run stage three yourself. Explaining uncertainty to a non-specialist is the skill you are best placed to judge.

Should I hire for ML or for analytics?

Analytics, unless a model is part of your product. Our guide to hiring ML and AI engineers covers the ML route.

What should the job description say?

Name the flavour and the decisions the hire will change. Start from our data scientist job description template.

Worth a brief chat about your next hire? Book a free call.

Hiring after a funding round?

First candidates in 7 days. Fixed-fee, no commission.

See pricing Book a free call Try the Growth Planner