Thumbtack Interview Questions

Thumbtack Interview Questions

Practice 24 real Thumbtack interview questions for 2026. Thumbtack interview questions and interview preparation focused on real questions from actual interviews with detailed solutions. Coverage leans on Coding & Algorithms and System Design first, then Data Manipulation (SQL/Python), Analytics & Experimentation, Behavioral & Leadership, and Machine Learning. Expect interviews across Software Engineer and Data Scientist roles, with conversations that test coding fluency, analytical rigor, product sense, and clear stakeholder communication rather than purely theoretical exams. For Data Scientist applicants the loop frequently drills into practical themes you can rehearse: NLP preprocessing and n‑gram choices, designing cross‑validation and explaining bias–variance tradeoffs, choosing clustering versus regression and how KNN fits, and robust implementations (min/mean/median) and data‑structure tradeoffs (list vs dict). You’ll also get SQL and streaming questions—monthly new‑vs‑returning metrics, weekly 3‑week rolling sums, and parsing JSON/CSV at scale—plus A/B experimentation diagnostics, power analysis, rapid ad‑hoc analysis, and framing project tradeoffs for stakeholders. Prep by practicing clear walk‑throughs, concise SQL, and one‑page summaries that link metrics to business impact.

24 Questions 1 Company01.09.2026
Showing 20 results
Thumbtack logo
Thumbtack
Medium
Data Scientist

Implement TF–IDF with sparse matrices

Implement TF–IDF from Scratch (Python + NumPy/SciPy) You are given a list of documents (plain strings). Implement a TF–IDF vectorizer from scratch — n...

Coding & Algorithms
16
0
138 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Detail NLP preprocessing and n‑gram choices

Describe your text preprocessing pipeline given the source modality: typed text, scanned/handwritten OCR, or speech-to-text. Specify language handling...

Machine Learning
10
0
81 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Hard
Data Scientist Locked

Build a defensible ML pipeline end-to-end

This question evaluates a data scientist's competence in designing and defending an end-to-end production ML pipeline for mixed tabular data, assessin...

Machine Learning
5
0
66 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Hard
Data Scientist

Design and evaluate an A/B test for launch

A/B Test Design: New Matching Model for a Two‑Sided Marketplace Context You are testing a new matching/ranking model that determines which providers a...

Analytics & Experimentation
7
0
67 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Test regional response-rate differences rigorously

Goal Assess whether provider response rates differ by region after adjusting for job category mix and time. Data You have job-level observations with ...

Statistics & Math
11
0
84 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Hard
Data Scientist

Design streaming new-vs-returning monthly metrics

Streaming design: Monthly NEW vs RETURNING request shares (event-time, with late/out-of-order and duplicates) Context You receive a high-volume event ...

Coding & Algorithms
8
0
79 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Hard
Data Scientist

Lead XFN decision under tight timeline

Scenario: 72-Hour VP-Level Recommendation on Expanding a New Quoting Workflow You have 72 hours to deliver a VP-level deck recommending whether to exp...

Behavioral & Leadership
11
0
97 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Write complex joins and window functions

You are given a simplified Thumbtack-like marketplace schema in PostgreSQL. Assume UTC timestamps and weeks start on Monday. Treat "today" as 2025-09-...

Data Manipulation (SQL/Python)
0
0
10 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Compute weighted response rates by job category

You are given a CSV with one row per job posting and the following columns: job_id, job_category, invitations_sent (integer >= 0), provider_responses ...

Data Manipulation (SQL/Python)
0
0
9 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Estimate Two Conditional Win Probabilities by Simulation

Estimate Two Conditional Win Probabilities by Simulation You are given a game with three closed doors, labeled 0, 1, and 2. Behind exactly one door is...

Coding & Algorithms
1
0
8 people solved
Jan 9, 2026
Thumbtack logo
Thumbtack
Medium
Data Scientist

Choose clustering vs regression; explain KNN

When would you use clustering vs. regression on a business problem with partially labeled outcomes? Specify the decision criteria (label availability,...

Machine Learning
6
0
79 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Implement min, mean, median robustly

Implement three functions in Python without using numpy/pandas: (1) my_min(nums) returning the minimum in O(n) time and O(1) space; (2) my_mean(nums) ...

Coding & Algorithms
4
0
45 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist Locked

Optimize red-ball draw probability, prove optimality

This question evaluates probabilistic reasoning, optimization and mathematical proof skills by asking how to allocate red and blue balls across two bo...

Statistics & Math
7
0
71 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Demonstrate rapid analysis and stakeholder debrief

Rapid Analysis and Stakeholder Debrief Plan You have 1 hour to analyze a provided dataset (no pre-read) followed by a 45-minute debrief with a product...

Behavioral & Leadership
10
0
72 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Explain power drivers and resolve unexpected A/B results

A/B Testing: Power, Sample Size, Allocation, and Diagnostics You are analyzing a two-proportion (binary conversion) A/B test with independent users, n...

Analytics & Experimentation
3
0
74 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist Locked

Forecast response-rate trends with backtesting

This question evaluates proficiency in time-series forecasting and model validation, including feature engineering, model selection, rolling-origin ba...

Machine Learning
4
0
62 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Write monthly new-vs-returning requests SQL

Given the schema and sample data below, write a single PostgreSQL query (no dynamic SQL) that returns, for every calendar month present in requests, t...

Data Manipulation (SQL/Python)
1
0
11 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Hard
Data Scientist

Define success metrics for Instant Book

Instant Book: Metrics, Measurement, Rollout, and Risk Plan Context You are evaluating an "Instant Book" feature that allows customers to immediately b...

Analytics & Experimentation
3
0
60 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Compare list/dict; parse JSON/CSV at scale

Compare Python list and dict precisely: for append/insert/lookup/update/delete, state average and worst-case time complexity, memory implications, and...

Data Manipulation (SQL/Python)
6
0
99 people solved
Oct 13, 2025
Thumbtack logo
Thumbtack
Medium
Data Scientist

Explain a project and justify choices

Walk me through your most impactful project end-to-end: what problem and success metric did you define, what alternatives did you evaluate and reject,...

Behavioral & Leadership
3
0
45 people solved
Oct 13, 2025

Frequently Asked Questions

How difficult are Thumbtack Data Scientist interview questions?
Thumbtack data scientist interviews are typically moderate-to-challenging, with an emphasis on applied analytics and product thinking rather than pure theoretical machine learning. Expect 24 targeted questions that probe SQL and Python data manipulation, experiment design, pragmatic model choices, and business-facing communication. Coding algorithm problems are less common and usually straightforward; complexity comes from framing ambiguous product problems, justifying tradeoffs, and writing robust SQL for real metrics like monthly new-vs-returning or rolling sums. Senior roles see more design and statistical depth; entry roles focus on clean analysis and clear stakeholder recommendations.
What is Thumbtack's interview process for Data Scientist roles and where do specific topics appear?
The typical process starts with a recruiter screen, then a technical screen or take-home assignment, followed by a hiring manager discussion and a virtual onsite or panel that may include a senior leadership chat. SQL and Python data manipulation questions usually appear in the technical screen or take-home, while analytics, experimentation, and product-metric questions surface in the hiring manager and panel rounds. Behavioral and leadership evaluation is woven throughout. Machine learning and NLP topics are asked when the role explicitly requires modeling, often as part of a take-home or an in-depth technical interview.
How should I structure my interview preparation timeline for Thumbtack Data Scientist interviews?
Plan 4 to 6 weeks of focused preparation, or 2 to 3 weeks of intensive review if time is limited. Start by drilling SQL and data-manipulation problems and practice writing monthly and rolling-window queries until you can produce correct, efficient SQL under time pressure. Next, review experimentation, power and A/B root-cause analysis while practicing verbal explanations of unexpected results. Allocate time for model selection, cross-validation design, and simple NLP preprocessing cases. Close with timed mock interviews, take-home practice, and refining STAR stories so you can concisely justify choices and communicate tradeoffs to stakeholders.
What key technical subtopics should I prioritize for Thumbtack Data Scientist interviews?
Prioritize SQL window functions, aggregates, CTEs and robust handling of NULLs for tasks like monthly new-vs-returning metrics and three-week rolling sums. In Python, focus on list versus dict tradeoffs, efficient JSON/CSV parsing at scale, and robust implementations of min/mean/median. For modeling, be ready to explain cross-validation design, bias–variance tradeoffs, KNN intuition, and when to choose clustering versus regression. Expect NLP preprocessing and n‑gram reasoning on text tasks, plus quick probability or optimization proofs for small puzzles that test mathematical rigor and clear justification.
What standout tips and common pitfalls should I know for Thumbtack Data Scientist interviews?
Lead with product context: frame the metric or business question before diving into queries or models. In SQL, validate edge cases and explain performance implications; in take-homes include a short README and reproducible steps. When modeling, justify choices with data assumptions and cross-validation strategy, and avoid overfitting. For experiments, check power, instrumentation, and alternative explanations for unexpected results. Common pitfalls are solving technical subtasks without tying them to business impact, missing NULL or edge-case handling in queries, and failing to communicate tradeoffs clearly to non-technical stakeholders.

Explore more Thumbtack interview questions

Jump straight to Thumbtack questions for a specific role or category.

By role
In-depth guides
Across all companies

Real Thumbtack interview experiences

First-hand reports from Thumbtack candidates — the rounds, the questions they were asked, and how it went.

All 2 Thumbtack interview experiences