Stabilize LLM inference and estimate needed repeats
Company: Citadel
Role: Data Scientist
Category: ML System Design
Difficulty: medium
Interview Round: Technical Screen
Overview: This question evaluates skills in designing reliable LLM inference pipelines and in statistical modeling of stochastic outputs, including reproducibility engineering, uncertainty quantification, and the use of correlation metrics (e.g., Pearson) to measure stability.
Read the full Citadel Data Scientist interview experience this question came from