Quant interview preparation
Prop market making and quantitative research, weighted the way the interviews actually are: probability and expected value, statistics and machine learning, market making logic, programming and options. Every question is either traced to a named firm from a public candidate report, or tagged at desk level when we could not trace it, and every probability answer shows the reasoning path rather than just the number.
100 questions, mapped to the firms that asked them
- Questions
- 100
- Traced to a firm
- 53
- Firms
- 15
- Updated
- September 2026
053How would you cross-validate a model on time series data, and why is standard k-fold wrong?Quant researchQuant trading
Say this
Standard k-fold trains on data that comes after your test set, which leaks the future. You need a forward-walking scheme: train on a window, test on the next block, roll forward, and put a gap between train and test so overlapping labels do not bleed across the boundary.
Then walk it
- Two distinct leaks. First, random folds put future observations in the training set, so the model learns things it could not have known. Second, features and labels are usually built from overlapping windows, so even adjacent-in-time observations share information across a fold boundary.
- The fix for the first is walk-forward or expanding-window validation: fit on 1 to t, test on t plus 1 to t plus h, roll. Expanding window mimics how you would actually retrain in production. A fixed rolling window is better if the process is non-stationary.
- The fix for the second is purging and embargoing, from Lopez de Prado. Remove training observations whose label window overlaps the test period, and embargo a short period immediately after the test block. On a 20-day forward return label you need at least a 20-day purge.
- Also beware the hidden leaks that sit outside the folds entirely: fitting a scaler, doing feature selection, or choosing hyperparameters on the full dataset before splitting. Every preprocessing step has to sit inside the fold.
- What I would actually report, and this is the part that matters: one final untouched hold-out period tested once, plus how many configurations I tried before I got there. Walk-forward validation run a hundred times is itself an overfitting device, and the number of trials is the honest measure of how much to discount the result.
Where candidates lose it
Saying you would use k-fold with shuffle turned off and stopping there. That fixes the ordering but not the overlapping-label leak, and interviewers at systematic shops probe exactly that. Mention purging and embargo, and mention that scalers and feature selection must live inside the fold.
Expect next
- How long should the embargo be?
- Expanding window or fixed rolling window, and why?
- How do you account for the number of configurations you tried?
Firm tags come from public, anonymous candidate reports on Wall Street Oasis: strong signal, not sworn testimony. Firms are named as the places a question was reported, not as partners of Fin Maverick. Answers are written for this page to show how to think out loud; they are not scripts to recite.

