Fin Maverick
Foundations VocabularyAccounting & ReportingEconomics & MacroQuant Methods & ProgrammingBusiness & Company AnalysisCorporate Finance & ValuationBehavioural Finance
Banking & Market InfrastructureFixed Income & RatesDerivatives & Structured ProductsPublic EquitiesTransactions & DealsPortfolio ConstructionFunds & AMCs
Private Markets & AlternativesRisk, Treasury & ControlAI & Digital FinanceStochastic Calculus & PricingWealth & Personal FinanceIndian Markets & RegulationProfessional Practice
CalculatorComparison
Frameworks
Explore Bootcamps
Equity ResearchPortfolio ManagementMutual Fund MasteryFinancial LiteracyInvestment Banking Analyst
Private Equity AnalystHedge Funds AnalystBreaking Into VCBreaking Into QuantsAI For Finance
Financial Analyst ProgramRisk Management ProgramPrivate Wealth ManagementDebt Capital MarketsDerivatives Foundation
Explore Internships
Equity Research InternMutual Fund Intern
Portfolio Management InternFinancial Literacy Intern
Explore Micro Courses

Equity Research6

Writing an Investment ThesisBuilding a Discounted Cash FlowReading an Annual Report FastReading a Sector Before a CompanySpotting Quality of Earnings Red FlagsBuilding a Revenue Forecast From Drivers

Portfolio Management3

Rebalancing: When, Why and What It CostsStrategic and Tactical Asset AllocationMeasuring Risk in a Portfolio

Mutual Fund Mastery3

Comparing Funds Without Being FooledHow a NAV Is Struck and Which Day You GetReading a Fund Factsheet Properly

Derivatives Unlocked4

Hedging a Real ExposureThe Greeks, PracticallyFutures, the Basis and What Moves ItReading an Option Payoff

AI For Finance2

Retrieval and Grounding for FinanceDocument Extraction in Finance

Breaking Into Quants4

Backtesting a StrategyHypothesis TestingCleaning Financial DataRegression for Finance

Breaking Into VC3

Sizing a MarketReading a Term Sheet as a FounderHow a Venture Round Actually Works

Financial Analyst Program4

Common Size and Trend AnalysisReading a Cash Flow StatementRatio Analysis That Says SomethingBuilding a Working Capital Schedule

Risk Management Program2

Credit Exposure and How It Is ReducedValue at Risk and What It Hides

Investment Banking Analyst3

Precedent Transactions and Why They DifferReading a Term Sheet StructurallyBuilding a Comparable Companies Table

Private Wealth Management3

Tax Aware Portfolio DecisionsBuilding a Client Risk ProfileGoal Based Planning Arithmetic

Debt Capital Markets3

Analysing an Issuer's CreditDuration and What It Does Not Tell YouBond Pricing and Yield Mechanics

Private Equity Analyst2

Fund Waterfalls and CarryThe LBO in Structure

Hedge Funds Analyst2

Short Selling MechanicsLong Short Mechanics
Courses
Explore Career Roadmaps
Investment Banking AnalystEquity Research AnalystVC AnalystPrivate Equity AnalystHedge Funds Analyst
Quant AnalystAI For FinanceFinancial Analyst ProgramPrivate Wealth ManagementDebt Capital Markets
Risk Management ProgramDerivatives FoundationPortfolio ManagementMutual Fund Mastery
PartnershipsShowdown
Log inSign up
Quant Analyst · CoreTrack
1Quantitative Methods, Financial Data & Programming
iProbability
Probability in FinanceRandom VariableProbability DistributionsThe Normal DistributionNormal Distribution ProbabilityThe Lognormal DistributionRandomness vs Uncertainty
iiStatistics and Inference
Population and SampleMean, Median and ModePrecision and AccuracyVariable TypesVariance, Standard Deviation and…Dispersion MeasuresStatistical BiasEffect SizeHypothesis TestingThe Sampling DistributionSkewnessKurtosisCovarianceConfidence IntervalArithmetic Mean vs Geometric MeanStatistical Significance vs Economic…Confidence Interval vs Prediction IntervalHow to Summarise a…
iiiCorrelation and Regression
RegressionCorrelation and CausationOrdinary Least SquaresInteraction TermsRegression CoefficientsRegression vs ClassificationHow to Build a…Spurious CorrelationRegression, Correlation and FitResidualsMulticollinearityAutocorrelation and Partial Autocorrelation
ivTime Series
Time Series in FinanceSimple, Weighted and Exponential…Moving Average CalculatorPrice, Return and Level SeriesHow to Prepare Time-Series…LagFrequencySeasonalityTimestampsTrendStationarity and the Unit RootHeteroskedasticityLeadRolling WindowsDifferencing
vSimulation and Numerical Methods
SimulationMonte Carlo SimulationHow to Run a…Numerical MethodsIterationResampling and the BootstrapPseudorandom Numbers and the SeedConvergence and ToleranceNumerical Stability
viOptimisation
OptimisationLocal and Global OptimaConstraintsConvex OptimisationThe SolverLinear ProgrammingThe Objective FunctionConstraint ViolationThe Feasible SetLagrange MultipliersQuadratic Programming
viiModelling Practice
Linear, Logistic, Ridge and…Training, Validation and Test…The ModelModel ErrorDependent and Independent VariablesThe ROC Curve and AUCWhat a Model HoldsMSE, RMSE, MAE and MAPEPrecision and RecallCross Validation and RegularisationOverfitting and UnderfittingReturn Series MeasuresSimple, Compound and Log Return
viiiBacktesting and Research Integrity
BacktestingBacktest vs Live PerformanceHow to Document a…How to Prevent Backtest…Out-of-Sample TestingWalk-Forward AnalysisMultiple TestingP-HackingData Snooping
ixData Quality and Structure
Data QualityThe DatasetSelection and Survivorship BiasVersioned DatasetsData Structures in FinanceData CleaningMissing Data and Null ValuesStructured Data vs Unstructured DataMissing Data vs ZeroData Validation vs Data CleaningOutliersDuplicate Records
xProgramming for Finance
Data PipelinesAPIs for Financial DataAPI vs CSV FileDatabases in FinancePython for FinanceJoinsSQL for FinanceThe Analysis Workflow
xiQuantitative Research
Research DesignThe Data Generating ProcessReproducibilityPeer Review in Analytical WorkThe Research HypothesisRobustness and Sensitivity
2Stochastic Calculus & Derivative Pricing Theory
iProbability Foundations
The Probability SpaceRandom VectorsSigma-AlgebraExpectationSample Space and EventsDensity and Distribution FunctionsRisk-Neutral ProbabilityState Price Density vs…
iiStochastic Processes and Jumps
Properties of a Stochastic ProcessMartingaleBrownian Motion and Its PropertiesBrownian Motion vs Geometric…Stopping TimeThe Markov PropertyState VariablesTransition ProbabilityQuadratic VariationQuadratic Variation vs Ordinary…Submartingale and SupermartingaleMartingale RepresentationMarkov Process vs MartingaleOptional StoppingFiltrationJump ProcessesThe Poisson ProcessLevy ProcessesJump Diffusion
iiiIto Calculus
The Ito IntegralThe Ito Integral vs the Riemann IntegralInfinitesimals in Stochastic CalculusQuadratic CovariationIto's LemmaHow to Apply Ito's…The Infinitesimal GeneratorIto Calculus vs Ordinary Calculus
ivStochastic Differential Equations
Stochastic Differential EquationsStochastic Differential Equation vs…Drift and DiffusionStrong and Weak Solutions ComparedDiscretisationGeometric Brownian Motion
vPricing Theory and No-Arbitrage
No-ArbitrageGirsanov, Radon-Nikodym and Change…Physical and Risk-Neutral Measures…The Fundamental Theorems of…The Law of One PriceThe Pricing KernelDiscount Factors and Zero-Coupon PricesReplication vs HedgingComplete Market vs Incomplete MarketClearing Margin Architecture
viOption Pricing Theory
European and American OptionsMonte Carlo European OptionThe Black-Scholes PDEBlack Scholes and the GreeksThe Payoff FunctionThe Binomial ModelBinomial Option PricingDelta Hedging in TheoryBoundary, Initial and Terminal ConditionsThe Exercise BoundaryHow to Check Put-Call…
viiVolatility Models
Constant, Local and Stochastic…Vasicek Model vs CIR ModelThe Heston ModelThe SABR ModelThe Volatility ProcessImplied VolatilityVolatility Smile vs Skew vs Surface
viiiInterest Rate Models
Interest-Rate DerivativesMean ReversionThe Zero-Coupon BondThe Ornstein-Uhlenbeck ProcessThe Discount CurveZero RatesShort-Rate Model vs Market Model
ixNumerical Pricing
Closed Form and Numerical…Monte Carlo PricingEuler and Milstein Schemes ComparedTree MethodsFinite Difference MethodsNumerical Error and StabilityVariance Reduction
xCalibration and Model Risk
Model OverrideMarket Price and Model PriceCalibrationHow to Document a Pricing ModelThe Educational Illustration LabelMarket ConventionsModel Uncertainty and LimitationsBacktesting a Pricing ModelIdentifiabilityCalibrated ParametersThe Calibration Loss Function

How to Document a Backtest So Somebody Can Rerun It

Three things are settled separately and are not rebuilt here: what it means to test a rule against a stretch of history, why a figure arrived at after choosing among options is a different kind of number from one arrived at without choosing, and what fitting on one part of a record and reading on another is for. Every count printed here comes from rerunning one calculation on one invented record of 72 monthly readings, and any word borrowed from finance is explained the first time it turns up.

The argument below rests on three things that can be checked rather than taken on trust. The first is the Ashwin rule, an invented rule settled earlier: it looks at how the Nakshatra unit moved last month, and when that movement clears a set thresholdA cut off. A reading either clears it or it does not, and which side of the cut off it lands on is what makes the rule say up rather than down. the rule calls the coming month up, and otherwise down. The second is the six year record itself, 72 monthly readings of the Nakshatra unit running from January 2019 to December 2024, every one of them made up so that a testing method could be watched from the outside.

The third is counting, and only counting. Nowhere below is a missing field called serious, or careless, or a matter of professional standards. Each one is measured instead, by counting how many different readings a write up lacking that field is still consistent with. Counting is the only currency in use, and counting is why the argument can be checked by anybody with the same 72 readings and an afternoon.

What is a backtest write up actually for?

The purpose of the document decides everything else. A backtest write up is adequate when a second person, holding only that document and the record, sits down and arrives at the same number. The standard stops there. Adequacy is not a matter of effort, or of how carefully the author thought, or of how long they have been doing this. Adequacy is a property of the document, and a document can be tested by handing it to somebody.

The shape is much easier to see away from finance. A recipe card that says bake until done can be followed by the person who wrote it and by nobody else. The person who wrote it knows that until done means the top has stopped wobbling, that this particular oven runs hot, and that the tin goes on the middle shelf rather than the top one. None of that is on the card. Handed to a neighbour, the card produces a different cake. Nobody was careless and nobody lied. The card was never a set of instructions. The card was a reminder, written by somebody who did not need reminding of the parts they left out.

A backtest write up fails in exactly the same way and for exactly the same reason. The author knows which months were counted, having been sitting there while counting them. The author knows that the months in which the Nakshatra unit finished where it started were dropped. At the time that seemed too obvious to write down: there is nothing to be right or wrong about in a month that stood still. The author knows that thirteen settings were tried before one was kept. None of that reaches the document unless somebody puts it there, and every one of those choices moves the number.

Notice what this standard refuses to do. The standard does not grade a write up out of ten, and it does not let a thorough document that is missing one field score nine. The second person either lands on the author's number or does not, and there is no partial credit in that test. The harshness is the useful part. Diligence is a question nobody outside somebody's own head can settle. The standard converts it into a question about a document, and anybody at all can settle that in an afternoon.

THE ONLY TEST A WRITE UP HAS TO PASS, AND IT HAS EXACTLY TWO OUTCOMES an invented run on an invented record. Nothing here is measured and nothing is recommended. A second person holds the document and the record, and nothing else. Do they arrive at the same number? YES NO ADEQUATE The document is finished, and nothing about its author entered the test. NOT ADEQUATE A field is missing, and which one it is is the only question left to ask. There is no third box. A write up is not partly reproducible, and care by the writer is not a substitute for a field.
A write up is adequate when a second person holding only the document and the record arrives at the same number, and that test has two outcomes with no partial credit between them.
Try it out

When is a backtest write up adequate?

Breaking Into Quants Bootcamp — Fin Maverick

Which six fields does a write up have to carry?

Six fields, and five of them are one line each. Two of the six describe the search rather than the run itself: the gridA written list of the settings a search will work through, decided up front, so that the list is itself a document rather than a recollection. of settings, and the count of attempts. One of them, the counting conventionA rule for counting, settled and put in writing beforehand, so that two people handed the same rows end up with the same total., describes a decision so small that people leave it out for being obvious. And the first of them, the record and its spanHow far a set of readings reaches: the date it opens on and the date it closes on., is the field most likely to be right. The record itself is usually sitting in the same folder. Here are all six, filled in for one honest run.

The six field write up, completed for one run of the Ashwin rule.
FieldWhat it has to sayThis run
1. The record and its spanwhich readings, and between which dates72 monthly readings of the Nakshatra unit, January 2019 to December 2024
2. The rule, in one sentencewhat it reads, and what it then callsthe Ashwin rule reads last month's reading and calls the coming month up when that reading clears the threshold
3. The grid searchedevery setting on the list, written before the search startsthirteen whole number settings, from minus 6.00 per cent up to 6.00 per cent
4. The split datewhere the record was cut in twothe end of December 2021
5. The counting conventionhow an awkward month is treateda month in which the unit did not move is left out
6. The count of attemptshow many settings were actually runthirteen

Field five is the one people are most surprised to find on a list of six, so look at what it does to the arithmetic rather than at how small it sounds. The record holds 72 months. January 2019 has nothing in front of it, so the rule has nothing to read and that month is never called, leaving 71. Six of those 71 are months in which the Nakshatra unit finished exactly where it started, and a call of up or down cannot be marked right or wrong against a month that did not move, so those six come out as well. The denominatorThe bottom half of a fraction. Twenty six right out of twenty nine and twenty six out of forty five are the same twenty six, and the bottom half is the whole reason they read so differently. is 65, and the only reason a second person knows it is 65 rather than 71 is that field five is in the document.

Field three and field six look like the same thing written twice, and they are not. The grid is the list of settings intended for trial, written down before any of them was tried. The count of attempts is how many were actually run. When a search goes as planned the two match, and on this run they both read thirteen. When a search does not go as planned they part company, and the gap between them is then the most informative line in the whole document. The gap records the settings that were added after somebody had already seen how the first ones came out.

Five of the six fields are one line each, and the sixth is a single word. The completed write up fits on a filing card with room to spare. The reason these fields go missing is never that they were hard to write.

THE SAME READING AT THE BOTTOM OF TWO DOCUMENTS, AND ONLY ONE CAN BE REBUILT the six field write up on the left, the same honest run written up loosely on the right. All entries invented. WRITE UP ONE, SIX FIELDS FILLED WRITE UP TWO, THREE FIELDS LEFT OUT 1 THE RECORD AND ITS SPAN 72 monthly readings of the Nakshatra unit, January 2019 to December 2024 2 THE RULE, IN ONE SENTENCE reads last month, calls the coming month up when last month clears the threshold 3 THE GRID SEARCHED thirteen whole number settings, from minus 6.00 per cent up to 6.00 per cent 4 THE SPLIT DATE the end of December 2021 5 THE COUNTING CONVENTION a month in which the unit did not move is left out, so the count sits on 29 6 THE COUNT OF ATTEMPTS thirteen 1 THE RECORD AND ITS SPAN 72 monthly readings of the Nakshatra unit, January 2019 to December 2024 2 THE RULE, IN ONE SENTENCE reads last month, calls the coming month up when last month clears the threshold 3 THE GRID SEARCHED left out of the document 4 THE SPLIT DATE left out of the document 5 THE COUNTING CONVENTION a month in which the unit did not move is left out, so the count sits on 29 6 THE COUNT OF ATTEMPTS left out of the document READING: 89.6552 per cent READING: 89.6552 per cent Two documents describing one run. The same figure sits at the bottom of both, printed in the same ink. The left one can be rebuilt by anybody. The right one cannot be rebuilt by the person who wrote it.
The same reading of 89.6552 per cent sits at the bottom of both forms, and only the one carrying all six fields can be reproduced by anybody else.
Try it out

There is time to write only one of the six fields before the analyst is pulled away. Which one, and what does it protect?

What does each missing field cost, measured in readings?

Now the measurement, the part worth the most. Take one honest run. The Ashwin rule at a threshold of minus 1.00 per cent, read over the first three years of the record with the unchanged months left out, calls 26 of 29 months correctly, a reading of 89.6552 per cent. Nothing about that run is wrong. Nothing about it is hidden. The figure is exactly what the arithmetic gives, and it would give the same figure to anybody who ran it that way.

So the question is not whether the run is honest. The question is how much of it a reader can pin down from the document. Take the same run and ask a different thing: given only what the write up actually says, how many other analyses of the same record would also have been fair to run, and what readings do those give?

The menu is short and every item on it is defensible. Five stretches of the record a person might reasonably read the rule over. Two ways of handling a month in which the unit did not move, either dropping it or counting it as a rise. And two signalsWhatever the rule looks at immediately before it decides. Here that is a single number, assembled only from months the record has already finished with. the rule might read, either last month on its own or the average of the last three. Five times two times two is twenty. Nobody has to bend anything to arrive at any one of the twenty; each is a choice somebody would make without blinking, and each is a choice this write up either records or does not.

TWENTY DEFENSIBLE ANALYSES OF ONE RUN, AND A FULL WRITE UP PICKS EXACTLY ONE five stretches by two counting conventions by two signals. Every reading below is invented. unchanged months left out unchanged counted as a rise last month average of the last three last month average of the last three all seventy two months 44 of 6567.69 per cent 39 of 6361.90 per cent 47 of 7166.20 per cent 42 of 6960.87 per cent the first thirty six 26 of 2989.66 per cent 19 of 2770.37 per cent 29 of 3582.86 per cent 22 of 3366.67 per cent the first forty eight 32 of 4178.05 per cent 26 of 3966.67 per cent 35 of 4774.47 per cent 29 of 4564.44 per cent the last forty eight 26 of 4557.78 per cent 26 of 4557.78 per cent 27 of 4856.25 per cent 27 of 4856.25 per cent the last thirty six 18 of 3650.00 per cent 20 of 3655.56 per cent 18 of 3650.00 per cent 20 of 3655.56 per cent The one lime cell is what a six field write up pins down. Without three of those fields, the document is consistent with all twenty.
Five stretches by two counting conventions by two signals gives twenty defensible analyses of one run, and the full write up picks the single cell reading 26 of 29.
Try it out

The three recoverable fields are about to be switched on one at a time. Before the control moves: does the band of readings shrink by roughly a third at each step?

Play with it

Watch the band close as each field goes in

One run, one record, one threshold, held completely still. The only thing that moves is how much of the run the document actually says. Every tick on the scale is one of the twenty analyses; the pine ticks are the ones this document still allows and the pale ones have been ruled out.

all three recorded
Analyses still allowed
1
Lowest reading
89.6552 per cent
Highest reading
89.6552 per cent
Band width
0.0000 points

Held still throughout: the record, the rule, the threshold of minus 1.00 per cent, and the run itself. Only the document changes.

Educational illustration. The menu of twenty analyses is the one set out above and no larger, and the threshold is fixed at minus 1.00 per cent throughout so that the band measures documentation and nothing else.

Recording the fields one at a time gives the ladder below, and those four rows carry the whole argument.

The band of readings one honest run is still consistent with, as each field goes into the document.
Fields recordedAnalyses still allowedBand of readingsWidth
nothing recorded2050.0000 to 89.6552 per cent39.6552 points
the stretch466.6667 to 89.6552 per cent22.9885 points
and the counting convention270.3704 to 89.6552 per cent19.2848 points
and the signal189.6552 per cent only0.0000 points

The band narrows at every step and closes completely at the last, and that is what turns the six fields from an opinion into a list. The top row is the least believable line above, and it deserves a second reading. A write up that names the record, states the rule and prints 89.6552 per cent, and says nothing about the stretch, the convention or the signal, is exactly as consistent with a run that reported 50.0000 per cent. Both documents are true. Both describe real arithmetic on the same 72 readings. Neither reader can tell which one they are holding.

The three widths are not evenly spaced at all, so do not read them as though each field were worth about the same. Naming the stretch closes 16.6667 points of the band. Adding the counting convention closes another 3.7037. Adding the signal closes the remaining 19.2848. The largest saving and the smallest sit on either side of the middle one, in that order, and there is no pattern in it to learn. A field is worth whatever the analyses it rules out happen to disagree about, and that has nothing at all to do with how important the field sounds when it is said out loud.

There is one limit worth naming before the panel above gets over-read. Only three of the six fields move the band, and those three are the ones that change which months got counted and how. The other two, the grid and the count of attempts, do not move the reading by a single month: run the same setting after trying twelve others first and it still calls 26 of 29. Neither the grid nor the count of attempts can show up in the band at all. Neither changes the reading itself; both change how much that reading is worth. Being invisible in the band is precisely why those two go missing.

WHAT THE DOCUMENT STILL ALLOWS, AS EACH FIELD GOES IN the band of readings a write up of one honest run is still consistent with. All figures invented. nothing recorded: 20 analyses survive, band 39.6552 points 50.0000 the stretch recorded: 4 analyses survive, band 22.9885 points 66.6667 and the counting convention: 2 analyses survive, band 19.2848 points 70.3704 and the signal: 1 analysis survives, band 0.0000 points 89.6552 50 60 70 80 90 per cent of months called correctly the reading the write up reports, 89.6552 per cent, sits at the dashed line in all four rows
Recording the stretch, then the counting convention, then the signal narrows the band of readings the document allows from 39.6552 points to 22.9885, then 19.2848, then nothing.
Try it out

A write up records the stretch and nothing else. How many analyses does it still allow, and how wide is the band of readings consistent with it?

AI For Finance Bootcamp — Fin Maverick

Which field goes missing most often, and why does nobody notice?

The field is the count of attempts, and the reason it goes missing has nothing to do with dishonesty. The count goes missing because trying thirteen thresholds does not feel like thirteen tests. It feels like one afternoon.

Picture the afternoon. The analyst sits down and tries minus 6.00 per cent; the reading is dull. Then minus 5.00 per cent; still dull. Working up the grid, the reading climbs, it peaks at minus 1.00 per cent, and that is the one written down. From the inside this is a single continuous act with a single result, in the way that looking for a set of keys in six rooms is one search and not six. Nothing in the experience marks a boundary between one attempt and the next, so nothing in the write up marks one either.

Now the part that makes this field different from the other five: the document cannot recover it from anything else in the document. The record is in the document, so a reader can count the months. The rule is in the document, so a reader can reread it. The reading is in the document, and a reader can recompute it from the first two. The number thirteen is in the document only if somebody typed it there. The number leaves no trace in the record, no trace in the rule and no trace in the reading, and once the afternoon is over it exists in one person's memory and nowhere else at all.

The cost falls due later, and it falls on the figure rather than on the person. A reading of 89.6552 per cent from a rule somebody proposed and then tested once is a different claim from the same 89.6552 per cent picked as the best of thirteen. Same number, same record, same arithmetic, different claim. When the write up cannot say which of the two it is, a careful reader has to assume the second. The second is the weaker reading, and it is the one the document does not rule out. A missing attempt count does not make the figure wrong; it leaves the figure unable to defend itself.

THE FIELD THAT GOES MISSING, AND THE ONE PLACE IT STILL EXISTS an invented write up card. No real record, result or organisation is described. BACKTEST RECORD CARD 1 RECORD 72 months, Jan 2019 to Dec 2024 2 RULE Ashwin rule, reads last month 3 GRID minus 6.00 to 6.00 per cent 4 SPLIT the end of December 2021 5 CONVENTION unchanged months left out 6 ATTEMPTS how many were tried? five lines typed, one line blank WHAT THAT BLANK COSTS The record is on the card. The rule is on the card. The reading is on the card. The number thirteen is in nobody's document at all. It is in one person's memory, and a memory is not a field. So the reading has to be read as the best of an unknown number of tries. Five of the six fields are one line each. The sixth is one word, and it is the one that goes missing.
The count of attempts cannot be recovered from anything else in the document, because the number exists nowhere except in the memory of whoever ran the search.
Try it out

Which field goes missing most often, and why does nobody notice it going?

How is it recorded when each number could first have been known?

There is a seventh thing on a good write up, and it is not a field of its own. The seventh thing is a date written beside every number the rule reads. Every number has a date on which it could first have been known, and that date is not always the date the number is filed under. Recording it costs a column and it is the difference between a document a reader can trust and a document a reader has to take on faith.

The earlier reading on this same record already worked the case through, so it does not need rebuilding here. Take a three month average of the price of the Nakshatra unit. Computed honestly, from this month and the two before it, that average sits Rs 4.9061/- away from the price it is meant to describe, averaged over the 69 months where it can be worked out at all. Now compute the same three month average as a centred averageAn average that borrows from both sides of the month it is written against, the ones before it and the ones after. Such an average fits the past neatly and belongs to no single date inside it., using the month before, the month itself and the month after. The centred average now sits Rs 2.5431/- away, an improvement of 48.1650 per cent.

Nothing about the second average is a better idea, and the whole of that improvement came from one month of information that had not happened yet. Where the number stays in a column, this is a data error: a column that describes history beautifully and cannot honestly be dated to any month in it. Inside a backtest the same thing is a backtest error, and the difference is not a quibble. The number did not stay in the column. The number went into a rule, and it came out as a reading somebody might quote.

Here is the backtest version of exactly that mistake, and it is deliberately unglamorous. Leave the Ashwin rule alone. Leave the record alone. Leave all thirteen settings alone. Swap only the signal: instead of reading last month, let the rule read the average of this month and the next. A two month average is perfectly ordinary arithmetic, and this one is dated one month too early. Scored on the same 65 months, the best reading goes from an honest 67.6923 per cent to a leaked 84.6154 per cent, a lift of 16.9231 points.

The leaked ladder reads higher at all thirteen settings and lower at none, and the lift runs from a single month at the lowest setting to thirteen months at a threshold of 1.00 per cent. The shape of the lift is what makes this failure so hard to catch by eye. The leak does not produce one absurd figure sitting next to twelve sensible ones. Anybody would query a figure like that. The leak lifts the whole shape, keeps it looking exactly like a ladder should look, and moves the best setting from minus 1.00 per cent to 1.00 per cent, so even the conclusion about which setting looked best is a leaked conclusion. 84.6154 per cent is the most flattering number in the whole comparison and the least real, so the word leaked stays attached to it every time it is printed.

AN HONEST WINDOW ENDS WHERE IT IS DATED. A LEAKED ONE REACHES ONE CELL PAST. the same mistake at two window widths, on invented months of an invented record. the price average the earlier reading already worked through, three cells wide honest: three months up to and including September May Jun Jul Aug Sep Oct Nov Dec leaked: the same three months centred on September, which needs October the solid pine bar in both panels marks the end of September, the date both signals are stamped with the two month reading average swapped into the rule here honest: two months up to and including September May Jun Jul Aug Sep Oct Nov Dec leaked: September and October, stamped September In both panels the honest window stops at the pine bar and the leaked one steps straight over it. The earlier reading calls this a data error. Inside a backtest it is a backtest error, because the number went into a rule and came back out as a reading.
An honest signal reads months that have finished and a leaked one reaches one month to the right, and that single cell is the only difference between the two columns.
THE LEAKED LADDER SITS ABOVE THE HONEST ONE AT EVERY SINGLE SETTING same rule, same record, same 65 scoreable months. Only the date on the signal moved. All figures invented. 45 55 65 75 85 per cent of the 65 months called correctly leaked best 84.6154 honest best 67.6923 the leaked ladder the honest ladder minus 6 minus 3 0 3 6 and the lift, in months, setting by setting 1 2 4 8 9 9 10 13 10 6 6 6 5 thirteen settings, from a threshold of minus 6.00 per cent on the left to 6.00 per cent on the right. No bar reaches zero.
The leaked ladder reads higher at all thirteen settings and lower at none, and the leaked best of 84.6154 per cent stands 16.9231 points clear of the honest best of 67.6923 per cent.
Try it out

Swapping in a signal that reaches one month forward lifted the best reading by 16.9231 points. Does that make the rule better, and where did the lift come from?

Try it out

There is a check that catches a leaked column in about five seconds and needs no arithmetic at all. What is it?

Reading an Option Payoff — free micro-course from Fin Maverick

What does a column with an empty last row indicate?

Scroll to the bottom of the column. Scrolling is the whole check, and it costs nothing.

The honest signal on this record is the month's own reading, and the record has one of those for every month. The honest column therefore has 72 filled rows and stops when the record stops. The leaked signal is the average of the month and the one after it, so it can be worked out for 71 months and not for the seventy second. December 2024 needs a January 2025, and the record has not got one. A column whose last row cannot be filled was built from the future, and finding out takes as long as scrolling to the bottom of it.

The check asks a question rather than settling one, and that is worth stating exactly. A column can be honest and still short at the bottom for a dull reason, such as somebody trimming a row while tidying. A signal that reaches two months forward leaves two rows empty rather than one, so the shape of the gap shows how far the reach went. And a column that is short at the top rather than the bottom is usually the ordinary consequence of an honest backward average. A backward average needs earlier months the record does not have. The empty last row gives a specific question, asked of a specific column, at no cost at all. In a folder of forty columns it is the cheapest check of the day.

THE CHECK THAT TAKES FIVE SECONDS AND NEEDS NO ARITHMETIC the last six rows of two signal columns on the same invented record. MONTH HONEST SIGNAL, LAST MONTH LEAKED SIGNAL, THIS AND NEXT Jul 2024minus 13.00 per centminus 10.50 per cent Aug 2024minus 8.00 per centminus 3.00 per cent Sep 20242.00 per centminus 0.50 per cent Oct 2024minus 3.00 per cent0.50 per cent Nov 20244.00 per cent9.00 per cent Dec 202414.00 per cent cannot be filled The signal stamped December 2024 needs a January 2025, and the record has not got one. A column whose last row cannot be filled was built from the future. Scrolling to the bottom of a column is the whole test. Nothing has to be recomputed and nothing has to be refitted.
The honest signal fills all 72 rows and the leaked one fills 71, because the signal stamped December 2024 needs a January 2025 the record has not got.

What this looks like when it goes wrong

An analyst writes the run up carefully. The record is described, the rule is stated in a sentence, the split date is given, the reading of 89.6552 per cent is printed, and the whole write up reads as the work of somebody being thorough. The work was thorough. Two fields are not on it: the grid, and the count of attempts. The grid and the count of attempts were left off because they felt like working notes rather than results. Feeling that about them is completely ordinary, and it is the reason this failure is common rather than rare.

Two years later a colleague picks the document up. The question they need answered is simple. Was one threshold proposed and then tested, or were thirteen tried until one landed? The document cannot say. Neither can the analyst. The number thirteen was never written down anywhere, and two years is longer than that kind of memory lasts. Rereading the document more carefully will not help, and no amount of care on the reader's side can put back a field the writer never wrote. The figure now has to be read as the best of an unknown number of attempts. Reading it that way is the weakest reading available, and the only route back to a stronger one is to rerun the entire search from the start.

The fix costs one line and it has to happen at the right moment: the grid goes down before the search rather than after it. A grid written in advance is a list of thirteen numbers on a card. The same grid written from memory afterwards is a recollection, and it will quietly leave out the two settings that were tried first, disliked, and moved on from. Writing it first is what converts a memory into a document, and there is no later moment at which that conversion can still be made.

The empty last row gives the leak away. See what the document carries.

How does anybody actually use a write up they have been handed?

A write up arrives from somebody else. The write up will probably not be rerun today and may never be rerun at all, so the practical question is what to read first and what to conclude from whatever is missing.

The six fields are read before the result. A lender looking at a credit team's screening rule, an analyst inheriting a spreadsheet from somebody who has left, a committee member reading a document prepared last year: all three are in the same position. All three are being asked to attach weight to a number they did not produce. The fields tell them how much weight the number can carry. A document with all six is a number they can use the way the author used it. A document missing the stretch is a number that could have been nineteen other numbers. A document missing the attempt count is a number that has to be read as the best of a search whose size nobody knows.

The household version has the same shape, and it makes the question easy to ask out loud. Somebody reports that a particular shop is cheaper. Cheaper on what, over what period, and was it the first shop checked or the eleventh? Nobody is lying in any version of that conversation, and the answer still changes completely depending on which of the three is known.

One sentence is worth carrying away: a second person, holding only the document and the record, arrives at the same number. Six fields, a date beside every number the rule reads, and the file that regenerates the figure. With those handed over, the write up is finished. Without them, what has been written is a reminder, and a reminder only works for the person who did not need it.

Try it out

An inherited write up carries the record, the rule and the reading, and nothing else. What has to happen before the figure is quoted?

What is left for other reading?

Two things are assumed rather than repeated: what a backtest is at all, and why a reading taken off history is not the same animal as one taken later while the rule is running. Covered separately: how to hold a search down before it starts, how to hold a stretch of the record back, and how to roll a test forward through time, each a separate question from what a document has to carry. Running many tests against one record needs arithmetic of its own and is covered separately too.

Nothing above says that any rule is worth testing, and nothing above says the Ashwin rule is worth anything at all. The Ashwin rule appears only as a thing being documented, and every figure attached to it is a hit rate or a count of months rather than anything a person could receive. Buying anything, selling anything, what a trade costs, and whether any rule deserves running at all: each is covered separately.

What was opened to produce the figures above?

Nothing was opened. Every count above came out of rerunning one calculation on 72 invented monthly readings, so the rows below record the working file that produced the numbers rather than an institution that published a table.

SourceDocumentSite
No outside source was usedA checking file kept in the same folder as these notes, rebuilding all 72 readings and every count printed above from their stated parts and halting if one has movedNone: no maintained record, published table or filed document is opened
The words write up, span, grid, convention, signal, threshold and denominatorOrdinary working vocabulary that turns up wherever people count things and argue about the countingNone required

The Nakshatra unit, the six year record and the Ashwin rule are invented.
Educational material. Not advice on any investment, tax, budget or market position.

← PreviousNext →
Fin Maverick Micro CoursesExplore Micro Courses
Fin Maverick BootcampsExplore Bootcamps
Fin Maverick

Finance education that ends in a job, not a certificate that gathers dust. Built for young India.

LEARN
CalculatorsFrameworksComparisonsCareersShowdown
RESOURCES
All CoursesMicro CoursesBootcampsInternships
COMPANY
AboutJob openingPartnership
LEGAL
Privacy PolicyTerms & ConditionsContent LicenseReturn & Refund Policy
© 2026 FIN MAVERICK / BUILT FOR INDIA.DO FINANCE, DO NOT JUST READ ABOUT IT.