Fin Maverick
Foundations VocabularyAccounting & ReportingEconomics & MacroQuant Methods & ProgrammingBusiness & Company AnalysisCorporate Finance & ValuationBehavioural Finance
Banking & Market InfrastructureFixed Income & RatesDerivatives & Structured ProductsPublic EquitiesTransactions & DealsPortfolio ConstructionFunds & AMCs
Private Markets & AlternativesRisk, Treasury & ControlAI & Digital FinanceStochastic Calculus & PricingWealth & Personal FinanceIndian Markets & RegulationProfessional Practice
CalculatorComparison
Frameworks
Explore Bootcamps
Equity ResearchPortfolio ManagementMutual Fund MasteryFinancial LiteracyInvestment Banking Analyst
Private Equity AnalystHedge Funds AnalystBreaking Into VCBreaking Into QuantsAI For Finance
Financial Analyst ProgramRisk Management ProgramPrivate Wealth ManagementDebt Capital MarketsDerivatives Foundation
Explore Internships
Equity Research InternMutual Fund Intern
Portfolio Management InternFinancial Literacy Intern
Explore Micro Courses

Equity Research6

Writing an Investment ThesisBuilding a Discounted Cash FlowReading an Annual Report FastReading a Sector Before a CompanySpotting Quality of Earnings Red FlagsBuilding a Revenue Forecast From Drivers

Portfolio Management3

Rebalancing: When, Why and What It CostsStrategic and Tactical Asset AllocationMeasuring Risk in a Portfolio

Mutual Fund Mastery3

Comparing Funds Without Being FooledHow a NAV Is Struck and Which Day You GetReading a Fund Factsheet Properly

Derivatives Unlocked4

Hedging a Real ExposureThe Greeks, PracticallyFutures, the Basis and What Moves ItReading an Option Payoff

AI For Finance2

Retrieval and Grounding for FinanceDocument Extraction in Finance

Breaking Into Quants4

Backtesting a StrategyHypothesis TestingCleaning Financial DataRegression for Finance

Breaking Into VC3

Sizing a MarketReading a Term Sheet as a FounderHow a Venture Round Actually Works

Financial Analyst Program4

Common Size and Trend AnalysisReading a Cash Flow StatementRatio Analysis That Says SomethingBuilding a Working Capital Schedule

Risk Management Program2

Credit Exposure and How It Is ReducedValue at Risk and What It Hides

Investment Banking Analyst3

Precedent Transactions and Why They DifferReading a Term Sheet StructurallyBuilding a Comparable Companies Table

Private Wealth Management3

Tax Aware Portfolio DecisionsBuilding a Client Risk ProfileGoal Based Planning Arithmetic

Debt Capital Markets3

Analysing an Issuer's CreditDuration and What It Does Not Tell YouBond Pricing and Yield Mechanics

Private Equity Analyst2

Fund Waterfalls and CarryThe LBO in Structure

Hedge Funds Analyst2

Short Selling MechanicsLong Short Mechanics
Courses
Explore Career Roadmaps
Investment Banking AnalystEquity Research AnalystVC AnalystPrivate Equity AnalystHedge Funds Analyst
Quant AnalystAI For FinanceFinancial Analyst ProgramPrivate Wealth ManagementDebt Capital Markets
Risk Management ProgramDerivatives FoundationPortfolio ManagementMutual Fund Mastery
PartnershipsShowdown
Log inSign up
Quantitative Methods, Financial Data & Programming
1Probability
Probability in FinanceRandom VariableProbability DistributionsThe Normal DistributionNormal Distribution ProbabilityThe Lognormal DistributionRandomness vs Uncertainty
2Statistics and Inference
Population and SampleMean, Median and ModePrecision and AccuracyVariable TypesVariance, Standard Deviation and…Dispersion MeasuresStatistical BiasEffect SizeHypothesis TestingThe Sampling DistributionSkewnessKurtosisCovarianceConfidence IntervalArithmetic Mean vs Geometric MeanStatistical Significance vs Economic…Confidence Interval vs Prediction IntervalHow to Summarise a…
3Correlation and Regression
RegressionCorrelation and CausationOrdinary Least SquaresInteraction TermsRegression CoefficientsRegression vs ClassificationHow to Build a…Spurious CorrelationRegression, Correlation and FitResidualsMulticollinearityAutocorrelation and Partial Autocorrelation
4Time Series
Time Series in FinanceSimple, Weighted and Exponential…Moving Average CalculatorPrice, Return and Level SeriesHow to Prepare Time-Series…LagFrequencySeasonalityTimestampsTrendStationarity and the Unit RootHeteroskedasticityLeadRolling WindowsDifferencing
5Simulation and Numerical Methods
SimulationMonte Carlo SimulationHow to Run a…Numerical MethodsIterationResampling and the BootstrapPseudorandom Numbers and the SeedConvergence and ToleranceNumerical Stability
6Optimisation
OptimisationLocal and Global OptimaConstraintsConvex OptimisationThe SolverLinear ProgrammingThe Objective FunctionConstraint ViolationThe Feasible SetLagrange MultipliersQuadratic Programming
7Modelling Practice
Linear, Logistic, Ridge and…Training, Validation and Test…The ModelModel ErrorDependent and Independent VariablesThe ROC Curve and AUCWhat a Model HoldsMSE, RMSE, MAE and MAPEPrecision and RecallCross Validation and RegularisationOverfitting and UnderfittingReturn Series MeasuresSimple, Compound and Log Return
8Backtesting and Research Integrity
BacktestingBacktest vs Live PerformanceHow to Document a…How to Prevent Backtest…Out-of-Sample TestingWalk-Forward AnalysisMultiple TestingP-HackingData Snooping
9Data Quality and Structure
Data QualityThe DatasetSelection and Survivorship BiasVersioned DatasetsData Structures in FinanceData CleaningMissing Data and Null ValuesStructured Data vs Unstructured DataMissing Data vs ZeroData Validation vs Data CleaningOutliersDuplicate Records
10Programming for Finance
Data PipelinesAPIs for Financial DataAPI vs CSV FileDatabases in FinancePython for FinanceJoinsSQL for FinanceThe Analysis Workflow
11Quantitative Research
Research DesignThe Data Generating ProcessReproducibilityPeer Review in Analytical WorkThe Research HypothesisRobustness and Sensitivity

MSE, RMSE, MAE and MAPE: Four Error Measures Compared

All four read the same ten misses and disagree only about how to add them up. Across the ten months where the Nakshatra unit is paired with the Vasant unit the mean squared error is 21.80, its square root is 4.6690, the mean absolute error is 3.60, and the mean absolute percentage error is 110.81 per cent. One unchanged straight rule, one unchanged column of misses, four different answers.

Two invented columns run down a ten month record. The Nakshatra unit is the input and the Vasant unit is the outcome, and earlier reading already put a straight rule through the ten pairs and worked out how far that rule sits from the truth in each month. Each of those distances is a miss. The ten misses are what all four measures read. Where the misses came from is covered under the fitting of the rule. The ten misses now have to be squashed into one number for the top of a report.

There is no neutral way to do that. Four measures are in common use. All four take exactly the same ten misses as their input. None of them touches the rule that produced those misses. And yet one of them announces 21.80, one announces 4.6690, one announces 3.60, and one announces 110.81 per cent. Each of those is a correct answer to a slightly different question, and the differences between the questions are decisions somebody made about what counts as bad.

What do all four measures actually read?

Here are the ten misses, in time orderThe order the months actually arrived in, first to tenth. Records like this one are cut along that order when a rule is tested on months it has not seen, so shuffling the rows would destroy something. and in percentage pointsThe unit that results from subtracting one per cent figure from another. A move from 4 per cent to 7 per cent is three percentage points, and calling it three per cent would mean something quite different.: a miss of 1 in month 1, 9 in month 2, 8 in month 3, nothing at all in month 4, minus 5 in month 5, nothing again in month 6, minus 3 in months 7 and 8, minus 2 in month 9 and minus 5 in month 10. A positive miss means the rule came in under the outcome and a negative one means it came in over. The ten add to exactly nothing, a property of how the rule was fitted rather than a coincidence.

All four measures start from that same row of ten. The rule does not change between the four measures. Every disagreement below is manufactured entirely by the adding up. When two reports about the same work quote different error figures, the honest first question is not which team built the better rule. The question is which of the four measures each team reached for.

A bus runs a ten stop route. At each stop it is a few minutes early or a few minutes late, and by the end of the day there are ten numbers on the conductor's sheet. Now somebody upstairs wants one number for how the timetable held up. One sheet adds up how far off the bus was, ignoring early and late. A second adds up the squares so that a fifteen minute mess counts for far more than five three minute wobbles. A third takes a square root at the end so the answer comes back in minutes. A fourth divides each delay by the gap it was meant to leave. Being three minutes off then matters more on a five minute headway than on an hourly one. Four sheets, four numbers, one bus, one day.

ONE ROW OF MISSES, FOUR WAYS OUT OF IT Nakshatra unit and Vasant unit, invented. Misses in percentage points, months in time order. 0 1 9 8 0 minus 5 0 minus 3 minus 3 minus 2 minus 5 m1 m2 m3 m4 m5 m6 m7 m8 m9 m10 square each, average 21.80 squared points mean squared error square, average, root 4.6690 percentage points root mean squared error take size, average 3.60 percentage points mean absolute error divide, then average 110.81 per cent mean absolute percentage error The row of bars is identical for all four. Only the box changes. The fourth box is the only one that also reads the outcomes, not just the misses.
The same ten misses feed all four measures, so the four different headline numbers are produced entirely by the adding up rather than by any change in the rule that made the misses.
Try it out

Two reports describe the same fitted rule on the same ten months. One quotes an error of 3.60 and the other quotes 21.80. What has happened?

What does squaring do that taking the size does not?

Take each of the ten misses, multiply it by itself, and average the results. Multiplying a negative number by itself gives a positive one, so direction disappears on its own without anybody having to strip it out. The ten squares are 1, 81, 64, 0, 25, 0, 9, 9, 4 and 25. The ten squares add to 218, and dividing by ten gives the mean squared error of 21.80. The 218 and the 21.80 are the same quantity with a division by ten between them, and the digits agreeing is arithmetic rather than a misplaced decimal point.

Now look at what the squaring did to month 2. Its miss was 9, the largest on the record. As a plain size it is a quarter of the total: 9 out of a summed size of 36. As a square it is 81 out of 218, a share of the totalOne item's contribution divided by the sum of every item's contribution, written as a per cent. The share answers how much of the whole one row is responsible for, and the shares always add to one hundred across all rows. of 37.16 per cent. The same month, the same miss, and its weight in the headline figure jumped by half again simply because somebody chose to square before averaging.

Push it further with a cleaner case. Compare one miss of 9 against three separate misses of 3. By plain size those are identical: nine points of error either way. By squaring they are not remotely identical. The single miss of 9 squares to 81. The three misses of 3 square to 9 each, 27 in total. Squaring rules that one big miss is three times as bad as three small ones adding to the same amount, and nothing in the formula announces that it is making such a ruling.

ONE MISS OF 9, OR THREE MISSES OF 3? Each panel splits its own total between the two candidates. Nine points of error either way. SQUARE FIRST, THEN COMPARE 75.00% 25.00% one miss of 9 squares to 81 three misses of 3 square to 27 together the single miss takes three quarters of the pair TAKE THE SIZE, THEN COMPARE 50.00% 50.00% one miss of 9 size 9 three misses of 3 size 9 together a dead heat, and the panel says so
Squaring gives one miss of nine three quarters of the total against three misses of three, while taking the size calls the two exactly equal at fifty per cent each.
Try it out

One miss of 9, or three separate misses of 3. Which arrangement does squaring prefer?

Breaking Into Quants Bootcamp — Fin Maverick

Mean Squared Error vs Mean Absolute Error: where exactly do the two part company?

Take the size of each miss, throw the direction away, and average the ten. The sizes are 1, 9, 8, 0, 5, 0, 3, 3, 2 and 5. The ten sizes add to 36, and the mean absolute error is 3.60. Under this measure month 2 carries 9 of the 36, or 25.00 per cent, against the 37.16 per cent it carried under squaring.

Now the honest part, and it matters more than the headline. On this record the squared measure and the size measure agree about the whole order, from the worst month to the best, and that has to be said plainly rather than dressing the pair up as rivals. Month 2 is first under both. Month 3 is second under both. Months 5 and 10 tie for third under both. A bigger size always squares to a bigger square, so squaring a set of sizes cannot reshuffle them. The two measures do not part company over which month they point at. They part company over how much of the blame that month is made to carry: 37.16 per cent against 25.00 per cent, half as much again.

HOW MUCH OF THE BLAME DOES MONTH 2 CARRY? Both bars are one hundred per cent tall. Only the size of the top slice differs. 37.16% month 2 the other nine months squared error, total 218 81 of it belongs to month 2 25.00% month 2 the other nine months absolute error, total 36 9 of it belongs to month 2 the slice shrinks but stays first either way
Month two heads the record under both measures, and the two differ only in how much of the total blame it is made to carry, thirty seven per cent against twenty five.
Try it out

Month 2 carries 37.16 per cent of the squared error and 25.00 per cent of the absolute error. Do the two measures disagree about which month went worst?

So if the two never reorder the months, is the choice between them empty? The choice is not empty, and there is a sharper way to see the difference than any share of a total. Put to each measure the simplest question a measure can be asked: forget the fitted rule entirely, and pick one single number to use as the answer in all ten months. Which single number would each measure choose?

Feed that question to the squared measure and it lands on the plain average of the ten outcomes, 2.00 per cent, and it lands there uniquely. Guess 2.00 and the mean squared deviation is 89.30. Guess anything else at all and it goes up: 90.30 at a guess of 1.00 or 3.00, 93.30 at nothing, 98.30 at minus 1.00. There is exactly one winner and every step away from it is punished.

Feed the same question to the size measure and something quite different happens. Guess minus 1.00 per cent and the mean absolute deviation is 7.50. Guess nothing at all and it is 7.50. Guess 1.00, or 2.00, or 2.50, and it is still 7.50. Every guess from minus 1.00 per cent up to 2.50 per cent ties for first place under the size measure, and only outside that stretch does the figure start to climb. Those two ends are the fifth and sixth outcomes when the ten are put in order, so what the size measure has actually chosen is the medianThe middle reading once a column is placed in order, smallest to largest. With an even count of rows there is no single middle, so anything between the two central readings sits equally in the middle., and with ten readings there is no single middle to land on.

The two measures punish different things. The squared measure punishes distance, so one outcome far away drags the answer toward itself, and the answer it settles on is the balance point of the whole column. The size measure punishes presence, so an outcome far away counts once, the same as one nearby, and the answer it settles on is whatever has half the column on either side. Month 2's outcome of 18.50 per cent pulls the squared answer up hard; it barely troubles the size answer at all. The two measures never disagree about the order of the misses, and they disagree completely about what number to aim at in the first place.

ONE GUESS FOR ALL TEN MONTHS. WHICH ONE? The fitted rule is set aside here. Both panels score a single flat guess against the ten Vasant unit outcomes. every guess in this band ties on size SQUARED: ONE WINNER, AND IT IS THE AVERAGE 89.30 at a guess of 2.00 98.30 at a guess of minus 1.00 125.30 SIZE: A WHOLE STRETCH OF JOINT WINNERS 7.50, flat all the way across minus 1.00 2.50 minus 4.00 6.00 the single guess offered to all ten months, in per cent. Vasant unit, invented.
Asked for one flat guess, the squared measure names 2.00 per cent and nothing else, while the size measure calls every guess from minus 1.00 to 2.50 per cent an equal winner.
Try it out

Under the size measure, every guess from minus 1.00 per cent to 2.50 per cent scores exactly 7.50. What does that stretch correspond to?

Why take a square root at the end?

The mean squared error of 21.80 has a problem that has nothing to do with weighting. Its unit is squared percentage points. Nobody has any feel for a squared percentage point. A squared percentage point is not a distance, it is not comparable with any other figure here, and setting it beside a miss of 9 in the same sentence compares quantities that are not the same kind of thing at all.

The square root of 21.80 is 4.6690. The root is back in percentage points, the same unit as the misses themselves, and now it can be read: the rule is off by about four and two thirds percentage points in a typical month. Along the row of ten sizes, 4.6690 sits sensibly among them, above the 3s and the 2 and below the 8 and the 9. Taking the root reorders nothing, improves nothing and changes no ruling about which miss counts more; it changes only whether a reader can interpret the number in front of them.

The root is easy to oversell here. Taking a square root of a positive number never flips an order. If one rule has a lower mean squared error than another, it has a lower root as well, always. The root is not a better judge. The root is a better label. The judging was all done by the squaring, several steps earlier, and it stays exactly as done.

THE SAME QUANTITY, ON TWO SCALES Only one of the two can be laid alongside a miss and read. SQUARED PERCENTAGE POINTS 0 10 20 21.80 mean squared error a unit nobody has a feel for PERCENTAGE POINTS 0 5 10 misses of 0 1 2 3 5 8 9 4.6690 root mean squared error sits among the misses square root no ranking moves
Taking the root moves the measure from squared percentage points onto the same scale as the misses themselves, without altering a single judgement about which miss counted more.
Try it out

Taking the square root of 21.80 gives 4.6690. What exactly does that step change?

AI For Finance Bootcamp — Fin Maverick

What does a percentage measure add, and what does it require?

The three measures so far all answer in points. Answering in points is fine when every month is the same size, and it stops being fine the moment they are not. Being off by 3 points in a month where the outcome was 17.00 per cent is a different kind of error from being off by 3 points where the outcome was 1.00 per cent, and none of the first three measures can tell those two apart.

The fourth measure can. Dividing each miss by the size of the outcome it missed and averaging those ten ratios gives the mean absolute percentage error. Month 1 was off by 1 on an outcome of 3.00 per cent, so its ratio is a third. Month 4 was off by nothing at all, so its ratio is nothing. Month 2 was off by 9 on an outcome of 18.50 per cent, so its ratio is a little under a half, and the biggest miss on the whole record turns into one of the smaller ratios. The measure is doing exactly what it was built to do: judging a miss against the size of what it was trying to hit.

Picture a vegetable seller. On a busy Sunday the stall takes Rs 4,000/-, and being wrong about that by Rs 300/- is a small mistake. On a wet Wednesday the stall takes Rs 100/-, and being wrong about that by the same Rs 300/- is not a small mistake at all. A measure in rupees calls those two errors identical. A measure in per cent calls the second one twelve times worse, closer to how the seller feels about it.

But that step relocated the outcome. The outcome now sits in the denominatorThe number underneath in a division, the one being divided by. As it shrinks toward nothing the answer to the division grows without limit. A denominator near nothing is always worth checking before the result can be trusted., and nothing whatever in the formula insists that it stay away from nothing. That is the requirement this measure carries and never states. Every outcome on the record must be comfortably clear of nothing, or the division will do something violent, and the formula will do it silently. Two of the ten outcomes on this record are not comfortably clear of nothing at all.

THE PERCENTAGE ERROR AGAINST THE SIZE OF THE OUTCOME Ten months, invented. Direction dropped on both axes, so only sizes are plotted. 0 100% 200% 300% the faint curve is a fixed miss of 3 points, divided by the outcome month 3, 320.00% missed by 8 on an outcome of 2.50 month 8, 300.00% missed by 3 on an outcome of 1.00 months 5 and 10, both 166.67% month 1, 33.33% month 7 month 9 month 6 month 4 month 2, 48.65% the biggest miss on the record the size of that month's outcome, in per cent, running from 0 on the left to 20 on the right
Percentage error stays flat and small wherever the outcome is large, and climbs steeply toward the left edge as the outcome the miss is divided by approaches nothing.
Try it out

Settle this one before the control below moves at all. Month 8 was missed by 3 percentage points on an outcome of minus 1.00 per cent. If that outcome shrinks toward nothing while the miss of 3 is held exactly where it is, what happens to month 8's percentage error?

Play with it

Shrink one outcome toward nothing and watch a single month swallow the whole measure.

One control, and it touches one number. The control moves month 8's outcome from minus 10.00 per cent up to minus 0.10 per cent, and every other figure on the record is pinned: month 8's miss stays at 3 percentage points throughout, the other nine months keep their own outcomes and their own misses, and the rule itself is never refitted. As the control slides, the bar for the whole record's percentage measure grows while the block belonging to the other nine months sits stubbornly still, the pair of bars on the right shows the miss holding steady against a vanishing outcome, and the strip along the foot reorders itself live as month 8 climbs the ranking. The default setting of minus 1.00 per cent is the record's own reading and reproduces 110.81 per cent exactly.

minus 0.10 per centmonth 8 at minus 1.00 per centminus 10.00 per cent
ONE DENOMINATOR, AND WHAT IT DOES TO THE HEADLINE Nakshatra unit and Vasant unit, invented. Only month 8's outcome moves. THE WHOLE RECORD, IN PER CENT 400% 200% 110.81% nine months, fixed the red block is month 8 alone MONTH 8, DRAWN TO SCALE the miss, held at 3.00 points 3.00 the outcome it is divided by minus 1.00 10 points across 3.00 divided by 1.00 is 300.00 per cent the numerator never moves. Only the thing underneath it does. THE TEN MONTHS RANKED BY PERCENTAGE ERROR, WORST FIRST Month 8 is the red tile. Exact ties are placed in month order.
Month 8's outcome
minus 1.00 per cent
Its own percentage error
300.00%
The record's measure
110.81%
Month 8's share of it
27.07%
Its place of ten
2nd
Educational illustration. Month 8's miss of 3 percentage points is held fixed at every setting of the control. The rule is neither improving nor worsening as the control moves. The mean squared error of 21.80, its root of 4.6690 and the mean absolute error of 3.60 do not appear in this panel for one reason: not one of them moves by any amount at any setting, because none of them ever looks at an outcome. All four figures are computed here from whole numbers of hundredths of a percentage point and rounded once, on the way to the screen.

Why does a rule that explains three quarters of the movement score above one hundred per cent?

The mean absolute percentage error on this record is 110.81 per cent. The R squaredA score running from nothing up to one that reports how much of an outcome column's variation a fitted rule managed to track. Its construction is covered separately. of the same rule on the same ten months is 0.7559. Both figures are correct. Neither is a typing error. And a reader meeting them in the same paragraph with no explanation is entitled to think one of them must be wrong.

Here is what is actually happening. Month 3's outcome was 2.50 per cent and month 8's was minus 1.00 per cent, the two smallest readings on the whole record. Month 3 was missed by 8, and 8 divided by 2.50 gives a percentage error of 320.00 per cent for that month alone. Month 8 was missed by 3, and 3 divided by 1.00 gives 300.00 per cent. Month 3 and month 8 carry 55.95 per cent of the entire percentage measure between them. The headline figure is more than half decided by the two months where the least was happening.

Notice what did not go wrong. The rule did not miss badly in those months. A miss of 3 is smaller than the record's average miss of 3.60, and it would not stand out at all in a list of sizes. The miss became enormous only after being divided by a very small number. The measure is not reporting a fault in the rule. The measure is reporting a fault in its own arithmetic, in exactly the same tone of voice it uses for everything else.

THE SMALL MISS THAT BEAT THE BIG ONE Both rows drawn on one scale, 12 pixels to a percentage point. Invented figures. MONTH 8 the miss 3 points, smaller than the record's average miss the outcome minus 1.00 per cent, the smallest but one on the record 300.00% 3 divided by 1 MONTH 2 the miss 9 points, the largest miss anywhere on the record the outcome 18.50 per cent, the largest on the record 48.65% 9 divided by 18.5 The smaller miss scores six times worse, and the only difference is what sat underneath. Nothing about the fitted rule differs between these two rows.
Month eight's miss of three scores six times worse than month two's miss of nine, purely because the outcome underneath it was eighteen times smaller.
Try it out

A rule reaching an R squared of 0.7559 on these ten months scores a percentage error of 110.81 per cent. Which of the two figures is wrong?

Risk Management Program Bootcamp — Fin Maverick Reading an Option Payoff — free micro-course from Fin Maverick

Do the four ever disagree about which month went worst?

One finding makes the choice of measure more than a matter of taste. Rank the ten months three times over. By squared error the worst is month 2, carrying 37.16 per cent of the total. By plain size the worst is month 2 again, carrying 25.00 per cent. By percentage error the worst is month 3, carrying 28.88 per cent, and month 2 has dropped to fifth place with 4.39 per cent of the total.

Two of the three measures agree on the entire order and the third reshuffles it, so the disagreement is not a difference of degree but a different answer about where to go and look. That distinction matters more than any change in share. If a share moves, the headline changes and the investigation goes to the same place. If the rankingThe list of items placed in order from worst to best by some measure. Two measures can hand back lists in different orders even when both are computed correctly from the same rows. moves, two people reading two correctly computed reports about the same rule will walk to two different months and start asking questions there.

THE SAME TEN MONTHS, RANKED THREE TIMES Worst at the top. Ten invented months, one unchanged fitted rule. BY SQUARED ERROR BY SIZE BY PERCENTAGE ERROR month 2 month 3 month 5 month 10 month 7 month 8 month 9 month 1 month 4 month 6 month 2 month 3 month 5 month 10 month 7 month 8 month 9 month 1 month 4 month 6 month 3 month 8 month 5 month 10 month 2 month 7 month 1 month 9 month 4 month 6 identical order The first two columns match row for row. The third is a different list entirely. The solid line follows month 2 from first place down to fifth. The dashed line follows month 8 up from sixth to second.
The squared and size rankings match row for row, while the percentage ranking lifts month three to the top and drops month two from first place to fifth.
Try it out

Under the percentage measure month 2 falls from first to fifth. Why does a reordering matter more than the change in its share of the total?

What should a report actually carry?

Somebody has to write one line at the top of a report and let a reader who will not check anything draw a conclusion from it. Four rules make that line honest, and none of them costs more than a few extra words.

Quote a measure in the units of the thing being predicted. That means the root rather than the raw squared figure: 4.6690 percentage points, not 21.80 squared percentage points. A reader can hold the first one against a miss of 9 and understand it. The second is a number in a unit that exists only inside the arithmetic.

Quote a second measure that weights the misses differently. Put the 3.60 beside the 4.6690. When the two sit close together, as they do here, the summary is not resting on how heavily one bad month was counted. When they sit far apart, one month is doing most of the work and the reader deserves to know that before drawing anything from either figure.

Never quote a percentage measure without also stating the smallest outcome in the record. On this record the smallest is minus 1.00 per cent, and once that is stated the 110.81 per cent stops looking like a verdict on the rule and starts looking like what it is. A percentage measure without its smallest denominator is a figure a reader cannot audit.

Name the worst single case under each measure quoted. Month 2 under the first two, month 3 under the third. Naming the worst case stops two readers walking away with two different investigations from one correctly computed report.

TWO WAYS TO WRITE THE SAME RESULT Both lines are arithmetically correct. Only one can be audited by the reader. THE LINE THAT CAN BE CHECKED Typical miss: 4.6690 percentage points Average size of a miss: 3.60 points Smallest outcome on the record: minus 1.00 per cent Worst month by both point measures: month 2 Worst by percentage error: month 3 Four lines. A reader can act on any of them. THE LINE THAT CANNOT 110.81% mean absolute percentage error Smallest outcome on the record Worst month under each measure A second measure for comparison One figure, and nothing to test it against. The struck lines are not wrong. They were simply left out, and leaving them out is the whole difference.
A report line carrying the root measure, a second measure and the smallest outcome can be audited, while the same result reduced to one percentage figure cannot be.
Try it out

What single fact has to travel alongside any percentage error quoted?

The failure: the figure that got quoted alone

Somebody puts 110.81 per cent at the top of a report and concludes that the rule is worthless. Or, worse, puts it in a summary two lines under an R squared of 0.7559 and leaves the reader to reconcile them. Neither figure is wrong and neither was computed carelessly. The rule really does account for about three quarters of the movement in the Vasant unit, and it really does miss month 8 by 3 percentage points on an outcome of minus 1.00 per cent, a percentage error of three hundred on that month by itself.

The measure was built for records whose outcomes stay well away from nothing, and this record is not one of them. Two of its ten outcomes sit at 2.50 per cent and minus 1.00 per cent, and those two months between them decide 55.95 per cent of the whole percentage figure. The problem is not in the rule. No amount of improving the rule fixes it. Improve every other month to a perfect hit and month 3 and month 8 would carry a hundred per cent of what remained rather than 55.95 per cent of it.

Here is the habit that prevents it, and it takes one look. Before a percentage error is quoted at all, the smallest outcome in the record has to be found and looked at. If it is anywhere near nothing, the measure is about to be dominated by the cases that matter least, and two choices remain: state the smallest outcome next to the figure so the reader can judge it, or quote a different measure. Printing the figure on its own and letting it stand as a verdict is not honest.

Where the misses come from in the first place, and which parts of them more months would shrink, is covered under the fitting of the rule. So is how a fitted rule is judged on months it has never seen, a different question from how the misses on months it has seen are added up. The measures used when the answer is a yes or a no rather than a size, where counting hits and misses replaces averaging distances, are covered separately, and none of the four here applies to them.

Whether an error of 4.6690 percentage points is large or small in any setting outside itself, and whether a rule fitting these ten invented months would be worth pointing at anything, are separate questions. Questions of that shape carry conditions of their own and belong where those conditions can be set out properly.

Three measures name one worst month and the fourth disagrees. See which metric decides.

Which of these figures could be knocked down, and how?

Every number above is either an invented reading or something a calculator settles in a minute, so the honest citation is an instruction to recompute rather than a link to follow. Each headline figure below carries the shortest route to catching it out.

The figureHow it was gotThe shortest way to catch it out
The ten misses, month by monthComposed for teaching, then subtracted: each outcome less what the straight rule returned for that monthAdd the ten together. They come to nothing at all. A column of misses that does not is not this column
21.80 and 4.6690The ten misses squared and averaged, then the square root of that average taken onceMultiply 4.6690 by itself and watch whether 21.80 comes back
3.60The size of each miss, direction thrown away, averaged over the ten monthsAdd the sizes: one, nine, eight, nothing, five, nothing, three, three, two, five. Divide by ten
110.81 per centEach miss divided by the size of the outcome it missed, those ten ratios averagedDo month 8 alone: three divided by one. If that is not three hundred per cent, nothing else here will close either
7.50, and the stretch that tiesThe ten outcomes scored against one flat guess, at every guess from minus 4.00 to 6.00 per centTry a guess of nothing and a guess of 2.00 by hand. Both give a summed size of 75 across the ten months
Any rate, threshold, period or standardNone is used anywhere above, because averaging a column of misses needs nothing issued by anybodyNothing to confirm, because no claim of that kind is made

The Nakshatra unit and the Vasant unit are invented.
Educational material. Not advice on any investment, tax, budget or market position.

← PreviousNext →
Fin Maverick Micro CoursesExplore Micro Courses
Fin Maverick BootcampsExplore Bootcamps
Fin Maverick

Finance education that ends in a job, not a certificate that gathers dust. Built for young India.

LEARN
CalculatorsFrameworksComparisonsCareersShowdown
RESOURCES
All CoursesMicro CoursesBootcampsInternships
COMPANY
AboutJob openingPartnership
LEGAL
Privacy PolicyTerms & ConditionsContent LicenseReturn & Refund Policy
© 2026 FIN MAVERICK / BUILT FOR INDIA.DO FINANCE, DO NOT JUST READ ABOUT IT.