Fin Maverick
Foundations VocabularyAccounting & ReportingEconomics & MacroQuant Methods & ProgrammingBusiness & Company AnalysisCorporate Finance & ValuationBehavioural Finance
Banking & Market InfrastructureFixed Income & RatesDerivatives & Structured ProductsPublic EquitiesTransactions & DealsPortfolio ConstructionFunds & AMCs
Private Markets & AlternativesRisk, Treasury & ControlAI & Digital FinanceStochastic Calculus & PricingWealth & Personal FinanceIndian Markets & RegulationProfessional Practice
CalculatorComparison
Frameworks
Explore Bootcamps
Equity ResearchPortfolio ManagementMutual Fund MasteryFinancial LiteracyInvestment Banking Analyst
Private Equity AnalystHedge Funds AnalystBreaking Into VCBreaking Into QuantsAI For Finance
Financial Analyst ProgramRisk Management ProgramPrivate Wealth ManagementDebt Capital MarketsDerivatives Foundation
Explore Internships
Equity Research InternMutual Fund Intern
Portfolio Management InternFinancial Literacy Intern
Explore Micro Courses

Equity Research6

Writing an Investment ThesisBuilding a Discounted Cash FlowReading an Annual Report FastReading a Sector Before a CompanySpotting Quality of Earnings Red FlagsBuilding a Revenue Forecast From Drivers

Portfolio Management3

Rebalancing: When, Why and What It CostsStrategic and Tactical Asset AllocationMeasuring Risk in a Portfolio

Mutual Fund Mastery3

Comparing Funds Without Being FooledHow a NAV Is Struck and Which Day You GetReading a Fund Factsheet Properly

Derivatives Unlocked4

Hedging a Real ExposureThe Greeks, PracticallyFutures, the Basis and What Moves ItReading an Option Payoff

AI For Finance2

Retrieval and Grounding for FinanceDocument Extraction in Finance

Breaking Into Quants4

Backtesting a StrategyHypothesis TestingCleaning Financial DataRegression for Finance

Breaking Into VC3

Sizing a MarketReading a Term Sheet as a FounderHow a Venture Round Actually Works

Financial Analyst Program4

Common Size and Trend AnalysisReading a Cash Flow StatementRatio Analysis That Says SomethingBuilding a Working Capital Schedule

Risk Management Program2

Credit Exposure and How It Is ReducedValue at Risk and What It Hides

Investment Banking Analyst3

Precedent Transactions and Why They DifferReading a Term Sheet StructurallyBuilding a Comparable Companies Table

Private Wealth Management3

Tax Aware Portfolio DecisionsBuilding a Client Risk ProfileGoal Based Planning Arithmetic

Debt Capital Markets3

Analysing an Issuer's CreditDuration and What It Does Not Tell YouBond Pricing and Yield Mechanics

Private Equity Analyst2

Fund Waterfalls and CarryThe LBO in Structure

Hedge Funds Analyst2

Short Selling MechanicsLong Short Mechanics
Courses
Explore Career Roadmaps
Investment Banking AnalystEquity Research AnalystVC AnalystPrivate Equity AnalystHedge Funds Analyst
Quant AnalystAI For FinanceFinancial Analyst ProgramPrivate Wealth ManagementDebt Capital Markets
Risk Management ProgramDerivatives FoundationPortfolio ManagementMutual Fund Mastery
PartnershipsShowdown
Log inSign up
Quantitative Methods, Financial Data & Programming
1Probability
Probability in FinanceRandom VariableProbability DistributionsThe Normal DistributionNormal Distribution ProbabilityThe Lognormal DistributionRandomness vs Uncertainty
2Statistics and Inference
Population and SampleMean, Median and ModePrecision and AccuracyVariable TypesVariance, Standard Deviation and…Dispersion MeasuresStatistical BiasEffect SizeHypothesis TestingThe Sampling DistributionSkewnessKurtosisCovarianceConfidence IntervalArithmetic Mean vs Geometric MeanStatistical Significance vs Economic…Confidence Interval vs Prediction IntervalHow to Summarise a…
3Correlation and Regression
RegressionCorrelation and CausationOrdinary Least SquaresInteraction TermsRegression CoefficientsRegression vs ClassificationHow to Build a…Spurious CorrelationRegression, Correlation and FitResidualsMulticollinearityAutocorrelation and Partial Autocorrelation
4Time Series
Time Series in FinanceSimple, Weighted and Exponential…Moving Average CalculatorPrice, Return and Level SeriesHow to Prepare Time-Series…LagFrequencySeasonalityTimestampsTrendStationarity and the Unit RootHeteroskedasticityLeadRolling WindowsDifferencing
5Simulation and Numerical Methods
SimulationMonte Carlo SimulationHow to Run a…Numerical MethodsIterationResampling and the BootstrapPseudorandom Numbers and the SeedConvergence and ToleranceNumerical Stability
6Optimisation
OptimisationLocal and Global OptimaConstraintsConvex OptimisationThe SolverLinear ProgrammingThe Objective FunctionConstraint ViolationThe Feasible SetLagrange MultipliersQuadratic Programming
7Modelling Practice
Linear, Logistic, Ridge and…Training, Validation and Test…The ModelModel ErrorDependent and Independent VariablesThe ROC Curve and AUCWhat a Model HoldsMSE, RMSE, MAE and MAPEPrecision and RecallCross Validation and RegularisationOverfitting and UnderfittingReturn Series MeasuresSimple, Compound and Log Return
8Backtesting and Research Integrity
BacktestingBacktest vs Live PerformanceHow to Document a…How to Prevent Backtest…Out-of-Sample TestingWalk-Forward AnalysisMultiple TestingP-HackingData Snooping
9Data Quality and Structure
Data QualityThe DatasetSelection and Survivorship BiasVersioned DatasetsData Structures in FinanceData CleaningMissing Data and Null ValuesStructured Data vs Unstructured DataMissing Data vs ZeroData Validation vs Data CleaningOutliersDuplicate Records
10Programming for Finance
Data PipelinesAPIs for Financial DataAPI vs CSV FileDatabases in FinancePython for FinanceJoinsSQL for FinanceThe Analysis Workflow
11Quantitative Research
Research DesignThe Data Generating ProcessReproducibilityPeer Review in Analytical WorkThe Research HypothesisRobustness and Sensitivity

The Research Hypothesis: Making a Claim That Can Fail

Here is a sentence that appears somewhere, in some form, in a great deal of written work. This study investigates whether the outcome is related to the input. The sentence sounds like the start of real work, careful and even modest. And it is not a claim at all. No result could be put in front of the person who wrote it that would make them say, out loud, that they were wrong. The whole difference between a hypothesis and a topic is that somebody can lose a hypothesis, and the losing is arranged in advance rather than argued about afterwards.

The same thing happens outside any research setting. Two people at a wedding argue about whether the caterer undercharged. One of them says the food was worth more than what was billed. The other says it was not. Neither of them has said what would settle it, so the argument can go around all evening. Now one of them says: the caterer served four hundred plates and billed under Rs 250/- a plate, and if the plate count turns out to be under three hundred then I am wrong. The plate count sentence can be lost. Somebody can go and count.

Taken as settled here

Several earlier subjects are assumed here, and none of them is taught again. How long a record has to be before it can settle anything, what a mechanism is against the record it produced, whether a second person can reach the same number from the same data, and how finished work gets challenged, are all covered separately and are taken as given here. Only one thing is at issue below: the wording of the claim itself.

The fifty month record and the rule behind it. Somebody wrote down a generatorA rule written out in full that says which values come out and how often each one does. Because it is written down, its own average is known before any record is drawn from it.: five possible monthly values, each with a weight, the weights adding to exactly one. The rule has an average of 1.00 per cent and a spread of 5.00 per cent, and both are known only because somebody wrote the rule out. Fifty months were then drawn from it. The fifty months average 0.50 per cent, with a standard errorA measure of how far a figure taken from a record is likely to sit from the figure the record was drawn from. It shrinks as the record gets longer. of 0.7035 per cent, and a 95 per cent intervalA range reported around an estimate, saying which values the record is consistent with rather than which single value it produced. running from minus 0.8788 per cent to 1.8788 per cent. Every one of those readings was worked out separately and is used here rather than rederived.

The ten paired months, and the two things laid against them. Separately, and on a different and much shorter record, ten months of an input and an outcome sit side by side. A straight line through them has a fitA single number between zero and one saying how much of the outcome's variation the line accounts for. Nearer one means the line tracks the outcome more closely. of 0.7559. An invented count of chairs put out each month in a hall, averaging 60 chairs and connected to nothing, reads 0.9029 against the same ten months. And a near duplicate of the input, nudged by five hundredths in eight of the ten, brings the fit to 0.7560. Ten months is not fifty, and nothing from one record is ever added to the other.

The rule for calling a claim refuted, written before anything below. A claim here says the true mean is at least some stated figure. The fifty month record refutes such a claim when that stated figure sits outside its 95 per cent interval, and leaves the claim standing when the figure sits inside. The convention is a decision about how to count, not a discovery, so it travels beside every verdict below.

What makes a claim a hypothesis rather than a topic?

A claim that can be tested has three parts, and most written claims have the first two. The first part is a quantity a record can actually produce. Not an impression, not a mood, but something with a value: the average monthly change over fifty months, say. The second part is a stated value or direction for that quantity. Not larger, not meaningful, but a figure: at least 1.00 per cent. The two parts feel like a claim, and they are what most people write down and stop at.

The third part is the one that turns the first two into something testable, and it is the part almost everybody leaves out: the reading that would end the claim. Written in advance, before the record is opened, and written specifically enough that a second person reading the sentence would write down the same reading. Here that third part reads: the claim is over when 1.00 per cent sits outside the record's 95 per cent interval. Very little judgement is left in that. Somebody works out the interval, looks at where 1.00 per cent falls, and the answer is not up for discussion.

THREE PARTS, AND ONLY THE THIRD ONE HAS A WAY OUT OF IT Most written claims stop after the second panel and are then discussed for years. PART ONE, THE QUANTITY PART TWO, THE VALUE PART THREE, THE WAY OUT The average monthly change over fifty months. Something the record can actually produce. At least 1.00 per cent. A stated floor. Not a direction, and not a hope. 1.00 per cent sitting outside the record's 95 per cent interval. A claim that can be left. THE DOOR Parts one and two can be discussed forever. Part three is what makes them testable.
A claim needs a quantity a record can produce, a stated value for it, and the reading that would end it, and only that third part gives anybody a way out of the claim.

The everyday version is a street food stall two streets away. The owner is thinking about moving the stall to a new pitch outside a college gate. Saying the new pitch is better is a position: on a slow day at the new pitch he will say the college was on holiday, and on a good day at the old one he will say a wedding party came past. Nothing settles it. Now he says something else. He says the new pitch takes more than Rs 4,000/- a day on average, and that if the first full month there averages under Rs 4,000/- a day he goes back to the old pitch. The stall owner can lose that claim, and the losing was arranged before the first day of trading.

ONE QUESTION SORTS EVERY WRITTEN CLAIM INTO TWO KINDS Is there a reading that would end this claim? NO YES A POSITION No record can go against it, so no record can end it either. A HYPOTHESIS A record can go against it, which is the only reason running it is worth it. The question is asked of the wording, before any record is opened, and it takes one minute.
If no reading would end a claim then no record can go against it, and what is being defended is a position rather than a hypothesis.
Try it out

A colleague offers this claim for testing: the average monthly change is positive. What is missing before anybody can test it at all?

What does a claim that cannot fail look like?

A claim that cannot fail is easier to see than to describe. Take a real one, put things in front of it, and count how many of them it turns away. So here is the claim, in the exact wording an unhurried person might write: the outcome is related to something. Three candidatesThe things being tried against a claim, one at a time, before any of them has been chosen or ruled out. were tried against it, all three read on the same ten paired months, and none of them was picked in advance to make a point.

Try it out

The claim is that the outcome is related to something. Three candidates get tried against it: the ordered input, a count of chairs put out each month in a hall, and a near duplicate of the input. How many of the three refute the claim?

The ordered input reads 0.7559. The count of chairs reads 0.9029. The near duplicate reads 0.7560. All three satisfy the claim, and the number of candidates that refute it is zero. The count of refuting candidates is the figure worth sitting with. Not one, not a small number: zero. There was no observation available anywhere in this exercise that could have gone against the sentence. Running the work was never going to change what anybody believed.

Notice what the count of chairs does to the argument. The count of chairs is not a weak candidate that squeaked through. A count of furniture put out in a hall, connected to the outcome by nothing whatsoever, reads higher than the ordered input on the very measure the claim was checked with. If a claim is satisfied by that, the trouble is not that the claim is weak. The set of things that would have refuted the sentence is empty, so the trouble is that the sentence is not a claim. A weak claim is one a record could go against and probably will not. A record could not go against this one at all.

THREE CANDIDATES TRIED, THREE ADMITTED, NONE TURNED AWAY All three readings sit on the same ten paired months, and none was picked to prove a point. THE CLAIM AS WRITTEN the outcome is related to something THE ORDERED INPUT reads 0.7559 THE COUNT OF CHAIRS reads 0.9029 THE NEAR DUPLICATE reads 0.7560 ADMITTED, THE CLAIM IS SATISFIED TURNED AWAY, THE CLAIM IS REFUTED 3 0 the ordered input, the count of chairs, and the near duplicate nothing at all An empty refusing bin is the whole diagnosis. The claim was never at risk from anything.
Three candidates were tried against the claim that the outcome is related to something, reading 0.7559, 0.9029 and 0.7560, and all three satisfy it while none refutes it.
Try it out

A colleague writes down that the input matters. Which rewriting turns it into a claim a reading could actually end?

How is a claim stated so that a reading can refute it?

Write the claim as a stated floor. Not the outcome is related to something, but the true mean is at least some figure, with the figure written out. Then fix the conventionA choice about how something will be counted, written down in advance so that a second person can repeat it and reach the same verdict. for refuting it, in the same breath, before anything is opened. The convention used throughout is one sentence long: this record refutes a claim that the true mean is at least some stated figure when that figure sits outside the record's 95 per cent interval. The interval runs from minus 0.8788 per cent to 1.8788 per cent, and it was fixed by the length of the record and the spread inside it, not by anybody's preference.

Six claims run down that convention give the following verdicts. Not one of the six verdicts is a judgement call. Somebody with the interval in front of them and no opinion about the subject would produce the same six answers.

The claim, as writtenDoes the stated floor sit inside the interval?This record saysAnd the claim is
the true mean is at least 0.00 per centYes, insidesurvivestrue
the true mean is at least 0.50 per centYes, insidesurvivestrue
the true mean is at least 1.00 per centYes, insidesurvivestrue
the true mean is at least 1.50 per centYes, insidesurvivesfalse
the true mean is at least 2.00 per centNo, it sits above the upper endrefutedfalse
the true mean is at least 2.50 per centNo, it sits above the upper endrefutedfalse

The verdict turns exactly once, at 1.8788 per cent, and it never turns back. Four claims survive and two are refuted, and the place where the answer changes was settled by the record's length before a single claim had been written down. The turning point is worth holding on to. Nobody chose 1.8788 per cent. The figure fell out of fifty months and the spread inside them, and it would have sat where it sits whichever six claims anybody had decided to line up against it.

One more thing about the ladder, and it is hygiene rather than a finding. The nearest claim on the ladder, the one stating at least 2.00 per cent, sits 0.1212 per cent away from the turning point. Nothing on this ladder is decided by a hair. If a claim had been placed at, say, 1.88 per cent, its verdict would turn on the fourth decimal place of a figure nobody can measure that finely, and the honest thing then would be to say the record cannot separate that claim from its neighbour rather than to print a verdict.

SURVIVAL IS A POSITION ON A BAR, NOT A JUDGEMENT ABOUT A CLAIM The band was fixed by the record before any of the six claims below it had been written down. the turning point, 1.8788 per cent the record's 95 per cent interval minus 0.8788 1.8788 the claim's floor verdict and the claim is 0.00 0.50 1.00 1.50 2.00 2.50 survives survives survives survives refuted refuted true true true false false false Read the bottom two rows together: the verdict turns between 1.50 and 2.00, the truth turns between 1.00 and 1.50.
Across claims of at least 0.00 up to at least 2.50 per cent the verdicts run survives, survives, survives, survives, refuted and refuted, turning once at 1.8788 per cent.
Try it out

The record's 95 per cent interval runs from minus 0.8788 per cent to 1.8788 per cent. Before anything in the panel below is moved, is a claim that the true mean is at least 2.50 per cent refuted, or left standing?

Play with it

Slide the claim's floor and watch the verdict flip at a point nobody chose

One control, one consequence. The band is the record's 95 per cent interval and it never moves. The record is already written, and nothing done here can lengthen it. The claim's stated floor is what moves. As the floor slides through the six settings, two things happen at once: the verdict flips, exactly once, at a place fixed before any claim existed; and the grid underneath fills in. One cell of that grid stays empty on this record, and one fills up in a way that should give pause.

THE BAND IS THE RECORD. THE MARKER IS THE CLAIM. Only one of the two can be moved, and it is not the one that decides the verdict. true mean minus 1.00 minus 0.50 0.00 0.50 1.00 1.50 2.00 2.50 3.00 the claim: at least 1.00 per cent Shaded band: the record's 95 per cent interval, minus 0.8788 to 1.8788 per cent. It does not move. THIS RECORD LEAVES THE CLAIM STANDING and the claim is true, because the rule behind the record has a mean of exactly 1.00 per cent WHERE EVERY CLAIM VISITED SO FAR LANDS THIS RECORD LEAVES IT STANDING THIS RECORD REFUTES IT THE CLAIM IS TRUE THE CLAIM IS FALSE at least 1.00 per cent 1 visited so far nothing lands here 0 visited so far, and this stays empty nothing here yet 0 visited so far nothing here yet 0 visited so far
floorat least 1.00 per cent

Educational illustration. The fifty month record and the generator behind it are inventions built for teaching, and every figure in this panel is illustrative. The band is the record's own 95 per cent interval, worked from a mean of 0.50 per cent and a standard error of 0.7035 per cent, and it is the same band at every setting because the record does not change when the claim does. The true mean of 1.00 per cent is knowable here only because the generator was written down first. No real record offers that luxury. The convention for calling a claim refuted was fixed before any of the six claims was written, and nothing in this panel decides it.

Breaking Into Quants Bootcamp — Fin Maverick

Does surviving a test count as evidence for the claim?

Look again at two of the six claims, slowly. The claim that the true mean is at least 1.00 per cent survives, and it is true: the generator's mean is exactly 1.00 per cent, and somebody wrote that generator down. The claim that the true mean is at least 1.50 per cent also survives, and it is false, for the same reason and by the same figure. A true claim and a false claim came through the same record with the same verdict, so the verdict separated nothing.

Most readers of published research get the next sentence wrong, and it is worth putting bluntly. Surviving is not a small amount of evidence. Surviving is not weak support, not a hint, not a promising sign that firms up when somebody runs it again. On this record surviving is a statement about where a claim's floor sits relative to 1.8788 per cent, and 1.8788 per cent is a fact about how many months were collected. A claim at 1.50 per cent survived because fifty months cannot see the difference between 1.00 per cent and 1.50 per cent, not because there is anything to be said for 1.50 per cent.

ONE TRUE CLAIM, ONE FALSE CLAIM, AND ONE VERDICT FOR BOTH Both are read against the same fifty months, under the same convention, at the same moment. THE CLAIM: AT LEAST 1.00 PER CENT THE CLAIM: AT LEAST 1.50 PER CENT IS THE CLAIM TRUE? IS THE CLAIM TRUE? YES. The generator's mean is exactly 1.00 per cent. NO. The generator's mean is 1.00 per cent, below 1.50. DOES THE FLOOR SIT INSIDE THE INTERVAL? DOES THE FLOOR SIT INSIDE THE INTERVAL? Yes. It is under 1.8788. Yes. It is under 1.8788. AND THE VERDICT UNDER BOTH IS THE SAME: THIS RECORD LEAVES IT STANDING
A claim of at least 1.00 per cent is true and a claim of at least 1.50 per cent is false, and this record leaves both standing, so surviving separates nothing.

Set the six claims out against two questions instead of one. Is the claim true? Only the generator can answer that. Does this record refute the claim? Only the convention and the interval can answer that. Two questions give four cells to fill. Only three of the four fill up on this record, and the empty one is the cell where a true claim gets refuted. Everything true about this generator sits comfortably inside the band, so nothing lands in that cell at any setting on the ladder. The record never makes the loud, obvious mistake. The record makes the quiet one instead.

The quiet one is the cell holding the claim of at least 1.50 per cent: false, and standing. A survived claim cannot be reported as a supported claim, and that one cell is why. Had 1.50 per cent been written down before the record was opened, the verdict now in hand would read exactly like the verdict a correct claim gets. Nothing in the arithmetic distinguishes them, and nothing in a longer write up would either.

FOUR CELLS, THREE OCCUPIED, AND THE DANGEROUS ONE IS NOT EMPTY The row a claim sits in is settled by the generator. The column is settled by the record. THIS RECORD LEAVES IT STANDING THIS RECORD REFUTES IT THE CLAIM IS TRUE THE CLAIM IS FALSE at least 0.00 per cent at least 0.50 per cent at least 1.00 per cent three of the six claims land here nothing lands here at all no true claim on this ladder is refuted, so the record never errs loudly at least 1.50 per cent false, and left standing. This one cell is why surviving proves nothing. at least 2.00 per cent at least 2.50 per cent two of the six claims land here A verdict gives the column. Only the generator knows the row, and no real record comes with one.
On this record a true claim and a false claim both survive, so the verdict says which side of 1.8788 per cent a claim sits on and nothing about whether it is true.
Try it out

A claim of at least 1.00 per cent survives on this record, and so does a claim of at least 1.50 per cent. Exactly one of the two is true. What does surviving actually say about either of them?

AI For Finance Bootcamp — Fin Maverick

What does failing to reject a claim of nothing actually mean?

A second kind of claim turns up constantly, and it is the flattest one available: the true mean is zero, and nothing is going on. Run against the fifty month record under the usual arrangement, it produces a reading of 0.7107, with a two sidedCounting a departure in either direction as evidence against a claim, rather than only a departure one way. chance of about 0.4772. So the claim of nothing is not rejected. A great deal of careless writing begins with that sentence.

Try it out

The record's reading against a mean of zero comes out at 0.7107. Before anything else is examined, is the two sided chance attached to that reading likely to come out large or small?

Failing to reject a claim of nothing does not say the mean is zero, and on this record it is flatly not zero: the generator's mean is 1.00 per cent, positive, and written down before a single month was drawn. The failure says something narrower and much less interesting: fifty months cannot tell 1.00 per cent apart from zero. The statement is about the length of the record, and the design knew it before the record was opened. How far apart two values have to be before a record of a given length can separate them is arithmetic available in advance.

Here is the household version. A shopkeeper weighs a sack of rice on a bathroom scale that reads to the nearest kilogram, and it shows 50 kilograms both before and after a customer takes a handful. The scale has failed to detect a change. Nobody sensible concludes that no rice left the sack. The reasonable conclusion is that a handful is smaller than what this instrument can see, and that if the question really matters somebody should fetch a better scale rather than write down that the sack is unchanged.

NOT REJECTED IS A DISTANCE, AND ZERO IS A PLACE Two different scales, because the two statements are not about the same thing at all. the record's reading, 0.7107 the two sided threshold, 1.9600 0.00 0.50 1.00 1.50 2.00 2.50 in here a mean of zero is not rejected, and this record sits here out here it would be rejected AND SEPARATELY, ON A SCALE OF PER CENT, WHERE THE TRUTH ACTUALLY SITS a mean of zero the true mean, 1.00 per cent minus 0.8788 1.8788 Both sit inside the record's interval, which is exactly why fifty months cannot separate them. The upper line says what the record could see. The lower line says what was there. They are not the same statement.
The record's reading against zero is 0.7107 with a two sided chance of about 0.4772, so a mean of zero is not rejected, and the true mean is 1.00 per cent and positive.
Try it out

The record fails to reject a mean of zero. Is the mean zero?

Why is the claim written before the record is opened?

Wording is only half of the discipline. The other half is order, and order costs nothing at all: the same words, written at a different time. Once the interval has been seen the turning point is visible, so a claim written afterwards can always be placed on the surviving side of it. Nothing dishonest has to happen for this to go wrong. Somebody works out the interval, looks at it, thinks about what they always suspected, and writes down a claim that fits. Nobody is lying. The person genuinely believes they suspected it.

The smallest version of the problem is right here on this ladder. A claim of at least 1.50 per cent survives. A claim of at least 2.00 per cent does not. Anybody who has seen that the interval stops at 1.8788 per cent can choose which of those two they had believed all along, and both choices will look reasonable in writing. Neither wording is any more informative than the other, and both cost nothing to produce. A verdict that costs nothing to obtain is worth nothing when obtained.

THREE DECISIONS ARE LOCKED BEFORE THE FOURTH GATE OPENS Same words, different order. The order is the entire safeguard, and it costs nothing. GATE ONE GATE TWO GATE THREE GATE FOUR The question and the quantity are written down. The convention for refuting the claim is fixed. The claim's stated floor is written down. The record is opened and read, and the verdict falls out. locked locked locked opened here Everything left of gate four was decided without seeing a single month of the record.
The convention for refuting the claim is fixed before the record is opened, so the turning point at 1.8788 per cent was decided by the record's length rather than by anybody's preference.

There is a very ordinary version of this in any household that has ever bet on a cricket match after the fact. Claiming afterwards to have known the chase was gone once the fourth wicket fell costs nothing and cannot be checked, and everybody in the room has an equally good memory of having known it. Saying it before the fourth wicket falls costs something, and it is the only version anybody can be held to. Research works the same way and for the same reason, and writing the claim down first is the whole of the mechanism.

Try it out

Somebody reads the interval first, and then writes down the claim they say they had held all along. Which claims are available to them?

Risk Management Program Bootcamp — Fin Maverick Bond Pricing and Yield Mechanics — free micro-course from Fin Maverick

What does a written claim have to pass before any data is read?

A working analyst, a credit reviewer or anybody signing off an internal note can run the check that follows on a Tuesday morning without any arithmetic whatsoever. Four questions, asked of the wording alone, before the record is opened. Each question takes about a minute, and every one can be answered from the sentence itself. Nothing in research is cheaper quality control.

One. Is there a reading that would end this claim? Not in principle, not eventually: it has to be written out. If that sentence is vague, the claim is vague, and no amount of careful work downstream will fix that. Two. Is that reading one this record could actually produce? A claim ended only by something a fifty month record cannot deliver is untestable here even if it is testable somewhere. Three. Would somebody else, handed the same wording, write down the same refuting reading? If two competent readers write down different endings, the claim has not been stated, it has been gestured at, and whichever ending suits the result will be the one that gets used.

Four, and this is the one that catches a claim written to survive: is the claim narrow enough that a count of chairs could not satisfy it? The count of chairs is the sharpest test available, so the question is worth asking in exactly those words. If something unrelated can be imagined coming through the claim, the claim is admitting everything, and the work about to be signed off will be reported as holding no matter what the record says. Tests one to three catch sloppiness. Test four catches a sentence that was written so that it could not lose, perhaps without anybody meaning to.

FOUR NOTES IN THE MARGIN, WRITTEN BEFORE THE RECORD IS OPENED Every one of them is answered from the sentence itself, with no arithmetic anywhere. THE CLAIM, AS DRAFTED The true mean of the monthly change is at least 1.00 per cent. This record refutes the claim when 1.00 per cent sits outside the record's 95 per cent interval, which runs from minus 0.8788 to 1.8788 per cent. Written and dated before the record was opened. TEST ONE, A NAMED READING Yes. The claim says which reading would end it. TEST TWO, IS IT REACHABLE Yes. Fifty months produce an interval, which is all it needs. TEST THREE, WOULD ANOTHER READER Yes. The convention is one sentence and leaves no choice. TEST FOUR, NARROW ENOUGH Yes. A count of chairs cannot satisfy a stated floor on this quantity. Test four is the one that catches a claim written to survive, and it is the one people skip.
The fourth test asks whether the claim is narrow enough that a count of chairs could not satisfy it, and it is the test that catches a claim written to survive.

A lender's credit committee uses the same four questions without ever calling them that. When somebody brings a paper saying a borrower's collections have improved, the useful question in the room is never whether the paper is well argued. The question is: what would the committee have had to see this quarter to be told the collections had not improved, and was that written down last quarter or is it being invented now? A household saves the same way. Deciding to try a cheaper vegetable market for a month is only a decision if somebody says in advance what the monthly bill would have to come to before the household goes back to the old one.

Try it out

A written claim passes the first three tests cleanly and fails the fourth. What has gone wrong with it?

The failure: a claim that held, and could not have done anything else

An analyst frames the work as a question about whether the outcome is related to anything at all, does the analysis carefully, and reports that the claim held. And it did hold. The holding is what makes this failure so hard to catch from the outside: nothing in the write up is false, no arithmetic is wrong, and the person is not being careless.

The claim held because it could not have done anything else. Three candidates were tried and all three satisfy it: the ordered input at 0.7559, the invented count of chairs at 0.9029, and the near duplicate at 0.7560. The number of candidates that would have refuted the claim is zero. There was no observation available anywhere in the exercise that could have gone the other way, and a verdict with no losing case behind it is not a result.

The cost lands later and lands hard. The work gets cited as a confirmed finding. Somebody builds the next study on top of it, and somebody after that builds on them. When the whole line eventually falls over, nobody can point to the cell, the month or the reading that should have stopped it. There never was one. A missing reading is a much worse position than a wrong number. Somebody can at least find a wrong number.

And the fix is a wording change that costs nothing. State the claim as a stated floor with a figure in it, and state the convention for refuting it in the same breath. Before the record is opened, both the writer and the reader can then name the readings that would have ended the claim. Same work, same record, same afternoon. Only the sentence changes, and the sentence is what made the difference between a finding and a formality.

Where this guide stops. Fixing the whole design, choosing the record and working out what a record of a given length can settle are covered separately. The difference between a mechanism and the record it produced, and whether a second person can reach the same number from the same data, is also covered separately. Structured challenge of finished work is covered separately. Whether a result survives a changed assumption, and which assumption it is most sensitive to, is covered separately. The arithmetic of a test statistic, how a threshold gets chosen and what happens when many claims are tested at once are all covered separately.

And what none of this claims. A claim that survives is not therefore probably wrong, and a refuted claim was not therefore better written. The verdict and the truth are two different questions, three of their four combinations actually occur on this record, and the wording of the claim decides whether the verdict was ever capable of telling a reader anything.

Four questions of the wording alone. See what a written claim must pass.

Is there a source to check, and what happens when there is not?

Every reading in this guide was produced by writing a rule down, drawing records from it and working the arithmetic; those derivations are covered separately. A figure with no outside source still has to be checkable, and one of these is checked by rerunning the script that made it rather than by looking it up. Putting the rule on paper first is the whole point: the true mean of 1.00 per cent is knowable only because somebody wrote it, and no record anywhere outside a lesson arrives with its own answer key attached.

Reading used hereWhat produced itWhere it is recheckedOutside source
The five value rule, its mean of 1.00 per cent and its spread of 5.00 per centWritten out as five values and five weights adding to exactly oneThe checking script kept beside these notesNone used
Record one's mean of 0.50 per cent, its standard error of 0.7035 per cent, and its interval from minus 0.8788 to 1.8788 per centFifty months drawn from that ruleThe same script, recomputed rather than copied acrossNone used
The reading against zero of 0.7107 and its two sided chance of about 0.4772The same fifty months, worked against a mean of zeroThe same scriptNone used
The three candidate readings of 0.7559, 0.9029 and 0.7560Ten paired months, with a count of chairs and a near duplicate input laid against the same tenThe same script, which also checks that all three sit on ten months and not fiftyNone used

The fifty month record, the five value rule behind it, the ten paired months of an input and an outcome, the count of chairs put out in a hall and the near duplicate input are invented.
Educational material. Not advice on any investment, tax, budget or market position.

← PreviousNext →
Fin Maverick Micro CoursesExplore Micro Courses
Fin Maverick BootcampsExplore Bootcamps
Fin Maverick

Finance education that ends in a job, not a certificate that gathers dust. Built for young India.

LEARN
CalculatorsFrameworksComparisonsCareersShowdown
RESOURCES
All CoursesMicro CoursesBootcampsInternships
COMPANY
AboutJob openingPartnership
LEGAL
Privacy PolicyTerms & ConditionsContent LicenseReturn & Refund Policy
© 2026 FIN MAVERICK / BUILT FOR INDIA.DO FINANCE, DO NOT JUST READ ABOUT IT.