Fin Maverick
Foundations VocabularyAccounting & ReportingEconomics & MacroQuant Methods & ProgrammingBusiness & Company AnalysisCorporate Finance & ValuationBehavioural Finance
Banking & Market InfrastructureFixed Income & RatesDerivatives & Structured ProductsPublic EquitiesTransactions & DealsPortfolio ConstructionFunds & AMCs
Private Markets & AlternativesRisk, Treasury & ControlAI & Digital FinanceStochastic Calculus & PricingWealth & Personal FinanceIndian Markets & RegulationProfessional Practice
CalculatorComparison
Frameworks
Explore Bootcamps
Equity ResearchPortfolio ManagementMutual Fund MasteryFinancial LiteracyInvestment Banking Analyst
Private Equity AnalystHedge Funds AnalystBreaking Into VCBreaking Into QuantsAI For Finance
Financial Analyst ProgramRisk Management ProgramPrivate Wealth ManagementDebt Capital MarketsDerivatives Foundation
Explore Internships
Equity Research InternMutual Fund Intern
Portfolio Management InternFinancial Literacy Intern
Explore Micro Courses

Equity Research6

Writing an Investment ThesisBuilding a Discounted Cash FlowReading an Annual Report FastReading a Sector Before a CompanySpotting Quality of Earnings Red FlagsBuilding a Revenue Forecast From Drivers

Portfolio Management3

Rebalancing: When, Why and What It CostsStrategic and Tactical Asset AllocationMeasuring Risk in a Portfolio

Mutual Fund Mastery3

Comparing Funds Without Being FooledHow a NAV Is Struck and Which Day You GetReading a Fund Factsheet Properly

Derivatives Unlocked4

Hedging a Real ExposureThe Greeks, PracticallyFutures, the Basis and What Moves ItReading an Option Payoff

AI For Finance2

Retrieval and Grounding for FinanceDocument Extraction in Finance

Breaking Into Quants4

Backtesting a StrategyHypothesis TestingCleaning Financial DataRegression for Finance

Breaking Into VC3

Sizing a MarketReading a Term Sheet as a FounderHow a Venture Round Actually Works

Financial Analyst Program4

Common Size and Trend AnalysisReading a Cash Flow StatementRatio Analysis That Says SomethingBuilding a Working Capital Schedule

Risk Management Program2

Credit Exposure and How It Is ReducedValue at Risk and What It Hides

Investment Banking Analyst3

Precedent Transactions and Why They DifferReading a Term Sheet StructurallyBuilding a Comparable Companies Table

Private Wealth Management3

Tax Aware Portfolio DecisionsBuilding a Client Risk ProfileGoal Based Planning Arithmetic

Debt Capital Markets3

Analysing an Issuer's CreditDuration and What It Does Not Tell YouBond Pricing and Yield Mechanics

Private Equity Analyst2

Fund Waterfalls and CarryThe LBO in Structure

Hedge Funds Analyst2

Short Selling MechanicsLong Short Mechanics
Courses
Explore Career Roadmaps
Investment Banking AnalystEquity Research AnalystVC AnalystPrivate Equity AnalystHedge Funds Analyst
Quant AnalystAI For FinanceFinancial Analyst ProgramPrivate Wealth ManagementDebt Capital Markets
Risk Management ProgramDerivatives FoundationPortfolio ManagementMutual Fund Mastery
PartnershipsShowdown
Log inSign up
Quant Analyst · CoreTrack
1Quantitative Methods, Financial Data & Programming
iProbability
Probability in FinanceRandom VariableProbability DistributionsThe Normal DistributionNormal Distribution ProbabilityThe Lognormal DistributionRandomness vs Uncertainty
iiStatistics and Inference
Population and SampleMean, Median and ModePrecision and AccuracyVariable TypesVariance, Standard Deviation and…Dispersion MeasuresStatistical BiasEffect SizeHypothesis TestingThe Sampling DistributionSkewnessKurtosisCovarianceConfidence IntervalArithmetic Mean vs Geometric MeanStatistical Significance vs Economic…Confidence Interval vs Prediction IntervalHow to Summarise a…
iiiCorrelation and Regression
RegressionCorrelation and CausationOrdinary Least SquaresInteraction TermsRegression CoefficientsRegression vs ClassificationHow to Build a…Spurious CorrelationRegression, Correlation and FitResidualsMulticollinearityAutocorrelation and Partial Autocorrelation
ivTime Series
Time Series in FinanceSimple, Weighted and Exponential…Moving Average CalculatorPrice, Return and Level SeriesHow to Prepare Time-Series…LagFrequencySeasonalityTimestampsTrendStationarity and the Unit RootHeteroskedasticityLeadRolling WindowsDifferencing
vSimulation and Numerical Methods
SimulationMonte Carlo SimulationHow to Run a…Numerical MethodsIterationResampling and the BootstrapPseudorandom Numbers and the SeedConvergence and ToleranceNumerical Stability
viOptimisation
OptimisationLocal and Global OptimaConstraintsConvex OptimisationThe SolverLinear ProgrammingThe Objective FunctionConstraint ViolationThe Feasible SetLagrange MultipliersQuadratic Programming
viiModelling Practice
Linear, Logistic, Ridge and…Training, Validation and Test…The ModelModel ErrorDependent and Independent VariablesThe ROC Curve and AUCWhat a Model HoldsMSE, RMSE, MAE and MAPEPrecision and RecallCross Validation and RegularisationOverfitting and UnderfittingReturn Series MeasuresSimple, Compound and Log Return
viiiBacktesting and Research Integrity
BacktestingBacktest vs Live PerformanceHow to Document a…How to Prevent Backtest…Out-of-Sample TestingWalk-Forward AnalysisMultiple TestingP-HackingData Snooping
ixData Quality and Structure
Data QualityThe DatasetSelection and Survivorship BiasVersioned DatasetsData Structures in FinanceData CleaningMissing Data and Null ValuesStructured Data vs Unstructured DataMissing Data vs ZeroData Validation vs Data CleaningOutliersDuplicate Records
xProgramming for Finance
Data PipelinesAPIs for Financial DataAPI vs CSV FileDatabases in FinancePython for FinanceJoinsSQL for FinanceThe Analysis Workflow
xiQuantitative Research
Research DesignThe Data Generating ProcessReproducibilityPeer Review in Analytical WorkThe Research HypothesisRobustness and Sensitivity
2Stochastic Calculus & Derivative Pricing Theory
iProbability Foundations
The Probability SpaceRandom VectorsSigma-AlgebraExpectationSample Space and EventsDensity and Distribution FunctionsRisk-Neutral ProbabilityState Price Density vs…
iiStochastic Processes and Jumps
Properties of a Stochastic ProcessMartingaleBrownian Motion and Its PropertiesBrownian Motion vs Geometric…Stopping TimeThe Markov PropertyState VariablesTransition ProbabilityQuadratic VariationQuadratic Variation vs Ordinary…Submartingale and SupermartingaleMartingale RepresentationMarkov Process vs MartingaleOptional StoppingFiltrationJump ProcessesThe Poisson ProcessLevy ProcessesJump Diffusion
iiiIto Calculus
The Ito IntegralThe Ito Integral vs the Riemann IntegralInfinitesimals in Stochastic CalculusQuadratic CovariationIto's LemmaHow to Apply Ito's…The Infinitesimal GeneratorIto Calculus vs Ordinary Calculus
ivStochastic Differential Equations
Stochastic Differential EquationsStochastic Differential Equation vs…Drift and DiffusionStrong and Weak Solutions ComparedDiscretisationGeometric Brownian Motion
vPricing Theory and No-Arbitrage
No-ArbitrageGirsanov, Radon-Nikodym and Change…Physical and Risk-Neutral Measures…The Fundamental Theorems of…The Law of One PriceThe Pricing KernelDiscount Factors and Zero-Coupon PricesReplication vs HedgingComplete Market vs Incomplete MarketClearing Margin Architecture
viOption Pricing Theory
European and American OptionsMonte Carlo European OptionThe Black-Scholes PDEBlack Scholes and the GreeksThe Payoff FunctionThe Binomial ModelBinomial Option PricingDelta Hedging in TheoryBoundary, Initial and Terminal ConditionsThe Exercise BoundaryHow to Check Put-Call…
viiVolatility Models
Constant, Local and Stochastic…Vasicek Model vs CIR ModelThe Heston ModelThe SABR ModelThe Volatility ProcessImplied VolatilityVolatility Smile vs Skew vs Surface
viiiInterest Rate Models
Interest-Rate DerivativesMean ReversionThe Zero-Coupon BondThe Ornstein-Uhlenbeck ProcessThe Discount CurveZero RatesShort-Rate Model vs Market Model
ixNumerical Pricing
Closed Form and Numerical…Monte Carlo PricingEuler and Milstein Schemes ComparedTree MethodsFinite Difference MethodsNumerical Error and StabilityVariance Reduction
xCalibration and Model Risk
Model OverrideMarket Price and Model PriceCalibrationHow to Document a Pricing ModelThe Educational Illustration LabelMarket ConventionsModel Uncertainty and LimitationsBacktesting a Pricing ModelIdentifiabilityCalibrated ParametersThe Calibration Loss Function

The SABR Model: Fitting the Smile in Rates and FX

The SABR model is built to reproduce a pattern of prices across strikes directly. Where the Heston model starts from a process and asks what pattern it produces, this one starts from the pattern and works back. Four parameters control it. One of them, the exponent, decides how the randomness scales with the level, and it has no counterpart in the previous model.

The reversal of direction repays a second reading. Everything below follows from it. One model is an argument that runs forwards: here is how variance moves, and the prices that come out are worked out from it. The other is an argument that runs backwards: here is the shape prices make across strikes, and a small set of numbers is then found that reproduces that shape. The two are not rival descriptions of the same thing; they are opposite directions of travel, and a direction of travel decides what a model is good for.

Consider something concrete. Suppose five thermometer readings have been taken along a corridor, at five marked points on the floor, and a rule is wanted for the temperature anywhere in the building. One way is to work out how heat moves: where the vents are, how the air circulates, what the walls do. With that right, something can be said about every room, including rooms nobody walked into. The other way is to draw the smoothest curve through those five marks. The curve will pass through all five marks beautifully. Asked about the basement, it has nothing to say. A curve through five marks was never a theory of heat.

The Heston model is the first way. The stochastic alpha, beta, rho (SABR) modelA model organised around the pattern of prices across strikes rather than around a process. is the second way, done carefully, with a process written down so that the curve is at least consistent with something moving. Being the second way is not a defect; it is a design choice with a specific payoff and a specific bill.

What is covered elsewhere

Fitting the model to observations is covered separately

The word fitting sits in the name this model is known by. Choosing parameters so that a model agrees with a set of observations is taken up separately and later, and a model and a fitted model are different objects. Every number below is a computed consequence of parameters written down by hand, never a parameter recovered from an observation.

What is this model built to do?

The standard process running through this whole subject area is a single invented traded quantity, written S with a time subscript, starting at Rs 100/-, carrying a volatility of 20 per cent a year, watched over one year against a risk-free rate of 5 per cent.

Under the constant treatment, that 20 per cent is one number for every strike. The earlier treatment of a single volatility across five strikes showed what happens when one number is asked to serve five contracts at five different strikes: it can be made to agree with any one of them and then misses the other four, cheap below and dear above, in a pattern that swings sign exactly once. Error scatters. A miss that changes sign exactly once does not. So the number is not enough, and something has to replace it.

The SABR model replaces it with a shape. The model sets down a process whose randomness is itself random and scales with the level of the process by an adjustable amount, and an expression then turns those parameters straight into a volatility for every strike. The whole apparatus exists so that a handful of parameters produces a curve across strikes, and the curve is the output the model was designed to deliver.

The ambition is a different one from the previous model's. The previous one wanted variance to behave: to be positive, to pull back to a long-run level, to be correlated with the process. Prices across strikes were a consequence to be worked out afterwards. Here, prices across strikes are the target, and the process is written down in whatever form makes that target reachable.

The model, as it is set down
$$ dS_t \;=\; \alpha_t\,S_t^{\,\beta}\,d\tilde W^{(1)}_t, \qquad d\alpha_t \;=\; \nu\,\alpha_t\,d\tilde W^{(2)}_t, \qquad d\tilde W^{(1)}_t\,d\tilde W^{(2)}_t \;=\; \rho\,dt $$
\(S_t\)the standard process, the single invented traded quantity, at time \(t\)
\(\alpha_t\)the level parameter, itself a random process, starting at \(\alpha_0=0.20\)
\(\beta\)the exponent, a fixed number between nought and one, set to 1 here
\(\nu\)the volatility of the level parameter, set to 0.30 here
\(\rho\)the correlation between the two Brownian motions, set to minus 0.7 here
\(\tilde W\)a Brownian motion under the pricing measure Q, one for each equation
What it says in wordsThe process moves with a randomness equal to a level parameter multiplied by the process raised to a fixed exponent, the level parameter is itself a random process with no drift of its own, and the two sources of randomness are correlated. Four numbers fix the whole thing: the starting level parameter, the exponent, the volatility of that level parameter, and the correlation.

Two things about that pair of equations are worth naming before anything else. First, there is no drift term on the process. The model is conventionally set down on a quantity that does not drift under the measure it is written under. Dropping the drift keeps the two equations as short as they are, and wherever a drift is needed it is carried separately. Second, the level parameter has no long-run level to return to. The level parameter has no mean reversion at all: it wanders, and where it wanders to is where it stays. The missing pull is the second thing this model gives up.

What does it mean to be organised around the pattern rather than the process?

A model organised around a process is judged by what its process does. Does variance stay positive? Does it pull back? Does the shape of its long-run distribution look like something? Each of those questions has an answer inside the model, and the answer is checkable without looking at a single price.

A pattern organisedBuilt to reproduce a shape across strikes rather than derived from a process and its consequences. model is judged by the shape it lays down across strikes. The process underneath is not fictional, and it obeys the same calculus as everything else in this subject area. The point is that the process was chosen for the shape it delivers, and the shape is the deliverable. The model was never organised to answer a question about the path of the process between two dates, even though it has a process and could be made to mumble an answer.

The everyday version runs as follows. A wall chart of measured temperatures at five recorded hours gives the temperature at those five hours perfectly. The chart was made from them. A curve drawn smoothly through those five marks also gives a reading at half past two. The curve connects two readings that were actually taken, so the half past two figure is real work. Asked about three in the morning, before the first mark, the same chart is extending a shape into a region no reading was taken in. The chart is not lying. Extending a shape is all it does, and there is nothing underneath the shape to keep it honest.

Same problem. Opposite ends. Read each panel downward. STARTS FROM A PROCESS Write an equation for how variance moves: positive, mean reverting, correlated. Then ask: what pattern of prices across strikes does that process produce? The pattern is the consequence. STARTS FROM A PATTERN Take the shape prices make across strikes as the thing to be reproduced. Then work back: which four numbers put that shape in place? The pattern is the target. A direction of travel decides what a model is good for. Reproducing the target is not the same achievement as deriving the consequence.
The Heston model starts from a process and asks what pattern of prices follows, while this model starts from the pattern across strikes and works back to the parameters that reproduce it, which is the reversal every other difference between the two comes from.
Try it out

Which of the two models starts from a process, and which starts from a pattern?

What are the four parameters, and what does each one control?

Four numbers, and each one has a job that can be stated in a sentence. The four are not interchangeable. Take them one at a time. A reader who blurs two of them will misread every result that follows.

The level parameterThe overall scale of the randomness at the start, set to 0.20 here, matching the locked volatility of the standard process. is the overall size of the randomness at the outset. The level parameter is set to 0.20, exactly the 20 per cent a year the standard process carries throughout this subject area. The match is neither a coincidence nor a fit. The value was written down deliberately, and the model therefore starts life sitting exactly where the constant treatment sits. Unlike the other three, it is not a fixed constant inside the model. The level parameter has its own equation, and it moves.

The exponentThe parameter deciding how the randomness scales with the level of the process, set to 1 here. decides how that randomness changes as the process moves to a different level. The exponent is the one parameter with no counterpart at all in the previous model, and it spans the two treatments set against each other earlier. A block of its own follows below.

The volatility of volatilityHow random the level parameter itself is, set to 0.30 here, the same numeral the previous model carries in the corresponding role. says how random the level parameter itself is. Set it to nought and the level parameter never moves, and the model collapses back to something with no randomness in its randomness. Set it to 0.30, the value used here, and the level parameter wanders. The number 0.30 is deliberate: the previous model carries the same numeral in the role that corresponds to it.

The correlationHow the level parameter moves with the process, set to minus 0.7 here, the same figure the case parameters carry. says how the two move together. At minus 0.7 the level parameter tends to rise when the process falls. The single negative sign tilts the resulting shape across strikes rather than leaving it symmetric. The same minus 0.7 appears elsewhere in this subject area, where the two locked paths have a quadratic covariation of exactly minus 0.028000 and therefore an implied correlation of exactly minus 0.7.

Three of those four numbers were chosen to match the previous model exactly. The exponent is the only thing that is new. Matching them is the whole point of the setting. On two unrelated parameter sets, every difference in output could have come from any of eight numbers, and a comparison like that teaches nothing. With three held fixed, the fourth is isolated.

Four parameters. Three are the previous model numbers. One is new. PARAMETER WHAT IT CONTROLS VALUE STATUS LEVEL PARAMETER a process of its own the overall size of the randomness 0.20 carried over unchanged VOL OF VOL a fixed constant how random the level parameter is 0.30 carried over unchanged CORRELATION a fixed constant how the two move together minus 0.7 carried over unchanged EXPONENT a fixed constant how randomness scales with level 1 NO COUNTERPART Hold three fixed and the fourth is the only thing being examined.
Three of the four parameters take the previous model values of 0.20, 0.30 and minus 0.7 unchanged, which isolates the exponent as the single addition and makes any difference in behaviour attributable to it alone.
Try it out

How many of the four parameters are the previous model numbers, carried over deliberately?

What does the level parameter do once it is allowed to move?

The level parameter is not a constant. Its own equation has no drift and a volatility of 0.30. Those two facts make it the simplest kind of wandering positive quantity: it cannot reach zero, it has no level it is pulled toward, and its average across all the futures the model contemplates stays exactly where it started.

Work that through with the locked numbers and one useful gap appears. The equation carries no drift, so over the one year horizon the average of the level parameter is 0.200000, exactly its starting value. Its middle value is not 0.200000. The middle value is 0.191199. The average of a quantity that wanders multiplicatively sits above its middle, and the gap here is the factor 0.955997. The average and the middle of the level parameter are two different numbers, 0.200000 against 0.191199, and the difference is produced entirely by the volatility of volatility. Its spread after one year comes to 0.061376.

The standard process itself shows the same shape of gap over the same horizon, where the average finish is Rs 108.33/- and the middle finish is Rs 106.18/-. Same mechanism, different quantity. A reader who has met that gap once already does not need it explained again. Notice instead that the gap is now happening to the volatility rather than to the price. Giving volatility its own randomness means exactly that.

Quantity, one year onFigureWhere it comes from
Level parameter at the outset0.200000written down, matching the locked volatility
Its average one year on0.200000its equation carries no drift
Its middle value one year on0.191199the starting value times 0.955997
Its spread one year on0.061376computed from a volatility of volatility of 0.30
Gap between average and middle0.008801produced entirely by the volatility of volatility
Try it out

The level parameter has no drift in its equation. What does that make its average one year on?

What does the exponent decide?

Here is the parameter that has no counterpart in the previous model, and the reason it deserves its own block. The randomness in the process at any moment is the level parameter multiplied by the process raised to the exponent. Everything about how that randomness responds to the level of the process is inside that one power.

The randomness at a level
$$ D(S) \;=\; \alpha\,S^{\,\beta} $$
\(D(S)\)the instantaneous randomness of the process when it sits at level \(S\), in rupees a year
\(\alpha\)the level parameter at that moment, 0.20 at the outset
\(S\)the level of the process, in rupees
\(\beta\)the exponent, a fixed number between nought and one
What it says in wordsThe size of the randomness at any level is the level parameter multiplied by that level raised to the exponent, so the exponent alone decides whether moving to a higher level makes the randomness bigger, and by how much.

Putting the two ends in shows what comes out. At an exponent of 1 the randomness is the level parameter times the level, so a process sitting at twice the level carries twice the randomness. An exponent of 1 is exactly the proportional treatment the standard process assumes everywhere in this subject area, and it is why a 20 per cent volatility means the same thing at Rs 50/- as at Rs 200/-. At an exponent of nought the level raised to nought is one, so the randomness is the level parameter and nothing else, the same size wherever the process happens to be. An exponent of nought is the absolute treatment. The Vasicek and Cox, Ingersoll and Ross rate models covered earlier assume it when they add a fixed amount of randomness regardless of where the rate sits.

The two ends of the exponent
$$ \beta=1:\;\; dS_t=\alpha_t\,S_t\,d\tilde W_t \qquad\qquad \beta=0:\;\; dS_t=\alpha_t\,d\tilde W_t $$
\(\beta=1\)the proportional treatment, randomness scaling with the level
\(\beta=0\)the absolute treatment, randomness of the same size at every level
What it says in wordsSetting the exponent to one gives the proportional form the standard process uses, and setting it to nought gives the absolute form the rate models use, so the two treatments set against each other earlier are the two endpoints of a single parameter.

Between them the exponent is neither, and it is not a compromise in any vague sense. The exponent is a precise statement about elasticity: multiply the level by any factor and the randomness is multiplied by that factor raised to the exponent.

The exponent as an elasticity
$$ \frac{d\ln D(S)}{d\ln S} \;=\; \beta \qquad\Longrightarrow\qquad D(cS) \;=\; c^{\,\beta}\,D(S) $$
\(c\)any factor the level is multiplied by
\(\beta\)the exponent, read here as an elasticity
\(D(S)\)the randomness at level \(S\), as defined above
What it says in wordsThe exponent is the percentage change in the randomness for a one per cent change in the level, so multiplying the level by four multiplies the randomness by four at an exponent of one, by two at an exponent of a half, and by one, meaning not at all, at an exponent of nought.

One parameter runs the whole distance between the two treatments set against each other earlier, and that is why it is worth a block of its own. The proportional and the absolute were presented as different worlds with different equations. Here they are two settings of one dial, and every value in between is a real model that is neither.

Bars at Rs 50/-, Rs 100/- and Rs 200/-, at three settings of the exponent. Each panel is scaled to its own tallest bar, so the shape is what changes. EXPONENT NOUGHT 0.200000 0.200000 0.200000 Rs 50/- Rs 100/- Rs 200/- same at every level EXPONENT A HALF 1.414214 2.000000 2.828427 Rs 50/- Rs 100/- Rs 200/- grows with the square root EXPONENT ONE 10.000000 20.000000 40.000000 Rs 50/- Rs 100/- Rs 200/- proportional to the level 0 0.5 1 the absolute treatment the proportional treatment One dial. Two worlds at its ends. Everything in between is a real model.
At an exponent of nought the randomness is the same at every level, at one it is proportional to the level, and at a half it grows with the square root, so a single parameter runs continuously between the absolute treatment and the proportional treatment.
Try it out

What does an exponent of 1 correspond to?

Breaking Into Quants Bootcamp — Fin Maverick

What do the four parameters look like when they are written out on the case numbers?

Here is the whole setting in one place, on the standard process, every value written down by hand.

ParameterValueWhat it doesWhere the number came from
Level parameter0.20overall size of the randomnessmatches the locked 20 per cent volatility
Exponent1how randomness scales with the levelno counterpart in the previous model
Volatility of volatility0.30how random the level parameter isthe previous model number, unchanged
Correlationminus 0.7how the two move togetherthe previous model number, unchanged

Now read the randomness those settings produce, at three levels of the process. At an exponent of 1 and a level parameter of 0.20, a process sitting at Rs 50/- carries a randomness of 10.000000 rupees a year, at Rs 100/- it carries 20.000000, and at Rs 200/- it carries 40.000000. Four times the level, four times the randomness. Set the exponent to nought instead. The level no longer enters, and the three become 0.200000, 0.200000 and 0.200000, all identical. Set it to a half and they become 1.414214, 2.000000 and 2.828427, where four times the level has produced exactly twice the randomness.

The nine figures are the entire content of the exponent, and every one of them is the level parameter multiplied by a level raised to a power. Nothing was sampled and nothing was fitted. Move the exponent in the control below and watch them recompute.

Try it out

At an exponent of 0.5, how does the randomness at Rs 200/- compare with the randomness at Rs 50/-?

Try it out

The exponent is about to be set to nought. Before the control moves: what happens to the randomness at the three different levels?

Play with it

Move the exponent from nought to one

Held fixed: the level parameter at 0.20, the volatility of volatility at 0.30 and the correlation at minus 0.7. Only the exponent moves. The top panel is the randomness in rupees a year at three levels of the process. The bottom panel is the same randomness read as a percentage of the level it sits at, and it is the mirror image. The track at the foot shows where between the absolute treatment and the proportional treatment the current setting sits. All three redraw together.

exponent 0, absoluteexponent 1.00exponent 1, proportional
RANDOMNESS IN RUPEES A YEAR each panel scaled to its own tallest bar 10.000000 20.000000 40.000000 process at Rs 50/- process at Rs 100/- process at Rs 200/- THE SAME RANDOMNESS AS A PERCENTAGE OF THE LEVEL 20.000000 20.000000 20.000000 at Rs 50/-, per cent a year at Rs 100/-, per cent a year at Rs 200/-, per cent a year absolute halfway proportional
At Rs 50/-
10.000000
At Rs 100/-
20.000000
At Rs 200/-
40.000000

At an exponent of 1.00 the randomness is 10.000000, 20.000000 and 40.000000 rupees a year at Rs 50/-, Rs 100/- and Rs 200/-. Four times the level gives four times the randomness, which is the proportional treatment the standard process assumes.

Educational illustration. Every figure is computed from the level parameter multiplied by the level raised to the exponent, on every move of the control. The computation is exact rather than sampled, so the default reproduces the worked instance above exactly. The three settings named in the text read as follows. At an exponent of 1: 10.000000, 20.000000 and 40.000000. At an exponent of 0.5: 1.414214, 2.000000 and 2.828427. At an exponent of nought: 0.200000, 0.200000 and 0.200000. The level parameter is held at 0.20 throughout, the volatility of volatility at 0.30 and the correlation at minus 0.7. The standard process and its parameters were written down for teaching and describe no market.

Does holding the level parameter at 0.20 hold anything fixed?

Holding the numeral at 0.20 holds nothing fixed, and the trap sits inside the control above. The units of the level parameter change with the exponent, and the same numeral therefore does not mean the same thing at every setting. Reading the equation dimensionally makes the answer fall out.

The units of the level parameter
$$ [\alpha] \;=\; \mathrm{Rs}^{\,1-\beta}\;\mathrm{year}^{-1/2} $$
\([\alpha]\)the units the level parameter is measured in
\(\beta\)the exponent, which appears in the units themselves
What it says in wordsThe units of the level parameter depend on the exponent, so a level parameter of 0.20 at an exponent of one is a pure proportion, meaning 20 per cent a year, while a level parameter of 0.20 at an exponent of nought is an amount of money, meaning twenty paise a year, and the two are not comparable quantities at all.

Work that through. At an exponent of 1 and a level parameter of 0.20, the randomness at Rs 100/- is 20.000000 rupees a year. Read as a percentage of the level, that is 20.000000 per cent, the locked volatility exactly. At an exponent of nought and the same 0.20, the randomness at Rs 100/- is 0.200000 rupees a year, or 0.200000 per cent of the level: one hundredth as much. The simulation above holds the numeral fixed and lets the meaning change. Holding it that way is what makes the shape visible, and a reader must not mistake it for a like for like comparison.

So what would a like for like comparison look like? Ask instead which level parameter each exponent needs in order to produce the same randomness at Rs 100/- as the locked volatility does, namely 20 rupees a year. Solve it and the answers are exact: 20.000000 at an exponent of nought, 2.000000 at a half, and 0.200000 at one. Feed those back in and the three settings agree exactly at Rs 100/- and fan out differently away from it. The fanning is the honest picture of the work the exponent performs.

Rescale each exponent to agree at Rs 100/-. One common scale. Watch them fan. PROCESS AT Rs 50/- PROCESS AT Rs 100/- PROCESS AT Rs 200/- 40 rupees a year 20.000000 14.142136 10.000000 all three agree: 20.000000 20.000000 28.284271 40.000000 exponent nought, level parameter rescaled to 20.000000 exponent a half, level parameter rescaled to 2.000000 exponent one, level parameter 0.200000, the locked setting
Rescaling the level parameter to 20.000000, 2.000000 and 0.200000 makes the three exponent settings agree exactly at Rs 100/- and fan apart at Rs 50/- and Rs 200/-, which is the like for like reading of what the exponent decides.
Try it out

Rescaled to agree at Rs 100/-, which exponent setting gives the largest randomness at Rs 50/-, below where the three meet?

Derivatives Foundation Bootcamp — Fin Maverick

What does being organised around the pattern buy, and what does it cost?

The model buys a great deal, and an account that listed only costs would mislead. Take the gains first. A model organised around a pattern across strikes gives a curve with a small number of handles, each of which moves the curve in a way that can be described: one sets the height, one tilts it, one bends it, and one decides how the whole thing shifts when the process moves to a different level. Four handles that behave that predictably are unusual, and they are why the model is reached for so often where a shape across strikes has to be produced quickly and consistently.

The model also buys speed of a specific kind. Because the model was organised to deliver a volatility for a strike, it does so with an expression rather than with an integral or a lattice or a grid of simulated futures. A model that has to be solved before it answers is a different tool from one that answers directly, and the difference is not academic when the same question has to be answered at many strikes at once.

Now the bill. Two items, and both come straight from the direction of travel.

The first is that the model has little to say about the process. Because the previous model was built out of a description of how variance behaves over time, it earned statements about exactly that. The SABR model has a process written down, but the process was chosen for the shape it delivers. The path the quantity takes between two dates is a matter of mechanism, and the mechanism here is scaffolding for a shape rather than a claim about anything.

The second is that the good behaviour is local. InterpolationFilling in between the strikes a model was built on, which is where a pattern organised model does its most defensible work. between the strikes the model was organised around is real work and it is the work the model is best at. ExtrapolationGoing beyond the strikes a model was built on, which is where a shape is being extended with nothing underneath it. beyond them is a different activity wearing the same clothes. Outside the region it was built on the model does not change, and neither does its confidence. The unchanged confidence is precisely the problem.

One question at the top. The answer is decided by what kind of question it is. What is the question actually about? a shape across strikes Reach for the pattern organised model. It was built to deliver exactly this and it delivers it with four handles. what the process does over time Reach for the process organised model. A good fit to a pattern does not qualify a model to answer this one. The choice is made by the question, never by which model fitted better.
A question about the shape of prices across strikes suits the pattern organised model, while a question about the path of the process over time suits the process organised one, and fit quality never settles that choice.
Try it out

Someone needs to know the path of the process between two observation dates. Which model does that call for?

Risk Management Program Bootcamp — Fin Maverick

Where does the model break down?

The breakdown has a location, and naming the location is more useful than naming a failure mode. The location is wherever the strikes stop.

Inside the range of strikes a pattern organised model has been given, the model is doing something defensible. The model is connecting readings. The four parameters are constrained on both sides, the shape between two marks is pinned at both ends, and if the shape is wrong in the middle it is wrong by a small amount. Inside the range is the region the model is good at, and being good at it is not nothing.

Outside that range there is no reading on one side. The shape is being carried outward on the strength of its own form. Nothing pins it. The same parameters that pinned the middle are now doing all the work at a distance. A small error in the tilt or the bend inside the range becomes a large error far outside it. The model gives no signal that this has happened. It returns a number in the same format with the same apparent authority, and no measure of fit computed inside the range says anything at all about it.

There is a second and quieter breakdown worth naming. The expression that turns the four parameters into a volatility for a strike is an approximation, and approximations have regions where they are good and regions where they are not. Approximations of this kind get worse for strikes far from the current level and for long horizons, and the reason is structural rather than accidental: the further an expansion is asked to reach, the less an expansion holds. The expression itself, and the size of its error, are covered separately.

Where the strikes stop, the work changes. The model does not. EXTRAPOLATION INTERPOLATION EXTRAPOLATION no mark on this side pinned at both ends by the marks no mark on this side shape carried outward errors stay small shape carried outward lowest mark highest mark strike, increasing to the right The dashed line is the same curve carried outward. It looks identical in every way that a reader can check. No measure of fit computed inside the marks says anything about outside them.
Between the strikes the model was built on it is connecting readings that pin it at both ends, while beyond them it is carrying a shape outward with nothing to pin it, and the output looks identical in both regions.
Bond Pricing and Yield Mechanics — free micro-course from Fin Maverick

What does someone actually do with a model organised this way?

Strip out the vocabulary and the practical use is a familiar one. Someone has a handful of readings and needs a value at a point where no reading exists. A pattern organised model turns those readings into a small set of numbers, and those numbers then produce a value anywhere the reader asks. The four parameters are the compressed form of the readings, and compressing is the useful act: five numbers become four handles that move in describable ways, and the four handles can be compared with the four handles from a different set of readings or a different horizon in a way that five raw numbers cannot.

Compression is the honest use, and it has an honest discipline attached. Before taking a value out of a model like this, the reader asks one question: is the point being asked about inside the readings or outside them? Inside, the answer carries most of the authority of the readings themselves. Outside, it carries the authority of a shape. The same discipline applies to the household version. Given a measurement of the electricity used in a two bedroom home and in a four bedroom home, a rule connecting them will do reasonable work for a three bedroom home. Asked about a hospital, it will still return a number, in the same units, with the same confidence, and the number will be worthless.

The whole practice reduces to knowing where the readings stopped, and that is information the model itself does not carry. It has to be kept alongside, deliberately, by whoever reads the output.

The failure: reading a close fit as evidence

The model reproduces the pattern it was organised around, closely. What has the close reproduction established?

Nothing beyond the fact that it did what it was built to do. The failure does not feel like an error at all, and that is why it catches careful readers. A model that matches its target looks like a model that has been validated, and every instinct trained on ordinary testing says a close fit is evidence.

A close fit is not evidence here, for a structural reason. A pattern organised model was constructed so that a small set of parameters could reproduce that shape. Reproducing it is the design specification being met. The match is the same as a curve drawn through five marks passing through those five marks: true, checkable, and empty of information about anything else.

The damage is not the wrong conclusion; it is the extension that follows from it. A reader who takes a close fit as confirmation will start using the model for questions it was never organised for, and the first such questions are always about the region between and beyond the strikes it was built on. Between and beyond those strikes the performance of the model is unmeasured, and no fit statistic computed inside the range will ever raise a flag about it. The failure is silent by construction.

A report card with nothing to report on the half that matters. INSIDE THE MARKS: THE FIT REPORT MARK MISS VERDICT mark one very small good mark two very small good mark three very small good mark four very small good mark five very small good Every row passes. Every row was a target. OUTSIDE THE MARKS no rows nothing was measured here The report is silent, not reassuring. A model meeting its own specification is not evidence about anything else.
Every row of a fit report is a strike the model was organised around, so a clean report card confirms the design specification was met and carries no information at all about the region where nothing was measured.
Try it out

The model reproduces the strikes it was built on closely. What has that established?

Universal

Where does this hold, and where would rules come in?

The mathematics in this guide is universal. Whether a model is organised around a process or around a pattern, and what an exponent does to the scaling of randomness, are not matters of jurisdiction. Conduct duties do apply to what anyone does with a model in any particular place, and those are settled elsewhere.

Fitting the model to observations is taken up separately and later, despite the word fitting sitting in the name this model is known by, and a model and a fitted model are different objects. The Heston model is set out separately and comes before this. The expression that turns the four parameters into a volatility for a strike, and the size of its approximation error, are covered separately. What any contract pays is settled in a different subject area, arrives here already known, and is used only as the function whose curvature the mathematics is about.
The model is organised around the shape it must reproduce. See what that buys.

References

SourceDocumentWhere
arXiv, Quantitative FinancePreprints on stochastic volatility models and their asymptotic expansionsarxiv.org
Social Science Research NetworkWorking papers on smile modelling and parameter interpretationssrn.com
Heston, 1993A closed-form solution for options with stochastic volatility, the process organised model set against this oneReview of Financial Studies
Black and Scholes, 1973; Merton, 1973The founding papers behind the constant volatility treatmentJournal of Political Economy; Bell Journal of Economics
Vasicek, 1977; Cox, Ingersoll and Ross, 1985The short rate models whose randomness does not scale with the levelJournal of Financial Economics; Econometrica

The standard process and its four locked parameters are invented.
Educational material. Not advice on any investment, tax, budget or market position.

← PreviousNext →
Fin Maverick Micro CoursesExplore Micro Courses
Fin Maverick BootcampsExplore Bootcamps
Fin Maverick

Finance education that ends in a job, not a certificate that gathers dust. Built for young India.

LEARN
CalculatorsFrameworksComparisonsCareersShowdown
RESOURCES
All CoursesMicro CoursesBootcampsInternships
COMPANY
AboutJob openingPartnership
LEGAL
Privacy PolicyTerms & ConditionsContent LicenseReturn & Refund Policy
© 2026 FIN MAVERICK / BUILT FOR INDIA.DO FINANCE, DO NOT JUST READ ABOUT IT.