Skip to content

Live vs Benchmark · The Comparison Stack

Five metrics, plus behavior.

How MARS uses this

The five-metric comparison stack ends with behavior, and tier usage is where behavior hides. Live EV and drawdown can both pass while tier dwell quietly drifts aggressive — which is why the stack includes this audit: the model's expected tier profile against what the operator actually deployed.

How it benefits you

Aggression creep gets caught while it is still a pattern on a chart rather than an oversized loss. Equally, unnecessary timidity shows up as a measurable gap, so you deploy the edge you have actually earned instead of leaving modeled expectancy on the table.

0%10%20%30%T1T2T3T4T5T6T7BENCHMARKLIVE

The behavioral row of the comparison stack: modeled tier dwell versus live deployment, T1 through T7.

The comparison stack

Five metrics, plus behavior.

Is live EV equal to or better than modeled EV? Is live drawdown equal to or better than modeled drawdown? Is live RAPF showing superior profit quality? Is live RAER converting risk as efficiently as assumed? Is live acceleration tracking the modeled growth velocity? And are gate dwell, tier usage, throttle behavior, and open exposure behaving normally?

Where it lives

Inside Live vs Benchmark.

This page expands one card of the Live vs Benchmark page into its own reference. For orientation, the module's own framing: The live comparison is deliberately not a shallow equity race. Five metric-level questions are asked against the model — EV, drawdown, RAPF, RAER, acceleration — plus behavioral checks on gate dwell, tier usage, throttle behavior, and open exposure.

Cadence

Five metrics, one review slot, no cherry-picking.

The stack is read as a set at scheduled review — EV, drawdown, tier behavior, quota adherence, throughput — because any single row can flatter. A hot EV row with degrading tier discipline is a worse report than a flat EV row with clean behavior, and only the full stack read in one sitting makes that ranking visible.

Doctrine

Band13 wk26 wk52 wkOperator read
P90+64.2R+131.8R+268.4Rhot path — do not extrapolate
P75+47.5R+96.3R+197.0Rstrong but ordinary
P50+31.1R+63.9R+130.6Rthe median story
P25+16.8R+35.4R+74.2Rslow — still inside the model
P10+4.3R+11.7R+27.9Rthe path sizing must survive

Cumulative R by horizon · 50,000 resampled paths

Further illustration

MARS reduces the fan to its actionable rows. Each percentile band carries a defined operator read - P50 is the planning story, P90 is explicitly quarantined from expectations, and P10 is the row the tier structure must survive. Reviews quote these rows directly instead of eyeballing a chart.

The percentile bands rendered as an operator table: cumulative R by horizon, and what each band is allowed to mean.

Connected inside MARS

This module doesn't work alone.

Go deeper

Operator briefs on this territory.

Every module ships in the complete MARS package.

One price. Eleven workbooks, three TradingView indicators, and the full manual library — $497.