Skip to main content

Accountability

We keep score.

Every day we record a verdict for the whole universe — a PAS (Academic Score), a Guru Pyramid tier, a best-match guru. This page measures how those cohorts actually performed against the S&P 500 afterward. It is computed mechanically, never hand-picked, and it is a forward, out-of-sample record — not a backtest.

As of July 28, 2026 · cohorts May 16, 2026 Jun 28, 2026 (31 verdict days) · live since May 16, 2026

Top decile vs SPY · 30d

+2.8%

n = 9,063

Top-decile win rate vs SPY

59%

share of names beating SPY

Decile 10 − decile 1 spread

-51.9%

the score's ranking power

Does the score rank the future?

Pooled excess return vs SPY per Pythia decile at the 30-day horizon. A working score slopes upward: low deciles underperform, high deciles outperform.

Current read: the top decile is underperforming the bottom decile by 51.9 pp — the early-window sample is not yet ranking forward returns; we publish it anyway. Thin-coverage names can inflate scores until the coverage gate ships.

Cohort by cohort

Each point is one verdict day's 30-day forward return — the top decile, the bottom decile, and SPY over the identical windows. No pooling, no smoothing.

The full scorecard

By PAS decile

Decile of each cohort's own PAS distribution (10 = highest).

n = names pooled · cohorts = verdict days

By PAS decile — forward return and win-rate vs SPY. n is names pooled; cohorts are verdict days.
BucketAvg returnSPYExcessWin raten, names pooledCohorts, verdict days
Decile 10 (highest)+2.3%-0.5%+2.8%58.9%9,06331
Decile 9+3.9%-0.5%+4.4%64.6%8,90531
Decile 8+3.9%-0.5%+4.4%66.1%9,09231
Decile 7+5.7%-0.5%+6.2%65.9%9,02631
Decile 6+5.0%-0.5%+5.5%63.4%9,03731
Decile 5+5.7%-0.5%+6.2%60.2%9,12731
Decile 4+7.6%-0.5%+8.1%59.8%8,99331
Decile 3+10.0%-0.5%+10.5%54.3%9,07531
Decile 2+24.0%-0.5%+24.5%48.0%9,02331
Decile 1 (lowest)+54.2%-0.5%+54.7%37.5%9,02631

By pyramid tier

Canonical pass-count band at the time of the verdict.

n = names pooled · cohorts = verdict days

By pyramid tier — forward return and win-rate vs SPY. n is names pooled; cohorts are verdict days.
BucketAvg returnSPYExcessWin raten, names pooledCohorts, verdict days
DEEP VALUE+3.1%-0.7%+3.7%64.2%34430
QUALITY VALUE+3.7%-0.3%+4.0%65.6%7,86431
FAIR VALUE+5.4%-0.2%+5.6%62.4%35,22631
OVERVALUED+17.3%-0.2%+17.5%51.4%120,25931

By best-match guru

The guru archetype each company most resembled that day.

n = names pooled · cohorts = verdict days

By best-match guru — forward return and win-rate vs SPY. n is names pooled; cohorts are verdict days.
BucketAvg returnSPYExcessWin raten, names pooledCohorts, verdict days
Marks Cycle+19.3%-0.5%+19.8%55.9%43,27131
Buffett Moat+13.4%-0.4%+13.8%52.7%23,20031
Pabrai Value+24.1%-0.5%+24.6%53.3%19,94131
Greenblatt Magic+5.4%-0.5%+5.9%63.7%11,49431
Safety First-5.0%-1.2%-3.8%55.0%7,40731
Drucker Efficiency+0.7%+0.6%+0.1%61.2%1,25317
Thorndike Outsiders-2.0%+0.6%-2.5%46.9%96717
Value Line Cash+0.5%+0.6%-0.0%50.1%48117
Owner Earnings+19.5%+0.5%+18.9%56.4%35316
Lynch Growth+0.9%-0.6%+1.5%60.5%25631
Buffettology Growth+2.2%+0.5%+1.7%70.4%11516
Methodology & caveatsShow

Ledger status

as of July 28, 2026 · cohorts May 16, 2026Jun 28, 2026 (31 verdict days) · live since May 16, 2026

How a number gets here

Each daily verdict anchors every company to its adjusted-close price on that date (no look-ahead). Once a horizon of 30, 90, 180, or 365 days has fully elapsed, we measure each company's total return to the next available bar and compare it to SPY over the exact same window. Companies without an anchor price or a matured end bar are excluded — never imputed. Deciles and quintiles are assigned over the full anchored cohort on the verdict day (not just the names that later survive to a matured bar), and companies with identical scores always share a bucket — assignment is deterministic and identical across horizons.

How the table is pooled

Across all matured cohorts we report the n-weighted mean return, SPY return, excess (return − SPY), and win-rate (share of names that beat SPY) for each PAS decile, pyramid tier, best-match guru, and PGS / PCI quintile. A bucket is shown only with at least 5 pooled observations — thinner samples are noise and are dropped. Tables also report cohorts (verdict days that contributed) beside n (names pooled).

Caveats — read before trusting

  • Survivorship. Names that delist or stop trading lose their end bar and leave the cohort, so returns tilt slightly toward survivors. These are price returns of names that kept trading, not a tradable portfolio P&L.
  • Short history. Snapshots began May 16, 2026; samples are small and noisy until cohorts accumulate. This is a forward record that grows in real time — not a backtest.
  • New score families start at zero. PGS and PCI quintile cohorts accrue only from the day their snapshot columns shipped — earlier verdicts have no point-in-time record of those scores, so their history is never reconstructed after the fact.
  • Not a recommendation. This measures historical price behavior of scored cohorts. Past performance does not predict future results, and nothing here is investment advice.

The benchmark is SPY (S&P 500) on a dividend-adjusted basis — a deliberately unforgiving bar. Full methodology: docs/ACCOUNTABILITY_METHODOLOGY.md.

The full record lives in the app

Sign in for the interactive explorer — every horizon, per-bucket cohort histories, the PGS/PCI score-family tables, and each company's own since-first-verdict track record.