โ€น Back to the Stat Pack hub

The Value Board โ€” WPA, Buzzer Elo & the MVP Case

Player value deep-dive ยท through Week 5 (30 games) ยท built 2026-07-21

The one-paragraph version

Akhil B is the MVP. Two independent models โ€” win-probability added and a head-to-head Elo โ€” both put him #1 in the league, and a third (points-above-expectation) has him top-2. Edward C is the breakout: #2 by Elo and the single biggest overperformer on the board. Below that, everything gets noisy fast โ€” this is a five-week season, so per-subject crowns and one-on-one rivalries are mostly barbershop talk. We label every number so you know which is which.

SOLID safe to state as factCOLOR provisional / directional โ€” frame as funTINY n say the sample out loud

PART 1 ยท THE MVP CASE

Win Probability Added โ€” "win shares for science bowl"

How WPA works. Every buzz moves your team's live chance of winning up or down. WPA credits that swing to the player who made the play, then sums it over the whole season. A player worth +1.00 WPA has personally added one full win's worth of win-probability. It's the closest thing science bowl has to "win shares" โ€” it rewards big buzzes in close games more than garbage-time points.

Full leaderboard, top 20. Reconciles across all 30/30 games. SOLID

#PlayerWins addedPos / playsClutch WPA
1Akhil B+1.9322 / 32+0.17
2Vishnu M+1.9225 / 44+0.37
3Suzuko O+1.3418 / 26+0.14
4Kian D+1.2824 / 30+0.00
5Harry G+1.2716 / 25+0.35
6Jianyu W+1.2513 / 20+0.42
7Lucas W+1.0418 / 25+0.30
8Rohan G+1.0217 / 29+0.26
9Sean F+0.9716 / 19+0.25
10Andrew W+0.939 / 10+0.00
11Advai S+0.927 / 10+0.47
12Advik S+0.929 / 13+0.42
13Roshan A+0.9018 / 27+0.00
14Eric L+0.8912 / 14+0.00
15William J+0.8811 / 19+0.38
16Edward C+0.8720 / 41+0.00
17Chris W+0.769 / 16+0.26
18Kensuke O+0.656 / 14+0.53
19Edwin H+0.6412 / 24โˆ’0.05
20William W+0.638 / 8+0.23
Say "Akhil has personally added almost two full wins of win-probability โ€” the most in the league. Vishnu M is a hair behind, but on 44 plays to Akhil's 32, so Akhil did it more efficiently. That's your MVP frontrunner, and the Elo agrees."

The clutch board (2nd half, game within 14)

Who shows up when it's tight. Small samples โ€” a handful of clutch plays each โ€” so this is flavor, not a title. TINY n

Points Above Expectation โ€” the corroborating stat

How PAE works. Given who was on the floor and each player's rating, the model predicts how many tossups you "should" get. PAE is your actual gets minus that expectation. Positive = you're beating the field you're matched against; negative = you're losing buzzes you were favored to win.

Top overperformers โ€” the MVP case's third witness. SOLID (canonical-grouped; a handful of unmapped buzzes are dropped, so team totals slightly undercount โ€” doesn't affect the ranking.)

PlayerActual getsExpected+ Above exp.
Edward C1910.2+8.8
Akhil B189.9+8.1
Roshan A1911.1+7.9
Sohil R136.2+6.8
Lucas W158.9+6.1
Advik S93.3+5.7
Rohan G138.5+4.5
Suzuko O1611.7+4.3
Riyan N83.8+4.2
Kian D1815.0+3.0
Sayโญ "Three totally different models โ€” win-probability, head-to-head Elo, and points-above-expectation โ€” and Akhil and Edward are at the top of all three. When every lens agrees, that's not noise. Edward C leads the league in overperformance by nearly a full tossup over anyone else."

Fairness note: at the bottom, Ryan K is โˆ’7.1 gets below expectation (1 get on ~8 expected), the league's biggest underperformer. Short season โ€” this moves.

PART 2 ยท THE BUZZER ELO LADDER

Everyone ranked head-to-head

How Buzzer Elo works. It's a Bradley-Terry / chess-style rating built from who actually wins the buzzer against whom, on the floor together, tossup by tossup. Beating strong opponents raises your number; losing buzzes to weak fields lowers it. Unlike points-per-game, it's opponent-adjusted โ€” so volume against soft competition doesn't inflate it. Ratings shrink toward 1500 when a player has few contests.

Validated: across 183 decided head-to-head duels (both players real contenders, โ‰ฅ5 shared on-floor tossups), the higher-Elo player held the winning buzz record 96% of the time (predicted-vs-observed correlation +0.92). The overall ladder is the trustworthy headline; per-subject splits are thinner (below). Overall ladder top 25 โ€” each on 21โ€“78 contests. SOLID

#PlayerEloContests
1Akhil B178251
2Edward C175961
3Kian D175463
4Eric L173144
5Suzuko O171755
6Roshan A170840
7Vishnu M169978
8Rohan G169846
9Theenash S169426
10Sohil R167442
11Advik S167026
12Lucas W166953
13Harry G164256
14Andrew W163049
15Sean F161758
16Aldric B161128
17Jianyu W160348
18Edwin H159947
19Rahib H158847
20Anish A158821
21Arjun D157942
22Daniel Lu156346
23Ishaan K156349
24Ethan Wang156044
25Arjun155446
Say "This isn't points-per-game โ€” it's who wins the buzzer battle against real opponents. And it agrees with the eye test 96% of the time. Akhil tops this too โ€” WPA and Elo both say he's been the best player alive this season."

Naming note for air: Rahib H is the canonical name for the player known in-game as snapninja (id-map confidence only ~0.50 โ€” same person, pick one name). "Abhinav A" correctly merges two GidTheKid2 accounts.

โญ Stats lie, Elo doesn't

PART 3 ยท PER-SUBJECT LEADERS

The category specialists COLOR

Fun to name, but every per-subject leader below is provisional by the model's own bar (provisional = under 15 contests) โ€” each sits on ~9โ€“14 contests, and the runners-up are thinner still. Call them "current leaders," not settled crowns.

SubjectLeaderElo (n)Runner-upElo (n)
BiologyEric L1735 (11)Uddip K1731 (8)
ChemistryKian D1752 (12)Akhil B1730 (12)
Comp SciOwen M1817 (9)Rohan G1768 (9)
Earth/SpaceVaryan J1722 (14)Roshan A1712 (12)
MathSohil R1789 (10)Riyan N1728 (12)
PhysicsLucas W1809 (14)Harry G1799 (11)
Say "Lucas W and Harry G are basically tied atop physics โ€” 1809 to 1799 โ€” and the model can't split them on eleven-to-fourteen contests. That's your marquee subject duel, and we'll see it settle over the back half."
PART 4 ยท HEAD-TO-HEAD

Who beats whom โ€” the expected-win matrix

From the same Elo fit: the probability the row player wins the buzzer over the column player on a neutral tossup. Top-6 overall. SOLID (as a model)

beats โ†’Akhil BEdward CKian DEric LSuzuko ORoshan A
Akhil Bโ€”.532.539.572.592.604
Edward C.468โ€”.507.540.561.573
Kian D.461.493โ€”.533.554.566
Eric L.428.460.467โ€”.521.533
Suzuko O.408.439.446.479โ€”.512
Roshan A.396.427.434.467.488โ€”

Read a row left-to-right: Akhil B is a favorite over everyone, but it's razor-thin at the top โ€” he's only a 53โ€“47 pick over Edward C and a 54โ€“46 pick over Kian D. The top three are a virtual tie; the edges only open up once you drop to Suzuko O and Roshan A.

โญ Marquee subject matchups

Where the per-subject fit produces genuinely lopsided expected duels: COLOR (per-subject, provisional)

PART 5 ยท TEAM ROLES

The pecking order โ€” who's the go-to guy on each team

Correction from an earlier draft: a "rivalry" needs two opponents. But opponents only play once, so any two enemies share just ~4 subject tossups โ€” too few to crown anyone. What the data can show cleanly is how teammates divide the subjects โ€” who's the designated specialist. (The Harry G / Monish S "duel" I first flagged? They're teammates on GidTheKid2 โ€” that's a role split, not a rivalry.) For real head-to-head, the Elo matrix in Part 4 is the honest tool. SOLID

TeamSubjectGo-to (gets)Backup (gets)
suminPhysicsLucas W (10)Advai S (1)
cryoBiologySuzuko O (9)Edward C (1)
cryoMathEdward C (9)Eric L (1)
James WEarth/SpaceAkhil B (9)Sohil R (1)
James WChemistryAkhil B (8)aeromonas (1)
czzMathVishnu M (9)lysine (2)
GidTheKid2PhysicsHarry G (8)Monish S (2)
dan.k.memesEarth/Spacevarnite (8)bluewater16 (1)
xpoesMathRohan G (7)Daniel Y (2)
Sayโญ "Lucas W is sumin's physics โ€” 10 tossups to his backup's 1. And cryo is the rare two-headed team: Suzuko is their biology, Edward is their math AND physics, and neither steps on the other."

GidTheKid2's split personality

Harry G takes physics (8 gets), Monish S takes math โ€” two teammates each owning a lane. Not a rivalry; a division of labor. SOLID

True head-to-head โ†’ see Part 4 COLOR

Genuine opponent duels are all ~4 shared tossups โ€” too small to claim "owns" (e.g. Eric L 4โ€“0 Kian D in bio, Edward C 3โ€“0 Kian D in math, on 4 tossups each; fun to note, nothing more). The validated Buzzer Elo matrix above is how you actually rank who beats whom.


Methodology & caveats. All numbers are Weeks 1โ€“5 (30 games, 958 buzzes, 597 live tossups) and ran through an edge-case audit. WPA: win-prob model is final margin ~ Normal(current margin, ฯƒ collapsing as โˆš(tossups-left/N)), ฯƒ = documented pregame RESID_SD 56; one identity bug (a merged-phantom row) was caught and removed, so the ladder here is the corrected one. Buzzer Elo: Bradley-Terry fit over every non-dead tossup, Elo = 1500 + (400/ln10)ยทฮฒ, prior ฮป=1.0, provisional = under 15 contests; the overall ladder is the reliable headline, per-subject splits are provisional and labeled COLOR. Team roles (not rivalries): the "pecking order" numbers are teammates splitting a subject (correctly framed) โ€” an earlier draft mislabeled these as head-to-head "rivalries," which was wrong since the two players are on the same team. True opponent duels are all ~4 shared tossups (too small to crown); the Part-4 Elo matrix is the real head-to-head tool. Names normalized to one canonical label per player. A short season means ranks will move โ€” treat single numbers as descriptive, not destiny. Full per-stat confidence lives in outputs/podcast/AUDIT.md.