The Value Board โ WPA, Buzzer Elo & the MVP Case
Player value deep-dive ยท through Week 5 (30 games) ยท built 2026-07-21
The one-paragraph version
Akhil B is the MVP. Two independent models โ win-probability added and a head-to-head Elo โ both put him #1 in the league, and a third (points-above-expectation) has him top-2. Edward C is the breakout: #2 by Elo and the single biggest overperformer on the board. Below that, everything gets noisy fast โ this is a five-week season, so per-subject crowns and one-on-one rivalries are mostly barbershop talk. We label every number so you know which is which.
SOLID safe to state as factCOLOR provisional / directional โ frame as funTINY n say the sample out loud
Win Probability Added โ "win shares for science bowl"
+1.00 WPA has personally added one full win's worth of win-probability. It's the closest thing science bowl has to "win shares" โ it rewards big buzzes in close games more than garbage-time points.
Full leaderboard, top 20. Reconciles across all 30/30 games. SOLID
| # | Player | Wins added | Pos / plays | Clutch WPA |
|---|---|---|---|---|
| 1 | Akhil B | +1.93 | 22 / 32 | +0.17 |
| 2 | Vishnu M | +1.92 | 25 / 44 | +0.37 |
| 3 | Suzuko O | +1.34 | 18 / 26 | +0.14 |
| 4 | Kian D | +1.28 | 24 / 30 | +0.00 |
| 5 | Harry G | +1.27 | 16 / 25 | +0.35 |
| 6 | Jianyu W | +1.25 | 13 / 20 | +0.42 |
| 7 | Lucas W | +1.04 | 18 / 25 | +0.30 |
| 8 | Rohan G | +1.02 | 17 / 29 | +0.26 |
| 9 | Sean F | +0.97 | 16 / 19 | +0.25 |
| 10 | Andrew W | +0.93 | 9 / 10 | +0.00 |
| 11 | Advai S | +0.92 | 7 / 10 | +0.47 |
| 12 | Advik S | +0.92 | 9 / 13 | +0.42 |
| 13 | Roshan A | +0.90 | 18 / 27 | +0.00 |
| 14 | Eric L | +0.89 | 12 / 14 | +0.00 |
| 15 | William J | +0.88 | 11 / 19 | +0.38 |
| 16 | Edward C | +0.87 | 20 / 41 | +0.00 |
| 17 | Chris W | +0.76 | 9 / 16 | +0.26 |
| 18 | Kensuke O | +0.65 | 6 / 14 | +0.53 |
| 19 | Edwin H | +0.64 | 12 / 24 | โ0.05 |
| 20 | William W | +0.63 | 8 / 8 | +0.23 |
The clutch board (2nd half, game within 14)
Who shows up when it's tight. Small samples โ a handful of clutch plays each โ so this is flavor, not a title. TINY n
- Kensuke O +0.53 ยท Advai S +0.47 ยท Advik S +0.42 ยท Jianyu W +0.42 ยท William J +0.38 lead the clutch splits.
- Cold under pressure: Daniel Y (โ0.09), Ishaan K (โ0.08), Sanjay S (โ0.06).
Points Above Expectation โ the corroborating stat
Top overperformers โ the MVP case's third witness. SOLID (canonical-grouped; a handful of unmapped buzzes are dropped, so team totals slightly undercount โ doesn't affect the ranking.)
| Player | Actual gets | Expected | + Above exp. |
|---|---|---|---|
| Edward C | 19 | 10.2 | +8.8 |
| Akhil B | 18 | 9.9 | +8.1 |
| Roshan A | 19 | 11.1 | +7.9 |
| Sohil R | 13 | 6.2 | +6.8 |
| Lucas W | 15 | 8.9 | +6.1 |
| Advik S | 9 | 3.3 | +5.7 |
| Rohan G | 13 | 8.5 | +4.5 |
| Suzuko O | 16 | 11.7 | +4.3 |
| Riyan N | 8 | 3.8 | +4.2 |
| Kian D | 18 | 15.0 | +3.0 |
Fairness note: at the bottom, Ryan K is โ7.1 gets below expectation (1 get on ~8 expected), the league's biggest underperformer. Short season โ this moves.
Everyone ranked head-to-head
Validated: across 183 decided head-to-head duels (both players real contenders, โฅ5 shared on-floor tossups), the higher-Elo player held the winning buzz record 96% of the time (predicted-vs-observed correlation +0.92). The overall ladder is the trustworthy headline; per-subject splits are thinner (below). Overall ladder top 25 โ each on 21โ78 contests. SOLID
| # | Player | Elo | Contests |
|---|---|---|---|
| 1 | Akhil B | 1782 | 51 |
| 2 | Edward C | 1759 | 61 |
| 3 | Kian D | 1754 | 63 |
| 4 | Eric L | 1731 | 44 |
| 5 | Suzuko O | 1717 | 55 |
| 6 | Roshan A | 1708 | 40 |
| 7 | Vishnu M | 1699 | 78 |
| 8 | Rohan G | 1698 | 46 |
| 9 | Theenash S | 1694 | 26 |
| 10 | Sohil R | 1674 | 42 |
| 11 | Advik S | 1670 | 26 |
| 12 | Lucas W | 1669 | 53 |
| 13 | Harry G | 1642 | 56 |
| 14 | Andrew W | 1630 | 49 |
| 15 | Sean F | 1617 | 58 |
| 16 | Aldric B | 1611 | 28 |
| 17 | Jianyu W | 1603 | 48 |
| 18 | Edwin H | 1599 | 47 |
| 19 | Rahib H | 1588 | 47 |
| 20 | Anish A | 1588 | 21 |
| 21 | Arjun D | 1579 | 42 |
| 22 | Daniel Lu | 1563 | 46 |
| 23 | Ishaan K | 1563 | 49 |
| 24 | Ethan Wang | 1560 | 44 |
| 25 | Arjun | 1554 | 46 |
Naming note for air: Rahib H is the canonical name for the player known in-game as snapninja (id-map confidence only ~0.50 โ same person, pick one name). "Abhinav A" correctly merges two GidTheKid2 accounts.
โญ Stats lie, Elo doesn't
- Patrick L is #14 in raw points-per-game but #53 in Elo (1415) โ big counting stats piled up against weak fields. The classic empty-numbers guy. SOLID
Say "Patrick L looks like a top-15 scorer on paper. Head-to-head against real competition, he's outside the top 50. Volume against weak defenses."
- Caveat in the other direction (keeping ourselves honest): the Elo slightly under-penalizes negs โ a correct buzz beats the whole opposing field, but a neg only loses to the one player who cleans it up. So a high-volume gunslinger like Ishaan K (9 gets but 9 negs in 4 games, zero net points) rates higher in Elo (#23) than his real game value โ we flag him as a draft bust elsewhere for exactly that reason. Don't over-trust a single rating. CAVEAT
The category specialists COLOR
Fun to name, but every per-subject leader below is provisional by the model's own bar (provisional = under 15 contests) โ each sits on ~9โ14 contests, and the runners-up are thinner still. Call them "current leaders," not settled crowns.
| Subject | Leader | Elo (n) | Runner-up | Elo (n) |
|---|---|---|---|---|
| Biology | Eric L | 1735 (11) | Uddip K | 1731 (8) |
| Chemistry | Kian D | 1752 (12) | Akhil B | 1730 (12) |
| Comp Sci | Owen M | 1817 (9) | Rohan G | 1768 (9) |
| Earth/Space | Varyan J | 1722 (14) | Roshan A | 1712 (12) |
| Math | Sohil R | 1789 (10) | Riyan N | 1728 (12) |
| Physics | Lucas W | 1809 (14) | Harry G | 1799 (11) |
Who beats whom โ the expected-win matrix
From the same Elo fit: the probability the row player wins the buzzer over the column player on a neutral tossup. Top-6 overall. SOLID (as a model)
| beats โ | Akhil B | Edward C | Kian D | Eric L | Suzuko O | Roshan A |
|---|---|---|---|---|---|---|
| Akhil B | โ | .532 | .539 | .572 | .592 | .604 |
| Edward C | .468 | โ | .507 | .540 | .561 | .573 |
| Kian D | .461 | .493 | โ | .533 | .554 | .566 |
| Eric L | .428 | .460 | .467 | โ | .521 | .533 |
| Suzuko O | .408 | .439 | .446 | .479 | โ | .512 |
| Roshan A | .396 | .427 | .434 | .467 | .488 | โ |
Read a row left-to-right: Akhil B is a favorite over everyone, but it's razor-thin at the top โ he's only a 53โ47 pick over Edward C and a 54โ46 pick over Kian D. The top three are a virtual tie; the edges only open up once you drop to Suzuko O and Roshan A.
โญ Marquee subject matchups
Where the per-subject fit produces genuinely lopsided expected duels: COLOR (per-subject, provisional)
- Physics โ Lucas W over Kian D, 77%. The single most lopsided marquee physics gap. Lucas W and Harry G, though, are a coin flip (51โ49).
- Chemistry โ Kian D over Roshan A, 70%, and over Eric L, 68%. Kian D is the clearest per-subject #1 on the board.
- Comp Sci โ Owen M over Edward C, 71%. Owen M's CS rating is the highest single-subject Elo in the league (1817).
- Biology โ Eric L over Kian D, 63%. Bio is the flattest subject up top: Eric L is barely ahead of Uddip K (51%) and Theenash S (52%).
The pecking order โ who's the go-to guy on each team
Correction from an earlier draft: a "rivalry" needs two opponents. But opponents only play once, so any two enemies share just ~4 subject tossups โ too few to crown anyone. What the data can show cleanly is how teammates divide the subjects โ who's the designated specialist. (The Harry G / Monish S "duel" I first flagged? They're teammates on GidTheKid2 โ that's a role split, not a rivalry.) For real head-to-head, the Elo matrix in Part 4 is the honest tool. SOLID
| Team | Subject | Go-to (gets) | Backup (gets) |
|---|---|---|---|
| sumin | Physics | Lucas W (10) | Advai S (1) |
| cryo | Biology | Suzuko O (9) | Edward C (1) |
| cryo | Math | Edward C (9) | Eric L (1) |
| James W | Earth/Space | Akhil B (9) | Sohil R (1) |
| James W | Chemistry | Akhil B (8) | aeromonas (1) |
| czz | Math | Vishnu M (9) | lysine (2) |
| GidTheKid2 | Physics | Harry G (8) | Monish S (2) |
| dan.k.memes | Earth/Space | varnite (8) | bluewater16 (1) |
| xpoes | Math | Rohan G (7) | Daniel Y (2) |
GidTheKid2's split personality
Harry G takes physics (8 gets), Monish S takes math โ two teammates each owning a lane. Not a rivalry; a division of labor. SOLID
True head-to-head โ see Part 4 COLOR
Genuine opponent duels are all ~4 shared tossups โ too small to claim "owns" (e.g. Eric L 4โ0 Kian D in bio, Edward C 3โ0 Kian D in math, on 4 tossups each; fun to note, nothing more). The validated Buzzer Elo matrix above is how you actually rank who beats whom.
Methodology & caveats. All numbers are Weeks 1โ5 (30 games, 958 buzzes, 597 live tossups) and ran through an edge-case audit. WPA: win-prob model is final margin ~ Normal(current margin, ฯ collapsing as โ(tossups-left/N)), ฯ = documented pregame RESID_SD 56; one identity bug (a merged-phantom row) was caught and removed, so the ladder here is the corrected one. Buzzer Elo: Bradley-Terry fit over every non-dead tossup, Elo = 1500 + (400/ln10)ยทฮฒ, prior ฮป=1.0, provisional = under 15 contests; the overall ladder is the reliable headline, per-subject splits are provisional and labeled COLOR. Team roles (not rivalries): the "pecking order" numbers are teammates splitting a subject (correctly framed) โ an earlier draft mislabeled these as head-to-head "rivalries," which was wrong since the two players are on the same team. True opponent duels are all ~4 shared tossups (too small to crown); the Part-4 Elo matrix is the real head-to-head tool. Names normalized to one canonical label per player. A short season means ranks will move โ treat single numbers as descriptive, not destiny. Full per-stat confidence lives in outputs/podcast/AUDIT.md.