D2VolleyballStats · Matchups
Methodology tested on every match

How we rank D2 volleyball

Every ranking is a claim about who's better. We test ours the only fair way: replay the season one day at a time, predict each match using only what happened before it, and keep score. Then we run every other kind of system through the same test, on the same matches.

79.3%
of 9,254 matches called before they were played · 2024–2026
+3.5
points better than the NCAA's NPI at picking winners
Every
head-to-head won against 6 other systems, every season, every gap statistically significant
78.9%
2026 so far · 2,085 matches · retested daily

The short version

How we test

The results

All 3 seasons together (2024, 2025, 2026): 9,254 matches

SystemLog lossBetter than a coin flipPicked the winnerCross-conferenceSeasons we beat it (p < .01)
D2VB Power (ours)
Points, conference hierarchy, recency, last season
0.430
38%
79.3%0.460 · 78%—
Pablo-style
Points (capped), recency
0.459
34%
77.9%0.516 · 75%3 of 3
Our old Power
Points, shrink to the D2 average
0.459
34%
78.2%0.524 · 75%3 of 3
Massey (1997 least squares)
Point margin
0.482
31%
78.1%0.572 · 76%3 of 3
Sets only
Sets won and lost
0.482
30%
76.9%0.548 · 74%3 of 3
Wins only
Wins and losses (the RPI/NPI/Colley family)
0.500
28%
75.1%0.551 · 71%3 of 3
Elo with margin (538-style)
Wins, scaled by margin, game by game
0.506
27%
75.0%0.554 · 72%3 of 3
RankingMatchesIts higher-ranked team wonOurs, same matchesWhere they disagreedSignificance
RPI (classic 25/50/25)9,25473.5%78.4%ours right 955, theirs right 504p < .001
NCAA Power Index (our calculation)9,22174.9%78.5%ours right 846, theirs right 523p < .001
AVCA coaches poll1,33284.5%85.3%ours right 54, theirs right 43p = .310 (not yet significant)

Averages are weighted by matches. The ranking test pools every match where the two rankings disagreed. The coaches' poll only ranks 25 teams and covers far fewer matches, so it takes several seasons of disagreements to separate it from ours; this test grows every week.

Season by season

Every system, predicting this season

SystemUsesLog lossBetter than a coin flipPicked the winnerCross-conferencevs. ours
D2VB Power (ours)Points, conference hierarchy, recency, last season0.436
37%
78.9%0.464 · 79%—
Our old PowerPoints, shrink to the D2 average0.486
30%
77.2%0.552 · 75%ours better, p < .001
Pablo-stylePoints (capped), recency0.489
30%
76.5%0.542 · 75%ours better, p < .001
Sets onlySets won and lost0.517
25%
75.8%0.576 · 73%ours better, p < .001
Massey (1997 least squares)Point margin0.523
24%
77.1%0.617 · 75%ours better, p < .001
Wins onlyWins and losses (the RPI/NPI/Colley family)0.530
24%
73.1%0.571 · 70%ours better, p < .001
Elo with margin (538-style)Wins, scaled by margin, game by game0.535
23%
73.3%0.563 · 72%ours better, p < .001

2,085 matches between D2 teams, August 20, 2026 to October 7, 2026, each predicted from only the matches before it. Lower log loss is better; a coin flip scores 0.693. Cross-conference and NCAA tournament columns: log loss · % picked. "vs. ours": a paired test of the log losses on the same matches.

Rankings without win chances

RPI, the NCAA's NPI and the coaches' poll only rank teams, so the test is simpler: did the better-ranked team win? Ours is scored on exactly the same matches.

RankingMatchesIts higher-ranked team wonOurs, same matchesWhere they disagreedSignificance
RPI (classic 25/50/25)
Every match between D2 teams, ranked from results before that day.
2,08571.6%78.7%ours right 249, theirs right 101p < .001
NCAA Power Index (our calculation)
Every match between D2 teams, ranked from results before that day.
2,07872.8%78.7%ours right 229, theirs right 106p < .001
AVCA coaches poll
Matches with at least one ranked team, using the poll out before the match (unranked counts below every ranked team).
27085.2%87.4%ours right 17, theirs right 11p = .345 (not yet significant)

What each ingredient adds

Starting from our old model and adding one piece at a time.

ModelLog lossPickedCross-conference log loss
Our old Power0.48677.2%0.552
+ conference hierarchy0.46677.5%0.520
+ recency (within conference)0.46477.6%0.518
+ robust weights0.46377.6%0.516
+ last season as a starting point0.44478.5%0.481
+ calibrated win chances0.43978.5%0.469
+ serve and tempo adjustment0.43678.9%0.464

Are the win chances honest?

When we call a team a 75% favorite, it should win about 75% of the time. The top bar is what we said, the bottom bar what happened.

We gave the favoriteMatchesAverage chanceFavorite won
50–60%31755.1%53.0%
60–70%29365.0%62.8%
70–80%30975.0%76.1%
80–90%43985.1%83.4%
90–100%72796.2%95.2%

Month by month

MonthMatchesOursOur old PowerWins only
August2400.514 · 77%0.575 · 75%0.594 · 70%
September1,5130.424 · 80%0.480 · 78%0.528 · 73%
October3320.435 · 76%0.447 · 76%0.496 · 76%

This season's results update every day as matches are played (last run Oct 8, 4:56 AM CT).

The systems

SystemWhat it usesStrength of scheduleBlowoutsHow we tested it
D2VB Power (ours)Rallies won, venue, dateSolved jointly, with conference strength measured from cross-conference playSurprising results count a little lessDirectly
RPIWins and lossesOpponents' win % (50%) and their opponents' (25%)IgnoredDirectly (classic 25/50/25)
NCAA Power Index (NPI)Wins, losses, venueAverage opponent NPI (75%), bonus for good winsIgnoredOur calculation of the published formula (used for projections)
AVCA coaches pollCoaches' votesVoters' judgmentVoters' judgmentDirectly, from each week's poll
Pablo (RichKern.com)Point share, venue, dateSolved jointly (least squares)Capped at about 59% of pointsA replica built from its published method; Pablo's site doesn't allow automated access, so we can't test the live ratings
MasseyPoint marginSolved jointly (least squares)Counted in full (1997 method)Directly (his original method)
Elo (FiveThirtyEight style)Wins, scaled by marginGame by gameDiminishing (log of margin)Directly
Sets onlySets won and lostSolved jointlyA 3–0 is a 3–0Directly
Wins only (Colley/RPI information)Wins and lossesSolved jointlyIgnoredDirectly
KenPom / T-RankPoints per possession (basketball)Solved jointlyExpected mismatches count lessTheir volleyball equivalent, a separate sideout and serving rating per team, predicted worse than one rating (below); their mismatch weighting is in ours

How the Power rating works

Résumé

For each team: how likely would the 25th-best team be to win at least as many matches against the same opponents, at the same sites? The less likely, the more impressive the record, and the higher the Résumé rank. Only wins and losses count; the power ratings judge how hard each match was. It's the same idea as ESPN's Strength of Record and Bart Torvik's Wins Above Bubble, which the NCAA's basketball committee now uses.

What we tried that didn't help

Keeping it honest

Sources: Pablo FAQ · Massey (1997) · NCAA D2 NPI overview · KenPom methodology · LRMC (Kvam & Sokol) · Ranking and schedule connectivity (Osting et al.) · RPI