Everyone who calls elections, scored
Models, markets, forecasters and pollsters on one board, graded on resolved US races from 2016 to 2024 and ranked by skill adjusted for how hard each race was. Calling thirty safe seats correctly is not the same achievement as calling thirty toss-ups, and an unadjusted score cannot tell them apart.
| # | Model / market / forecaster / pollster | Races | Win Brier | 90% range (wider = less data) |
|---|---|---|---|---|
| 1 | FiveThirtyEightinactive | 683 | 0.072 | |
| 2 | VotePredictor Elections+ensemblemodel + forecasters + pollsters | 804 | 0.074 | |
| 3 | VotePredictor Elections+model + forecaster consensus | 804 | 0.074 | |
| 4 | The New York Times/Siena College | 145 | 0.074 | |
| 5 | JHK Forecasts | 391 | 0.078 | |
| 6 | OnMessage Inc. | 29 | 0.067 | |
| 7 | Cook Political | 615 | 0.079 | |
| 8 | Race to the WH | 276 | 0.079 | |
| 9 | Fox News | 174 | 0.079 | |
| 10 | The Economist | 501 | 0.081 | |
| 11 | Sabato's Crystal Ball | 615 | 0.082 | |
| 12 | Inside Elections | 390 | 0.083 | |
| 13 | VotePredictor Electionsour model | 804 | 0.084 | |
| 14 | Split Ticket | 196 | 0.084 | |
| 15 | Polymarket | 43 | 0.080 | |
| 16 | DDHQ/Decision Desk | 531 | 0.090 | |
| 17 | Research Co. | 79 | 0.085 | |
| 18 | Suffolk University | 65 | 0.084 | |
| 19 | Princeton Election Consortium | 131 | 0.088 | |
| 20 | Beacon Research/Shaw & Co. Research | 41 | 0.083 | |
| 21 | Cygnal | 54 | 0.086 | |
| 22 | CNalysis | 276 | 0.092 | |
| 23 | Elections Daily | 254 | 0.093 | |
| 24 | Silver Bulletin | 40 | 0.085 | |
| 25 | Emerson College | 230 | 0.094 | |
| 26 | SSRS | 47 | 0.087 | |
| 27 | SurveyUSA | 94 | 0.092 | |
| 28 | University of Massachusetts Lowell | 25 | 0.085 | |
| 29 | Morning Consult | 50 | 0.091 | |
| 30 | Rasmussen Reports | 68 | 0.097 | |
| 31 | SurveyMonkey | 156 | 0.100 | |
| 32 | YouGov | 156 | 0.100 | |
| 33 | New York Times | 51 | 0.097 | |
| 34 | Ipsos | 83 | 0.100 | |
| 35 | AtlasIntel | 35 | 0.096 | |
| 36 | Mason-Dixon Polling & Strategy | 35 | 0.096 | |
| 37 | Redfield & Wilton Strategies | 33 | 0.098 | |
| 38 | Civiqs | 53 | 0.104 | |
| 39 | Remington Research Group | 35 | 0.103 | |
| 40 | InsiderAdvantage | 43 | 0.105 | |
| 41 | Siena College | 34 | 0.104 | |
| 42 | RMG Research | 29 | 0.105 | |
| 43 | University of New Hampshire | 46 | 0.109 | |
| 44 | Targoz Market Research | 26 | 0.106 | |
| 45 | Swayable | 43 | 0.110 | |
| 46 | Global Strategy Group | 33 | 0.110 | |
| 47 | Public Policy Polling | 99 | 0.119 | |
| 48 | Data for Progress | 69 | 0.119 | |
| 49 | Garin-Hart-Yang Research Group | 38 | 0.117 | |
| 50 | RealClearPolitics | 214 | 0.125 | |
| 51 | Marist College | 78 | 0.124 | |
| 52 | Change Research | 97 | 0.125 | |
| 53 | Quinnipiac University | 59 | 0.126 | |
| 54 | Monmouth University | 77 | 0.129 | |
| 55 | Gravis Marketing/Kaplan Strategies | 80 | 0.131 | |
| 56 | Trafalgar Group | 86 | 0.133 | |
| 57 | GQR | 27 | 0.133 | |
| 58 | Impact Research | 32 | 0.139 | |
| 59 | GBAO | 32 | 0.148 | |
| 60 | Susquehanna Polling & Research Inc. | 27 | 0.150 | |
| 61 | Lucid | 26 | 0.163 | |
| Prov. | Harris Insights & Analytics | 19 | 0.068 | |
| Prov. | Public Opinion Strategies | 23 | 0.078 | |
| Prov. | The Washington Post | 20 | 0.085 | |
| Prov. | Kalshi | 22 | 0.087 | |
| Prov. | Wick | 19 | 0.101 | |
| Prov. | Patriot Polling | 20 | 0.119 | |
| Prov. | Clarity Campaign Labs | 20 | 0.121 | |
| Prov. | OH Predictive Insights | 20 | 0.127 | |
| Prov. | Selzer | 22 | 0.141 | |
| Prov. | co/efficient | 20 | 0.151 | |
| Prov. | DCCC | 24 | 0.198 |
The one metric they all share is win-call Brier (a pollster's win probability is derived from its margin; forecasters publish one directly). A Bayesian two-way model adjusts for race difficulty, the bar is each entity's 90% credible interval (wide = few races), and the board is ranked by a conservative estimate so thin coverage can't top it. Pick an office; lower (left) is better. Sharp forecaster probabilities lead here — switch to pollster ratings for the margin-accuracy view.
Polymarketis included as a market model using our timestamp-matched, resolved race archive. The displayed sample is scoreable overlap, not Polymarket's lifetime coverage. Kalshi is evaluated the same way using its archived pre-election candlesticks. Their uncertainty bars reflect limited matched coverage; unresolved 2026 prices are never treated as scoring data.
VotePredictor+ is a meta-forecaster: it stacks our own model with the published expert consensus, leaning on the human shops where they've historically been sharper (the House) and on the model where it has (Senate, President), with the blend weight fit on prior cycles only. It is ranked in place like everything else, but read it knowing it ensembles the very forecasters it sits beside — an advantage none of them have. The like-for-like entry, and the one to judge the model itself by, is plain VotePredictor Elections.
VotePredictor+ensemble goes one further and folds in the pollster consensus too — but it trails VotePredictor+, and that's the interesting part: the poll signal is already inside the model, which de-biases and recency-weights the same polls far better than a raw pollster average. Adding that average back is redundant and adds no lift — evidence that the gains live in the forecasters' judgment, not in re-counting the polls.
The same field, unadjusted
Raw Brier over 615 resolved races from 2016 to 2024 — no difficulty model, nothing but what each forecaster said and what happened. It is here because it is checkable by hand. VotePredictor places 4th of the 4 forecasters that rated every one of them. Forecasters that rated only a subset are listed below the ranked rows, unranked, next to VotePredictor's Brier on exactly the races they chose — the subsets differ enormously, which is what the adjusted board above exists to correct.
What a Brier score is
The squared distance between what you said and what happened. Call a race at 90% and win it, you score 0.01. Call it at 90% and lose, you score 0.81. Confidence is rewarded only when it is earned, which is why lower is better and why hedging everything at 50% is not a way to win.
Why this board is not the ranking
Only 4 entries cover all 615 races, and only they carry a rank. Everyone else scored a subset they chose, and the safe races are the easy ones, so a thin sample flatters a score: Princeton Election Consortium posts the lowest Brier here on 131 races and sits 12th once difficulty is accounted for. The “VP, same” column is VotePredictor scored on exactly that forecaster's races — the only fair pairwise read a subset allows.
Two ways to be right, two rankings. Win-call skill grades the probability of picking the winner; margin accuracy grades how close the predicted point spread landed. Each view ranks by its namesake — win-call skill, or average miss — with a skill score alongside (how much you beat a naive prior-result baseline on the same races, so polling only easy races earns nothing). Split by competitiveness (under 10 points = the hard calls). A pollster's win probability is derived from its margin so it can join the win board; forecasters publish only a probability — a number that can't reconstruct a margin — so they sit out the margin board (switch to Win-call skill to see them). Click a column to sort, a name to drill in.
How close the predicted Dem−Rep margin landed (skill vs a prior-result baseline; lower avg miss is better). Forecasters publish only a win probability, so they don't appear here.All races, 2016–2024.
| # | Name | Type | Races | Margin skill | Avg miss |
|---|---|---|---|---|---|
| 1 | AtlasIntel | Pollster | 35 | 51% | 1.7 pts |
| 2 | InsiderAdvantage | Pollster | 43 | 43% | 3.6 pts |
| 3 | Beacon Research/Shaw & Co. Research | Pollster | 41 | 58% | 3.8 pts |
| 4 | Rasmussen Reports | Pollster | 68 | 40% | 3.8 pts |
| 5 | University of Massachusetts Lowell | Pollster | 25 | 60% | 3.8 pts |
| 6 | OnMessage Inc. | Pollster | 29 | 17% | 3.9 pts |
| 7 | Trafalgar Group | Pollster | 86 | 49% | 4.1 pts |
| 8 | Research Co. | Pollster | 79 | 45% | 4.1 pts |
| 9 | Suffolk University | Pollster | 65 | 60% | 4.2 pts |
| 10 | Emerson College | Pollster | 230 | 53% | 4.2 pts |
| 11 | The New York Times/Siena College | Pollster | 145 | 55% | 4.2 pts |
| 12 | Susquehanna Polling & Research Inc. | Pollster | 27 | 50% | 4.2 pts |
| 13 | Cygnal | Pollster | 54 | 53% | 4.3 pts |
| 14 | Civiqs | Pollster | 53 | 40% | 4.4 pts |
| 15 | SurveyUSA | Pollster | 94 | 39% | 4.4 pts |
| 16 | Morning Consult | Pollster | 50 | 27% | 4.4 pts |
| 17 | Marist College | Pollster | 78 | 42% | 4.5 pts |
| 18 | SSRS | Pollster | 47 | 32% | 4.5 pts |
| 19 | VotePredictor Elections | Model | 804 | 49% | 4.9 pts |
| 20 | Data for Progress | Pollster | 69 | 34% | 4.9 pts |
| 21 | YouGov | Pollster | 156 | 32% | 5.1 pts |
| 22 | Public Policy Polling | Pollster | 99 | 39% | 5.2 pts |
| 23 | Mason-Dixon Polling & Strategy | Pollster | 35 | 51% | 5.2 pts |
| 24 | Ipsos | Pollster | 83 | 33% | 5.3 pts |
| 25 | VotePredictor Elections House segmented | Model | 332 | 43% | 5.3 pts |
| 26 | Global Strategy Group | Pollster | 33 | 43% | 5.4 pts |
| 27 | Quinnipiac University | Pollster | 59 | 29% | 5.4 pts |
| 28 | Gravis Marketing/Kaplan Strategies | Pollster | 80 | 36% | 5.4 pts |
| 29 | Siena College | Pollster | 34 | 38% | 5.5 pts |
| 30 | GBAO | Pollster | 32 | 37% | 5.5 pts |
| 31 | Monmouth University | Pollster | 77 | 35% | 5.6 pts |
| 32 | Poll average (naive) | Baseline | 804 | 40% | 5.7 pts |
| 33 | Change Research | Pollster | 97 | 42% | 5.8 pts |
| 34 | Swayable | Pollster | 43 | -19% | 5.9 pts |
| 35 | Redfield & Wilton Strategies | Pollster | 33 | -18% | 6.0 pts |
| 36 | Targoz Market Research | Pollster | 26 | 32% | 6.1 pts |
| 37 | RMG Research | Pollster | 29 | -2% | 6.2 pts |
| 38 | Lucid | Pollster | 26 | 22% | 6.3 pts |
| 39 | Remington Research Group | Pollster | 35 | 23% | 6.3 pts |
| 40 | SurveyMonkey | Pollster | 156 | 13% | 6.4 pts |
| 41 | University of New Hampshire | Pollster | 46 | 14% | 6.4 pts |
| 42 | VotePredictor Elections House fundamentals | Model | 321 | 30% | 6.5 pts |
| 43 | Garin-Hart-Yang Research Group | Pollster | 38 | 13% | 6.6 pts |
| 44 | GQR | Pollster | 27 | 18% | 6.8 pts |
| 45 | Impact Research | Pollster | 32 | 37% | 6.8 pts |