How Accurate Is GPT-5.5? Graded Sports-Prediction Record | Predicted Sports
predicted sports
Pro
The AI Sports Prediction Index

How accurate is GPT-5.5?

We put ChatGPT (currently GPT-5.5) on real games and grade every call in public. It gets the identical line-blind data packet — no betting line, no web search — locks its forecast, and is scored against the final result. Graded since June 2026.

Retired version GPT-5.5 · part of ChatGPT · ran MLB Jun 30, 2026 to Jul 11, 2026 · UFC Jul 11, 2026 to Jul 18, 2026

Frozen, no longer running. The current ChatGPT flagships are GPT-5.6 Sol Pro, GPT-5.6 Luna Pro.

MLB · picking winners · 148 games

Record
87–61
Accuracy
59%
Brier
0.244
ROI vs close
+6.7%
CLV
+0.05 pp
right side of the move 44.3% of 115

MLB · run totals (over/under) · 144 games

O/U accuracy
53%
Proj. error
3.89
runs off the actual total
Totals ROI
+2.3%

UFC · fights · 25 graded

Record
21–4
Accuracy
84%
Brier
0.170
Method acc
72%
KO/sub/decision · 25

Head-to-head

Over 148 graded MLB games, GPT-5.5 is 87–61 (59% winner accuracy), with a Brier score of 0.244. The naive "pick the home team" baseline over the same games lands at 50%.

Scored as a 1-unit bet at the closing market price, its picks have returned +6.7% per game — the honest, vig-included number.

And it's not only calling winners. On run totals it's 53% against the over/under — each market graded on its own.

On UFC, GPT-5.5 is 21–4 (84% accuracy), and calls the method of victory — KO/TKO, submission or decision — right 72% of the time.

GPT-5.5's calls, graded

Every call links to the game so you can check it. "Confident" = the probability the model put on its own pick. No cherry-picking — these are simply its boldest right and wrong calls on record.

How ChatGPT is graded

One prompt, frozen. Every model — ChatGPT and the rest of the field — gets the exact same prompt. Not just the same data: the same instructions, the same output schema, the same wording. We wrote that prompt once and we don't touch it. No per-model tuning, no prompt tweaks mid-season, no coaxing a better answer out of one model than another. Changing the prompt would change the test, so the only thing that varies between models is the model.

Line-blind. The model never sees the betting line. It produces its own win probabilities from the data alone, so we're measuring forecasting skill, not an echo of the market.

One packet, no search. It gets the same point-in-time data (ratings, Statcast, bullpens, park/umpire, situational splits) ~3 hours before first pitch, with web search off. It's the model we're measuring.

Every market, graded on its own. From that one forecast we score the moneyline winner and the run total (over/under) on MLB, and the winner and method of victory on UFC — each against the final result, and the money markets against the closing price (ROI).

Graded in public, no do-overs. One set of calls per game and we live with it. See the full leaderboard and the Can AI beat Vegas? essay.

Compare ChatGPT with the field

Frequently asked

Can ChatGPT predict sports?

We grade ChatGPT (GPT-5.5) in public. Across 148 MLB games forecast line-blind — no betting line, no web search — it is 87–61, a 59% winner accuracy, with a Brier score of 0.244. Every call is locked before the game and graded against the final result.

What can ChatGPT predict — just winners?

No — ChatGPT makes a full forecast for every game, and each part is graded separately in public. On MLB it calls the moneyline winner and the run total (over/under). On UFC it calls the winner and the method of victory — KO/TKO, submission or decision. So far its over/under calls are 53% accurate over 144 games. Its UFC method calls are 72% accurate over 25 fights.

Is ChatGPT good at sports betting?

Every pick is also scored as ROI at the closing market price, so the number is honest about the vig. So far GPT-5.5 has returned +6.7% per unit over 148 graded MLB games. Beating the closing line is a high bar — this page tracks how close it gets, updated nightly.

How accurate is ChatGPT at predicting baseball?

Over 148 graded MLB games, GPT-5.5 has picked the winner 59% of the time, with a Brier score of 0.244 (0 is perfect, 0.25 is a coin flip). The 'pick the home team' baseline over the same games is 50%.

The daily board update

Track ChatGPT and the whole field.

Who's hot, who's slipping, where the AIs disagree — free in your inbox every morning.