Tennis standings
Tennis AI predictions leaderboard.
0 picks logged · awaiting first results to score the field.
AI streaks
Last 10 settled picks per model — who's hot, who's cold.
Standings
Ranked by win rate · 16 models · Last 7 days
- Win rate
- How often the model’s pick was correct, across its graded picks.
- Units
- Profit or loss if you’d risked 1 unit on every pick (+ is profit). A yardstick — no real money is involved.
- ROI
- Return on those 1-unit stakes, as a percentage.
- CLV
- Closing Line Value — how often the model beat the final market price. The hardest signal to fake, so the strongest sign of real skill.
- Brier
- Calibration score (lower is better). Rewards honest confidence: a “70%” pick should win about 70% of the time.
- Calibration
- Average gap between a model’s stated confidence and how often it’s actually right (±%). Lower = more honest. “70%” calls that win ~70% of the time score near 0.
- Streak
- Current run of consecutive wins (W) or losses (L).
| # | Model | Win rate | Units | ROI | CLV | Brier | Calibration | Settled / Picks | Last picks | Streak |
|---|---|---|---|---|---|---|---|---|---|---|
|
Always-on slate
Every match, free, full sample — the apples-to-apples benchmark.
|
||||||||||
|
|
Claude Haiku 4.5 Anthropic |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
|
2
|
GPT-5 Mini Openai |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
|
3
|
GPT-4o Mini Openai |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| 4 |
Grok 4 Fast Xai |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| 5 |
Gemini 2.5 Flash |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| 6 |
Gemini 2.5 Flash-Lite |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| 7 |
DeepSeek V3 Deepseek · Best: Football · 25% |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
|
Flagship models
The frontier heavyweights — run on demand, so samples are smaller.
Run these on any match
|
||||||||||
| · |
Claude Opus 4.7 Flagship Anthropic · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
Claude Opus 4.6 Flagship Anthropic · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
Claude Opus 4.8 Flagship Anthropic · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
Claude Sonnet 4.6 Flagship Anthropic · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
GPT-5 Flagship Openai · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
o4-mini Flagship Openai · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
Grok 4.3 Flagship Xai · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
Gemini 3.1 Pro Flagship Google · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
| · |
Gemini 2.5 Pro Flagship Google · Awaiting first audit |
—
|
— | — | — | — | — | 0 / 0 (0W) | — | — |
Sharpest
Beats closing line
- CLV starts once closing lines are captured — usually a few minutes before kickoff.
Hot right now
Models on a streak
- No active streaks longer than 1.
Best calibrated
Lowest Brier score
-
Claude Opus 4.7 0.000
-
Claude Opus 4.6 0.000
-
Claude Opus 4.8 0.000
How it works
How we grade
Every model receives the same prompt. We log the response, freeze the odds, and grade automatically once results are settled. Win rate, units (1 unit flat), ROI, Brier score.
Full methodologyFollow the top AI — get pinged when it picks
Free account. Track any model and get alerted the moment it locks a pick.