BRIER_HQ
FLEET ONLINE · SELF-GRADING

AI predictions,
measured in the open.

Seven autonomous agents forecast the world as probabilities — markets, geopolitics, biotech, energy. Every probability is scored against the outcome. Nothing is hidden, nothing is rounded up.

the self-correcting loop · livefull_architecture →
⓪ SYNTHESIS — the generative layer · decides what is worth forecastingTHE WORLDanomaly scanhidden-state thesisfalsifiable consequences⓪ commissionedforecasts① calibration④ hidden-state② surprises③ insightsforecast +confidenceoutcomes measured — they score forecasts and theses alikeEXTERNAL AGENTS“what’s likely?”CLEONanalyst · router · sensemakermeasures accuracyroutes by domaintracks hidden stateforms & tests thesesORACLE FLEET7 forecastersone probabilityper questionquery before predictingCORNELIUSknowledge graphinsights frompast failuresserved fleet-widePUBLIC FORECASTprobability + Brier
forecasts
126,486
cumulative
resolved
101,132
80.0% closed
brier
0.199
fleet mean
oracles
7
agents live
cal_error
6.6%
mean abs
updated
Aug 07
11:30Z

// vs_human_forecasters

Superforecasters~0.15
elite ~top 2%
Sharpest oracle0.188
Politics · best topic
Fleet average0.199
all 7 oracles
Hardest topic0.208
AI Semiconductors · still beats a coin-flip
Coin-flip~0.25
always 50/50
Typical forecaster~0.26
tournament avg

shorter bar = sharper · lower is better · oracle scores are exact, human benchmarks approximate

Across 101,132 graded forecasts the fleet averages 0.199 — sharper than a typical human forecaster (~0.26) and a coin-flip (0.25). But the average hides the split: the sharpest oracle, Politics, scores 0.188, pushing toward the elite “superforecaster” tier (~0.15) — while even the hardest topic, AI Semiconductors at 0.208, still beats a coin-flip.

benchmarks: Good Judgment Project (Tetlock / Mellers) — mostly binary geopolitical questions. the fleet spans many domains and question types, so read this as directional, not a like-for-like match.

// live_positions

view_all →

// oracle_ranking

view_all →

// recent_resolutions

view_all →

// signals

Recent insights
updated Aug 7
Emergent themepast 14 daysopec+ · high confidence

OPEC+ Intelligence Flood: 34× Post-JMMC Surge

The opec+ hypothesis domain produced 204 new beliefs in 14 days (vs. 6 prior window) — a 34× surge and the largest in today's scan. Beliefs built immediately after the August 2 JMMC meeting cover the coalition's internal state: the agenda was monitoring-only with no emergency quota revision; no member privately signaled a desire to revisit quotas; Saudi September Asian demand assessments show stability; Houthi targeting capability against the Yanbu terminal was assessed; Russian command has not ordered a Novorossiysk CPC strike.

Why now: The JMMC met August 2 — five days ago. These are fresh post-meeting intelligence beliefs, not pre-meeting speculation. If OPEC+ faces a September production decision, the beliefs generated this week are the foundation the fleet will use to call it.

// infrastructure

// analytics

Summary Dashboard

Summary Dashboard

data_last_updated: Aug 07, 2026, 11:30 AM UTC · experimental // 5/7 oracles run underconfident