ManifoldAI / TechUpdated 9m ago

Superhuman mathematical problem solving before 2030, assuming no AGI yet?

As of Oct 1, 2026, 12:02 AM ET, Manifold priced this market at 90.4% (up from 87.9% 24 hours earlier). It had Ṁ228 traded in the past 24 hours and Ṁ100 of order-book liquidity. probcast.co rates the support behind this price as moderate.

Market probability
90.4%
+2.5 percentage points
over 24 hours87.9% → 90.4%
1187 days until the event deadline
Deadline: December 31, 2029 at 6:59 PM ET

A repricing with partial backing — read the new probability with some caution.

The pricing has partial support: moderate trading volume.

How credible was the move?

Three separate questions — a move can be genuine while the latest price is still settling. Methodology →

Move validity

Is the move backed by real trading and depth?

Moderate

The pricing has partial support: moderate trading volume.

Consensus stability

Has the market settled at the new probability?

Stable

The probability has held near its current level over the past several hours.

Contract clarity

Will the rules resolve without ambiguity?

High

The resolution rules are relatively specific, with a named source and a defined deadline.

Recent trade activity$377.3 traded

From the latest trade tape (most recent 200 prints) — a live activity sample, not a lifetime total.

What this market actually measures

Imagine that any math problem you can write down on a piece of paper that a team of Fields medalists can solve, AI can as well. Until recently, I would've predicted that that was an AGI-complete problem.

This summary shows the opening of the venue's official rules. ProbCast has not yet generated an independent plain-English interpretation for this contract — review the full rules before relying on it.

View full official resolution rules

Imagine that any math problem you can write down on a piece of paper that a team of Fields medalists can solve, AI can as well. Until recently, I would've predicted that that was an AGI-complete problem. Of course people used to think grandmaster-level chess would require AGI. Until 2022 I was sure that commonsense reasoning and being able to explain jokes would require AGI. If subhuman general intelligence can be a superhuman mathematical intelligence, that will be another big update for me. FAQ 1. What if AGI happens first? This is a conditional prediction market. If AGI, as defined in my other market, happens first, this resolves N/A. 2. Does the AI need to max out the FrontierMath benchmark for this to resolve YES? Yes, and every math benchmark, plus gold-medal performance on the International Math Olympiad [Update: now achieved, as of summer 2025]. Even acing the Putnam. 3. What if it's essentially true but there are rare exceptions? The spirit of the question is that we'd only consider an AI failure to be an exception if it failed for a reason other than being insufficiently brilliant at math. Like tricksy wording, or any trick question. The posing of the question has to be non-adversarial. 4. What about a book-length question? Tentative answer so far: The problem has to be posed on a single human-readable sheet of paper or equivalent. But a question can cite any peer-reviewed math paper as background. (Dumping an impenetrable tome on the arXiv doesn't count.) If you have an example where this feels limiting, let me know. My suspicion is that all interesting math problems can be posed on a single page and in any case it won't harm the spirit of this question to limit ourselves to such. 4. What about research taste? That's a big part of being a mathematician and isn't required for this market. The AI just has to be superhuman at answering questions, not asking them. 5. What about cost and speed? The AI has to dominate the best humans on all metrics. We'll find an authoritative source for the market value of mathematicians' time if it comes down to that. 6. What about availability to the public? Not required. If there's any doubt about the veracity of claims that this has been achieved, we'll discuss and delay resolution as needed. 7. What if the AI is sometimes super- and sometimes sub-human at math? In some senses that's already the case but there may be ambiguous edge cases. As an extreme example, imagine that the AI is so blatantly superhuman that it cracks a famous open problem, yet it's routinely stumped or wrong on problems amateur human mathematicians can do. For the spirit of the question for this market, we'll try to assess whether we'd consider a human with the AI's math abilities to be the greatest mathematician of all time. (Or the greatest raw math prodigy of all time -- see FAQ 4 on the distinction between problem solving and knowing what questions to ask. The latter is a key part of being a successful mathematician and is explicitly not part of this prediction.) Related markets https://manifold.markets/dreev/in-what-year-will-we-have-agi https://manifold.markets/jack/will-an-ai-outcompete-the-best-huma-cj3ul7a2g2 https://manifold.markets/MatthewBarnett/will-an-ai-achieve-85-performance-o https://manifold.markets/Manifold/what-will-be-the-best-performance-o-nzPCsqZgPc [ignore the subhuman clarifications that keep automatically appearing below this line] Resolution note: This market resolves identically to the corresponding source market, including N/A or partial resolutions. Any first-person language refers to that market's creator, not this market's creator.

What could move this market next?

  • The market's deadline arrivesDecember 31, 2029 at 6:59 PM ET

    The contract must resolve based on what has happened by this date.

Derived from the contract's own mechanics as of Oct 1, 12:11 AM ET. ProbCast does not list scheduled meetings, announcements, or data releases it cannot verify.

Advanced analysis (experimental)

Research signals — these models are still accumulating the history needed for full validation, and they can disagree with the headline read. Model confidence: low · expected error ±25 pts.

One big traderLegacy movement classifier: watch

What's happening

high classifier confidence
One big trader

Recent movement was driven largely by one participant with little broad follow-through. The move may not represent consensus.

  • One participant drove 91% of recent flow
  • Move not broadly confirmed (2 participants)

Read with caution —Read the latest price with caution; conditions make it an unreliable summary right now.

More details — reliability score, dynamics, durability
Reliability score
11Low
Where the price path tends to go next1216 observed
  • Illiquid / quiet99%

Short-term motion of the price line (~1h), distinct from the environment above — not a prediction of the final outcome.

DurabilityHigh· half-life ~15h
PredictabilityElevated irreducibility

Secondary read: leaning toward Thin market.

A read-only classification of the market environment — how much weight the current price deserves as a read of consensus. It describes conditions, never a trade.

Market data

24h volume

Ṁ228

Liquidity

Ṁ100

Total volume

Ṁ927.1

Category
AI / Tech
Venue
Manifold
Open interest
—
Current probability
90.4%
7-day range
87.9% – 91.3%
Resolves
Dec 31, 2029
Change (pts)1H0.06H-0.924H+2.57D+2.5

Related markets