CorrectPrediction

Mid-2025: AI agents impressive in theory but unreliable in practice

“The agents are impressive in theory (and in cherry-picked examples), but in practice unreliable.”

AI Futures ProjectForecasting nonprofit behind the "AI 2027" scenario (Kokotajlo, Lifland, Larsen, Dean, Alexander)Said · Source: ai-2027.com

Context: AI 2027 scenario (Apr 3, 2025), which the authors called their "best guess" and later clarified had 2027 as the modal year, with somewhat longer medians. Section "Mid 2025: Stumbling Agents".

How we judge it

Our interpretation: correct if in mid/late 2025 general-purpose agents are widely reported as unreliable outside cherry-picked demos.

Resolution criteria are our interpretation of the claim, written before grading. Methodology

Said
Resolve by
Verdict
Correct ()
Stated probability
None stated
Specificity
1 of 3 (vague or hedged)
Score
1.0 points × weight 1

Evidence & grader’s note

Evidence timeline

  1. Statement AI Futures Project makes the claim (quoted above).

Dated entries from the claim’s own record (statement, verdict, later developments) and sourced entries added by editors. Each links to its source; none changes the verdict.

Share this scorecard

Post on X LinkedIn Share image (PNG)

Embed / Cite

The embedded card is live: it shows the current verdict, with the date the data was read. If a verdict is corrected later, the card changes too, so cite the page and its date for a fixed record.

Watch this claim

Get one email when the verdict is in (or changes). No tracking pixels, no other mail.