- Score
- 100%
- 95% interval
- 21%–100%
- Resolved
- 1
- Record (C–P–W)
- 1–0–0
- Open
- 0
- Too vague
- 0
Not ranked yet: leaderboards need at least 5 resolved claims (has 1). Score = specificity-weighted average of claim points; we grade claims, not people.
Calibration
No resolved claims with a stated probability yet, so there is nothing to calibrate. Calibration needs forecasts like “70% chance of X”.
Track record
Resolved (1)
CorrectPrediction
PhD-level AI for specific tasks in "a year and a half"
“And then in the next couple of years, we're looking at PhD-level intelligence for specific tasks.”
Embed / Cite
The embedded card is live and shows the current score and record, with the date the data was read.
Right of reply: send a correction or response.