# FRI 中期报告：顶尖 AI 专家大幅低估领域进展速度

- 来源：The Decoder：AI News（RSS）
- 作者：Matthias Bastian
- 发布时间：2026-09-25 03:18
- AIHOT 分数：54
- AIHOT 链接：https://aihot.news/items/cmufxejj6040xrogvla55twb8
- 原文链接：https://the-decoder.com/top-ai-experts-badly-underestimated-how-fast-the-field-is-moving-study-finds

## AI 摘要

Forecasting Research Institute 的中期报告显示，专家和超级预测员系统性低估了 AI 进展。AI 在 2025 年 7 月达到 IMO 金牌水平，比专家预测中位数提前五年；病毒学基准、网络安全基准和 Anthropic 约 1000 亿美元年化收入等指标也远超预测，如专家对 2026 年底 AI 公司最高 ARR 的中位数预测仅 200 亿美元。

## 正文

Nano Banana Pro prompted by THE DECODER

How fast is AI improving? That question usually goes to experts at top universities, heavily cited AI researchers, and seasoned economists.

Yet these same specialists significantly underestimated recent progress on benchmarks and some adoption metrics, according to an interim report from the Forecasting Research Institute (FRI).

Since mid-2022, FRI has collected forecasts on AI progress across several studies and projects. Its samples include senior specialists. The first round of LEAP (Longitudinal Expert AI Panel) drew 339 experts, including 76 computer scientists, 76 industry experts, 68 economists, and 119 AI policy specialists. The computer scientists included 30 professors at top-20 institutions and 10 of the 200 most-cited AI authors. The panels also included superforecasters, generalists with a proven record of accurate predictions.

AI hit major milestones years ahead of forecasts

The widest gap involves math. AI reached gold-medal level at the International Mathematical Olympiad in July 2025, five years before the median expert forecast and ten years before the median superforecaster forecast. Those predictions were gathered in 2022, before ChatGPT launched, but the pattern held afterward too, according to FRI.

Based on their forecasts, experts assigned an average probability of 24.6 percent to the benchmark results that actually happened, while superforecasters assigned just 9.7 percent. For gold-medal performance at the Math Olympiad, the figures dropped to 8.6 and 2.3 percent. | Image: Forecasting Research Institute

AI may also have solved a Millennium Prize Problem, though it's still unclear whether the solution meets the evaluation criteria. In a survey from August and September 2025, experts had put the median odds of such a solution by the end of 2027 at just 10 percent, and superforecasters at 5.4 percent.

In a study of AI capabilities in virology, experts predicted AI models wouldn't match a top team of virologists on a troubleshooting benchmark until 2030. Superforecasters said 2034. FRI says that likely happened as early as April 2025. A cybersecurity benchmark showed similar underestimates.

At the median, experts expected AI to match a top team on the Virology Capabilities Test by 2030, and superforecasters by 2034. FRI says it likely happened in April 2025. Respondents tied this milestone to higher expected biorisk, not to any documented rise in actual harm. | Image: Forecasting Research Institute

Economic forecasts were also far too conservative. Experts put the median for the highest annual recurring revenue (ARR) of any AI company at the end of 2026 at $20 billion. Economists said $16 billion, and superforecasters said $25 billion. FRI cites roughly $100 billion for Anthropic in September 2026 as a figure that has likely already been reached.

Annualized revenue at Anthropic and OpenAI (left) far exceeded the median forecasts of every surveyed group (right). Even the highest group estimate of $25 billion fell well short of the $65 billion reported in July 2026. | Image: Forecasting Research Institute

Real-world impact is harder to call

Not every forecast ran too low. Biosecurity experts predicted that 22.5 percent of participants using a language model would complete biological lab tasks. Virologists expected 40 percent, superforecasters 16.2 percent. In a controlled trial, only 5.2 percent succeeded with a language model and internet access, compared with 6.6 percent using the internet alone. The language model made no measurable difference, though the trial was small.

Experts may also have overshot on self-driving cars. Their median forecast for the share of autonomous US ride-hailing trips in 2027 was 7.3 percent, while an LLM projection puts it at 2.5 percent. FRI says forecasts on economic growth, employment, and major AI harms can't be reliably judged yet.

At the same time, respondents are revising their expectations upward. Among those who completed both surveys, the average probability assigned to AI becoming a "technology of the century" rose from 31 to 36 percent for experts and from 28 to 35 percent for superforecasters over nine months.

Experts and superforecasters now expect bigger societal effects from AI than they did nine months earlier. Both groups assign their highest average probability to "technology of the century," on par with electricity. | Image: Forecasting Research Institute

FRI is adding faster methods to keep pace with AI

Going forward, FRI will highlight a subsample of respondents who expect very rapid AI progress through 2040 and publish continuously updated LLM forecasts alongside the human ones. According to ForecastBench, some models already match superforecasters on certain question types. FRI also wants to find the most accurate LEAP panelists and feature their forecasts once enough data is in.

RI does flag a catch in its own data: underestimates become obvious as soon as reality overtakes a prediction, but overestimates only become clear once a deadline passes. That makes the interim report naturally tilted toward finding cases where forecasters were too cautious. Some of FRI's own assessments also rely on LLM projections that use information the original forecasters didn't have.
