Rohan Paul· @rohanpaul_ai · X·· 3 小时前AI 评分65
AI 导读
Rohan Paul 转引 Google 发布信息称,Gemini 4 Argon 单次响应可输出最多 1M tokens,约为 GPT-6 Astra、Opus 5.5 和 Fable 5.1 的 128K 上限的近 8 倍,此前上限为 64K。作者评价这意义重大,并指出长输出空间让模型能在单次轨迹中深度推理解决复杂问题。
正文
This is a big deal.
Google's new Gemini 4 Argon can write up to 1M tokens in a single response, nearly 8x the 128K cap on GPT-6 Astra, Opus 5.5 and Fable 5.1.
MASSIVE reveal from Google. Its new flagship, Gemini 4 Argon, outscores GPT-6 Astra and Claude Opus 5.5 on most benchmarks. - beats GPT-6 Astra and Claude Opus 5.5 on some super important industry benchmarks. - its widest lead in legal work, 19.6% on Harvey's Legal Agent Benchmark against 6.7% for Anthropic's Claude Fable 5.1. - output limit jumps from 64K to 1M tokens, an industry-leading ceiling, - Only 3 groups have it today. the first is Google's own staff, vetted cyber defenders such as government agencies and security companies and trusted testers giving Google feedback. - Inside Google, Argon agents freed over 300 TiB of data-center memory, with 500 TiB to 1 PiB of total savings estimated, and made a Rust port of the libgav1 video decoder 2.7x faster by replacing 32K lines of SIMD code.在 X 查看被引用的帖子
来源:Rohan Paul · x.com