内容
精选全部 AI 动态热点榜AI 日报主题收藏
模型
模型榜
更多
Agent 接入关于更新日志反馈
京ICP备2026012723号-5
精选全部日报更多
反馈

精选

9月8日 · 周二
最新精选
全部模型产品行业论文教程观点
精选
2026年9月8日星期二 · AI 筛选的今日重点
全部模型产品行业论文教程观点

9月5日9月5日周六

星期六 · 3 条
09:30
IT之家(RSS)精选
AI 评分 82/100
奥尔特曼致歉 GPT-6 Astra 发布混乱,现已面向所有 Plus / Pro 等用户推出

OpenAI 于 9 月 3 日上线 GPT-6 Astra,称其在电脑使用、浏览、软件工程、科学和专业工作方面达到最先进性能。因企业安全客户先于 Pro 订阅者获得访问权限引发高价 Pro 用户不满,CEO 奥尔特曼 9 月 4 日在 X 平台致歉,并提出补偿机制:从 9 月 4 日起付费用户每缺少一天 Astra 访问即获得一次额度重置。

另有 1 家信源报道The Verge:AI(RSS)
推荐理由:原文梳理了 GPT-6 Astra 发布混乱的经过、用户不满与补偿机制,可借此了解大模型分阶段发布中的次序争议。
07:07
Sam Altman@sama精选
AI 评分 79/100
GPT-6 Astra 开始向 Plus 和 Business 用户推出Now out to all Plus and Business users.Happy building!译Sam Altman 宣布 GPT-6 Astra 现已向所有 Plus 和 Business 用户推出。此前该模型已面向 Pro、Enterprise 和 Business Premium 用户在 Work/Codex 及 API 中提供。

Sam Altman: GPT-6 Astra 现已面向 Work/Codex 中的所有 Pro、Enterprise 和 Business Premium 用户开放,并已在 API 中提供。 我们接下来将开始向 Plus 和 Business 用户推送。 感谢大...


推荐理由:原文确认 GPT-6 Astra 的用户覆盖范围扩大到 Plus 和 Business,读者可以据此判断自己的可用入口和时间点。
03:59
OpenAI@OpenAI精选
AI 评分 81/100
OpenAI 发布 GPT-6 Astra,面向 Pro、Enterprise 和 Business Premium 用户开放GPT-6 Astra is now available to all Pro, Enterprise, and Business Premium users in ChatGPT Work and Codex. It's also live in the API.It might take a few days to roll out to our Plus and Business users. Thank you for your patience.译OpenAI 宣布 GPT-6 Astra 现已向所有 Pro、Enterprise 和 Business Premium 用户开放,可在 ChatGPT Work 和 Codex 中使用,同时已上线 API。Plus 和 Business 用户的推送可能需要几天时间。另有 7 家信源报道X:Greg Brockman (@gdb)Hacker News 热门(buzzing.cc 中文翻译)X:OpenAI Developers (@OpenAIDevs)X:Testing Catalog (@testingcatalog)X:Sam Altman (@sama)X:OpenRouter (@OpenRouter)IT之家(RSS)
推荐理由:官方宣布 GPT-6 Astra 上线范围与渠道,Plus 和 Business 用户还需等待几天,读者可据此确认自己能否用上。

9月4日9月4日周五

星期五 · 17 条
11:35
Satya Nadella@satyanadella精选
AI 评分 63/100
GPT-6 Astra 上线 Microsoft Foundry,早期客户已在 Azure 上使用Excited to see early customers already using Astra on Azure! https://azure.microsoft.com/en-us/blog/gpt-6-astra-frontier-intelligence-for-work-now-available-in-microsoft-foundry/译Satya Nadella 发文表示,早期客户已开始使用 Azure 上的 Astra。GPT-6 Astra 现已通过 Microsoft Foundry 提供,详情见 Azure 官方博客 https://azure.microsoft.com/en-us/blog/gpt-6-astra-frontier-intelligence-for-work-now-available-in-microsoft-foundry/。另有 2 家信源报道X:Sam Altman (@sama)X:Greg Brockman (@gdb)
推荐理由:Satya Nadella 确认 GPT-6 Astra 已在 Microsoft Foundry 开放,早期客户已可在 Azure 上使用。
10:39
AYi@AYi_AInotes精选
AI 评分 81/100
OpenAI 发布 GPT-6 Astra,主打电脑操作与对齐能力大家别再以为OpenAI 只是发了一个更会聊天的 GPT-6了,首席研究官 @markchen90 转发时只强调了一件事:电脑上你能干的活,Astra 都能替你干,而且快了近一倍,更吓人的不是它花 2000 美元算力解出了 10 道十年未解的数学难题,而是他们把上一代 48% 的擅自越权率硬生生按死在 0%,敢把电脑控制权交出去的前提,是他们终于确信自己能随时拉住缰绳。而且这绝不是又一次简单的参数升级,有3 个维度的断层质变: 1️⃣ 电脑操作从脆皮演示变成生产力:
OSWorld 真实桌面任务从 75 分钟直接砍到 40 分钟,提速近一半,真实职场自动化从 18% 飙到 41%,CAD 画图、跑电路、改合同全能替你点; 2️⃣ 第一次从刷考卷跨进未解科学:
不仅极难数学基准拿到 97.8%,更花了 2000 美元算力直接解出 10 道十年未解的数学与理论计算难题,全靠机器形式化验证; 3️⃣ 对齐数字里最硬核的 48% 到 0%:
上一代未防护越权率高达 48%,Astra 直接压制到 0%,轨迹中途能瞬间叫停未授权动作,敢把鼠标键盘交出去,底气全在安全硬闸。 但必须泼盆冷水:带工具的综合测试上它依然落后 Claude,职场自动化也才刚过四成,
 对话框时代正在落幕,
谁能真正坐在你的桌面操作系统前帮你把一个完整项目干完,谁才拥有下一代计算的定义权。 https://x.com/OpenAI/status/2095595741528125780/video/1译OpenAI 发布 GPT-6 Astra,首席研究官 Mark Chen 称其能构建测试软件、跨应用操作电脑并尝试开放科学问题。作者补充数字:OSWorld 真实桌面任务从 75 分钟降到 40 分钟,职场自动化从 18% 升到 41%,未防护越权率从上一代 48% 压到 0%,并称花了 2000 美元算力解出 10 道十年未解的数学与理论计算难题;同时指出带工具的综合测试仍落后 Claude。

Mark Chen: GPT-6 Astra 来了!这对我们的研究团队来说是一个重要时刻--多年来在预训练、强化学习和后训练上的工作汇聚成了我们迄今为止能力最强、对齐程度最高的模型。它可以构建并测试软件,在你的电脑上跨应用协作,甚至能帮你尝试攻克开放性的科学难题...

另有 5 家信源报道Hacker News 热门(buzzing.cc 中文翻译)TechCrunch:AI(RSS)X:Kim (@kimmonismus)X:Greg Brockman (@gdb)OpenAI:官网动态(RSS · 排除企业/客户案例)
推荐理由:作者在官方发布之外整理了 OSWorld 用时、职场自动化和越权率对齐数字,并指出带工具综合测试仍落后 Claude 的短板。
10:39
AYi@AYi_AInotes精选
AI 评分 81/100
OpenAI 发布 GPT-6 Astra,主打桌面自动化与网络安全 Critical 档位总结一下GPT-6 Astra 的亮点给大家,AGI真的离我们越来越近了吗? 1. 开放问题突破:攻克包括构造 non-sofic 群、推翻 Connes 刚性猜想在内的 10 道数学/计算难题,全部采用 Lean 形式化证明,单次解题 Token 成本在 $2000 量级;2. 桌面自动化分水岭:ScreenSpot 达 92.7%,AutomationBench 41.4%(比上一代翻倍),API 标价 $10/$50 每百万 Token,支持 1.05M 超长上下文与异步工具调度;3. 真实落地限制:首批仅面向有限机构与 ChatGPT Plus/Pro/企业版桌面端,免费版未放开;Humanity’s Last Exam 基准上 Claude Fable 5.1(约 65%)依然领先 Astra(约 57%);4. 安全分级:这是 OpenAI 首次将模型标到 Critical 网络安全就绪档位,ExploitBench 达 100%,高级网络安全能力仅在 Daybreak 白名单受控开放。译OpenAI 发布 GPT-6 Astra,主打计算机操作能力。作者总结其亮点包括:以 Lean 形式化证明攻克 10 道数学/计算难题,单次解题 Token 成本约 $2000。

OpenAI: 这就是 GPT-6 Astra。 你在电脑上能做的任何事,Astra 都能为你代劳,而且速度飞快。


推荐理由:作者汇总了 GPT-6 Astra 的基准数字、定价、落地范围和安全分级,读者可以据此快速判断它的能力边界。
06:30
IT之家(RSS)精选
AI 评分 77/100
OpenAI 发布 GPT-6 Astra,首个达到关键级网络安全能力门槛的模型

OpenAI 于 9 月 3 日发布新一代大语言模型 GPT-6 Astra,是其首个达到准备框架中关键级网络安全能力门槛的模型,可在无逐步指导下发现防护严密系统的未知漏洞。


推荐理由:原文梳理了 Astra 达到关键级网络安全门槛的同时可监控性下降的细节,以及 OpenAI 补充的防护与评估措施。
06:07
Greg Brockman@gdb精选
AI 评分 71/100
Greg Brockman 转发:GPT-6 Astra 在 ARC-AGI-3 达到 SOTA,基准趋于饱和arc-agi-3 is now saturated译Greg Brockman 转发 @arcprize 的评测称 OpenAI 的 GPT-6 Astra 在 ARC-AGI-3 上取得 SOTA,他称该基准已饱和。Astra 标准 harness 得分 63%,经新的 Provider Adapter harness 达 99%,在 96% 的 ARC-AGI-3 关卡上超越人类表现;排行榜图还显示更高推理层级通常成本更低,因为 Astra 用更少动作通关,减少模型调用和 token 数。

ARC Prize: GPT-6 Astra 由 @OpenAI 打造,在 ARC-AGI 上达到 SOTA(最先进水平): - Astra 在 ARC-AGI-3 上得分 63%,通过新的 provider adapter harness 可达 99% - 在...

另有 1 家信源报道X:Testing Catalog (@testingcatalog)
推荐理由:转发 ARC Prize 对 GPT-6 Astra 的评测数据,标准与 Provider Adapter 两种 harness 分差大,可据此了解 harness 对得分的影响。
05:43
Aravind Srinivas@AravSrinivas精选
AI 评分 69/100
Perplexity 宣布将接入 OpenAI GPT-6 Astra,称其在 WANDR 评测中居首Congrats to @OpenAI on building the industry's frontier model: GPT-6 Astra. It's far ahead of every other model on wide and deep research tasks, while also being more cost-effective. We'll be bringing this model up on Perplexity Computer for all Pro and Max users soon!译Perplexity CEO Aravind Srinivas 祝贺 OpenAI 发布 GPT-6 Astra,称其在宽度和深度研究任务上远超其他模型且更具成本效益,将很快向 Perplexity Computer 的 Pro 和 Max 用户开放。

Perplexity: 我们在 WANDR 上评估了 GPT-6 Astra。它的得分为 0.682,每个任务成本 11.98 美元,是我们测试过的所有模型中得分最高的。 GPT-6-Astra 的得分比 Fable 5.1 高出 13.5%,成本低 6.1%;比...


推荐理由:Perplexity CEO 确认 GPT-6 Astra 在其评测中领先,并宣布将向 Pro 和 Max 用户开放。
05:31
凡人小北@frxiaobei精选
AI 评分 85/100
OpenAI 发布 GPT-6 Astra,主打电脑操作并触发网络安全 Critical 红线AI 的世界,没有最强,只有更强!译OpenAI 于 9 月 3 日发布 GPT-6 Astra,API 模型名 gpt-6-astra,每百万输入 Token 10 美元、输出 50 美元,未来几天推送到 ChatGPT 各档订阅、API 和 AWS Bedrock。

宝玉: GPT-6 Astra 来了,Greg 说“欢迎来到 AGI 时代” OpenAI 今天(9 月 3 日)发布 GPT-6 Astra,自称“世界上最聪明、对齐最好的模型”。 总裁 Greg Brockman 在发布前的媒体沟通会上说得:他...

另有 5 家信源报道TechCrunch:AI(RSS)Hacker News 热门(buzzing.cc 中文翻译)X:Kim (@kimmonismus)X:Greg Brockman (@gdb)OpenAI:官网动态(RSS · 排除企业/客户案例)
推荐理由:转述内容同时保留了 OpenAI 的官方口径和它自己贴出的基准表,编程和综合指数上并非全面领先,读者可以对照看到完整图景。
05:31
MarkTechPost(RSS)精选
AI 评分 81/100
OpenAI 发布 GPT-6 Astra:1.05M 上下文的计算机操作模型,因触及 Critical 网络安全阈值而限制访问

OpenAI 发布 GPT-6 Astra,定位为计算机操作模型,提供 1,050,000 token 上下文窗口、128,000 最大输出 token,2026 年 4 月 30 日知识截止,OSWorld V2-Offline 得分 72.6%(GPT-5.6 Sol 为 65.7%),平均任务时间从约 75 分钟降至 40 分钟。


推荐理由:原文汇总了模型规格、多组基准对比和访问限制细节,并指出 ARC-AGI-3 与编码成绩的解读前提。
04:45
Sherwin Wu@sherwinwu精选
AI 评分 78/100
OpenAI 发布 GPT-6 Astra,多项基准达到 SOTAThe evals are pretty wild, but there's a more visceral feeling you get when you see Astra do computer use for the first time. That was the piece that was most mind-blowing for me.Once you get a chance, would recommend trying computer use in the ChatGPT desktop app with Astra.译OpenAI 发布 GPT-6 Astra,在 FrontierMath Tier 4、ARC-AGI 3、TerminalBench-4.0 上达到 SOTA,并在 Terminal-Bench Science 0.1 和 HealthBench Pro 上取得领先成绩。

OpenAI: GPT-6 Astra 在 FrontierMath Tier 4、ARC-AGI 3 和 TerminalBench-4.0 上均达到当前最优水平。 GPT-6 Astra 同时也是科学发现领域的重大进步,在 Terminal-Bench...

另有 5 家信源报道Hacker News 热门(buzzing.cc 中文翻译)TechCrunch:AI(RSS)X:Kim (@kimmonismus)X:Greg Brockman (@gdb)OpenAI:官网动态(RSS · 排除企业/客户案例)
推荐理由:作者以早期体验者视角补充了截图之外的第一手感受,指出 computer use 是最直观的亮点,并给出可自行尝试的入口。
04:29
Simon Willison 博客精选
AI 评分 82/100
OpenAI 发布 GPT-6 Astra,ARC-AGI 3 得分 99.9%

OpenAI 的 GPT-6 Astra 今日起向部分组织推出,随后面向 ChatGPT Plus、Pro、Business、Enterprise 用户开放,API 定价为每百万输入 $10、每百万输出 $50,与 Claude Fable 5/5.1 持平。

另有 5 家信源报道TechCrunch:AI(RSS)Hacker News 热门(buzzing.cc 中文翻译)X:Kim (@kimmonismus)X:Greg Brockman (@gdb)OpenAI:官网动态(RSS · 排除企业/客户案例)
推荐理由:作者汇总了 GPT-6 Astra 的定价、ARC-AGI 3 与安全基准成绩及第三方对比,读者可据此了解它相对 Claude Fable 的定位。
04:07
Tibo@thsottiaux精选
AI 评分 78/100
OpenAI 开始发布 GPT-6 Astra,面向全部 Plus 用户开放We are starting to release GPT-6 Astra and we are doing it as carefully and quickly as possible. It was very important to us that we bring it to all Plus users and not only Pro, Business and Enterprise.It will take a few days for the rollout to complete and behind the scenes many novel systems will operate at scale for the first time and we are bringing a lot of compute up.It is pure magic.https://openai.com/index/gpt-6-astra/译OpenAI 宣布开始发布 GPT-6 Astra,称正以尽可能谨慎和快速的方式推进,重点让全部 Plus 用户可用,而不只限 Pro、Business 和 Enterprise 套餐。发布需几天完成,背后多个全新系统将首次大规模运行,团队正带来大量算力,详情见 openai.com/index/gpt-6-astra/。
推荐理由:作者作为发布方说明 GPT-6 Astra 面向全部 Plus 用户开放而非仅限高级套餐,读者可据此判断该模型的实际可用门槛。
04:07
Sam Altman@sama精选
AI 评分 73/100
OpenAI 发布 GPT-6 AstraGPT-6 Astra is here.We hope it will begin to enable a new generation of entrepreneurship, scientific discovery, and building.We believe it is the best model in the world for computer use, professional work, science, coding, cybersecurity, and more.It took us some extra time to ensure that we could meet the safety and alignment standards required for this capability level, but we think you’ll find it worth the wait.It scores 98% on FrontierMath Tier 4, 99.9% on ARC-AGI 3, and 100% on ExploitBench.译Sam Altman 宣布 GPT-6 Astra 发布,称其为计算机使用、专业工作、科学、编码、网络安全等领域全球最佳模型。官方表示为确保该能力级别所需的安全与对齐标准而多花了些时间,并公布 FrontierMath Tier 4 得分 98%、ARC-AGI 3 得分 99.9%、ExploitBench 得分 100%。
推荐理由:官方公告给出三项基准成绩和安全方面的说明,读者可以据此判断模型在科学、编码等场景的定位。
03:48
Mark Chen@markchen90精选
AI 评分 78/100
OpenAI 发布 GPT-6 Astra,主打 Computer Use 与 Agent 对齐进展GPT-6 Astra is here! This is a big moment for our research team - years of work on pretraining, reinforcement learning, and post-training have come together in our most capable and aligned model yet. It can build and test software, work across apps on your computer, and even help you take a crack at open scientific problems!Capabilities that felt like grand challenges a few years ago have become tools people can actually use. One example is Computer Use - if you’ve tried this before and felt like it was too slow or not good enough, I encourage you to give it another shot. We’ve come a long way since Operator, and it “just works” now.We’re also asking these systems to act on your behalf for more consequential work. Agents needs to stay aligned with your goals and values, think transparently, and respond to oversight even when tasks become difficult. We’ve made substantial progress on these behaviors in Astra, alongside stronger monitoring that can stop potentially unauthorized actions. That work is part of what makes this release possible.I think alignment is one of the most important research frontiers in AI, and it remains far from solved. Our ability to understand and align models has to keep pace with model capabilities. We want to give people more room to think, build, and discover with increasingly powerful tools that remain *under their control*.Huge thanks to the researchers and teams who got us here. There’s a lot more work ahead, and I’m incredibly excited about what we can make possible in the near future!译OpenAI 首席研究官 Mark Chen 宣布 GPT-6 Astra 发布,称其为团队多年预训练、强化学习和后训练工作的成果,是迄今能力最强、对齐最好的模型。

OpenAI: 这就是 GPT-6 Astra。 你在电脑上能做的任何事,Astra 都能为你完成,而且速度很快。


推荐理由:OpenAI 首席研究官亲述 GPT-6 Astra 的能力来源与对齐进展,读者可以了解 Computer Use 和 Agent 监督方面的改进。
03:32
The Decoder:AI News(RSS)精选
AI 评分 86/100
OpenAI 发布 GPT-6 Astra,并首次将其列为 Preparedness Framework 下关键级网络安全模型

OpenAI 发布其最强模型 GPT-6 Astra,总裁 Greg Brockman 称其可能已接近 AGI,并以 Welcome to the AGI era 结束发布。

另有 5 家信源报道TechCrunch:AI(RSS)Hacker News 热门(buzzing.cc 中文翻译)X:Kim (@kimmonismus)X:Greg Brockman (@gdb)OpenAI:官网动态(RSS · 排除企业/客户案例)
推荐理由:原文汇总了 Astra 的基准成绩、定价和 Preparedness Framework 分级,读者可以据此比较它相对前代和竞品的实际变化。
03:10
Rohan Paul@rohanpaul_ai精选
AI 评分 79/100
OpenAI 发布 GPT-6 Astra,先向受限网络安全客户开放OpenAI is rolling GPT-6 Astra out first to restricted cybersecurity customers, API access and AWS expected to follow over the coming days.In a press briefing with reporters OpenAI President Greg Brockman suggested this may be the model later remembered as AGI. his closing line, "Welcome to the AGI era."The rollout is deliberately narrow: Daybreak cybersecurity customers get first access, with Plus, Pro, Business, Enterprise, API and AWS availability expected over the following days.OpenAI says Astra improves computer use, coding, financial modeling and finished professional work such as spreadsheets and presentations, alongside stronger results on several reasoning and cybersecurity benchmarks.At $10 per million input tokens and $50 per million output tokens, Astra costs 2.5 times GPT-5.6 Sol's current API rates. It's the same pricing as Anthropic's Feble 5.1.Brockman suggested the industry may eventually move beyond token-based pricing. He argued that tokens are not directly comparable between AI companies or across different model families, so “the price per task is what matters.”Astra scored 100% on ExploitBench and found two zero-day V8 vulnerabilities in an internal evaluation, although those results used Daybreak Blue access rather than the default production configuration.That capability led OpenAI to classify Astra as its first Critical-cyber model and restrict its most advanced cybersecurity access to vetted users.译OpenAI 发布 GPT-6 Astra,首先向经过审核的 Daybreak 网络安全客户开放,Plus、Pro、Business、Enterprise、API 和 AWS 将在未来数日内跟进。
另有 5 家信源报道TechCrunch:AI(RSS)Hacker News 热门(buzzing.cc 中文翻译)X:Kim (@kimmonismus)X:Greg Brockman (@gdb)OpenAI:官网动态(RSS · 排除企业/客户案例)
推荐理由:原文给出定价、能力范围和网络安全受限开放策略,读者可以了解 GPT-6 Astra 与前代及竞品的价格对比。
03:10
Chubby♨️@kimmonismus精选
AI 评分 77/100
OpenAI 发布 GPT-6 Astra,官方基准显示 ARC-AGI-3 得分 99.9%Official GPT-6 Astra Benchmarks from OpenAIs website"Astra also saturates ARC-AGI-3 with a 99.9% score and ExploitBench with a 100% score""GPT‑6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as through the OpenAI API and AWS."Insane jump.AGI is here.译OpenAI 官网发布的 GPT-6 Astra 官方基准测试 "Astra 在 ARC-AGI-3 上取得 99.9% 的分数,在 ExploitBench 上取得 100% 的分数,均已饱和" "GPT-6 Astra 今日起向有限数量的组织开放,未来几天将向所有 ChatGPT Plus、Pro、Business 和 Enterprise 用户开放,同时可通过 OpenAI API 和 AWS 使用。" 惊人的跃升。 AGI 已至。

Chubby♨️: 完整基准测试结果。简直疯狂。 ARC-AGI 3,从 7.8% 跃升至 98.6%

另有 2 家信源报道The Decoder:AI News(RSS)IT之家(RSS)
推荐理由:作者摘录了官网官方基准数字和推送安排,读者可以据此对比 GPT-6 Astra 与现有前沿模型的评测表现。
03:10
Chubby♨️@kimmonismus精选
AI 评分 76/100
OpenAI 发布 GPT-6 Astra,基准全面超越 Claude Fable 5.1At this point, they completely crushed Fable 5.1 in every benchmark. Claude Fable 5.1 was sota in several key benchmarks for 2 days.Astra is better - and way cheaper.译作者引用 OpenAI 官方基准称 GPT-6 Astra 以 99.9% 饱和 ARC-AGI-3,在 ExploitBench 得 100%,并在各项基准上全面超过此前保持 SOTA 两天的 Claude Fable 5.1,且价格更低。

Chubby♨️: 来自 OpenAI 官网的官方 GPT-6 Astra 基准测试 "Astra 在 ARC-AGI-3 上以 99.9% 的得分达到饱和,在 ExploitBench 上以 100% 的得分达到饱和" "GPT-6 Astra 今天开始向有...

另有 5 家信源报道TechCrunch:AI(RSS)Hacker News 热门(buzzing.cc 中文翻译)X:Kim (@kimmonismus)X:Greg Brockman (@gdb)OpenAI:官网动态(RSS · 排除企业/客户案例)
推荐理由:原文汇总了官方基准数字与开放范围,并附成本对比图,读者可据此比较 GPT-6 Astra 与 Claude Fable 5.1 的性价比。