马斯克披露算力部署:Colossus 2 计划年底前配 110 万张英伟达 GB300,目标约 6 个月达到领先地位
马斯克 9 月 25 日在 X 平台披露旗下公司算力部署:Colossus 1 含 15 万张 H100、5 万张 H200 和 3 万张 GB200,Colossus 2 含 11 万张 GB200 和 44 万张 GB300。
马斯克 9 月 25 日在 X 平台披露旗下公司算力部署:Colossus 1 含 15 万张 H100、5 万张 H200 和 3 万张 GB200,Colossus 2 含 11 万张 GB200 和 44 万张 GB300。
Arena 宣布 Grok 4.7 (xHigh) 进入 Agent Arena 榜单第16位,净改进分 +3.96%。
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
Elon Musk 宣布 Grok 已可用在 Tesla 中,可通过 Connectors 免手动完成收件箱管理、日历清理以及对话已有文件/聊天/任务等实际工作。
.@Grok in your Tesla can now do meaningful work for you With Connectors, you can manage your inbox, clean up your calendar, or talk through existing files/chat/tasks – all hands-free
SpaceXAI 发布新旗舰模型 Grok 4.7,面向编码、智能体任务和知识工作,定价与 Grok 4.6 相同,为每百万 token 输入 $2、输出 $6。
Grok 4.7 is behind only Anthropic models on AA-Briefcase, ranking just behind Opus 5 at ~50% of its Cost per Task Grok 4.7’s improvements over Grok 4.6 are clear in AA-Briefcase-Lite, our public due diligence scenario where models are tasked with building market models and target assessment decks. Grok 4.7 gains significantly in Analytical Quality Elo (1698 → 1994) with a slight regression in Presentation Elo (1531 → 1499). API cost to produce example decks: Grok 4.7 (xhigh) ~$8 vs. Grok 4.6 (xhigh) ~$4.40
xAI 发布 Grok 4.7,称其为编码与知识工作最强模型,采用更大的新基础模型并经过更长强化学习训练,定价与 Grok 4.6 相同,每百万输入 token $2、输出 token $6。
Grok 4.7 发布,官方称在同等价格和速度下较 Grok 4.6 有显著提升。Harvey Legal Agent Benchmark 上 Grok 4.7 得分 19.6%。
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
Grok 4.7 by @SpaceXAI and @elonmusk is now in the Agent Arena! Your votes drive the @arena leaderboards, head over and bring your toughest prompts. In Agent Arena, we measure models on millions of real-world, long-horizon agentic tasks from a global community of users. Models can access web search, filesystem, and terminal tools to complete complex workflows. The leaderboard measures model performance on outcomes relative to the average model using a causal tracing methodology. In addition to Agent Arena, Grok 4.7 is in Battle Mode for: Text, Vision, Code, and Document.
Elon Musk 引用 Artificial Analysis 评测称,Grok 4.7 使 xAI 在智能体编码上排名第三,仅次于 Anthropic 和 OpenAI。
Grok 4.7 scores 46 on the Artificial Analysis Intelligence Index to bring SpaceXAI into the top 4 AI labs. Coding Agent Index performance has also improved, overtaking GPT-5.6 Sol Grok 4.7 scores +2 points over Grok 4.6 on the Intelligence Index, with strong performance on agentic knowledge work tasks. We evaluated the new model at xhigh reasoning effort. Congratulations to @SpaceXAI and @ElonMusk on the release! Key takeaways: ➤ Grok 4.7 joins the frontier of agentic knowledge work: Grok 4.7 gains +111 Elo over Grok 4.6 (high) on AA-Briefcase, our private benchmark for long-horizon agentic knowledge work, scoring 1657 Elo and placing it alongside Claude Opus 5 and Claude Fable 5.1 at the frontier. On GDPval-AA, it scores 1695 Elo, +90 ahead of Grok 4.6 (high). ➤ A leap in coding agent performance: Grok 4.7 (xhigh) with Grok Build scores 56 on the Artificial Analysis Coding Agent Index, up +9 points from Grok 4.6 (xhigh). Among models in their native harnesses, Grok 4.7 + Grok Build now ranks 4th, behind only Claude Fable 5.1, GPT-6 Astra, and Claude Opus 5. ➤ Incremental performance changes elsewhere: Outside of agentic knowledge work, Grok 4.7 broadly matches Grok 4.6 (high) on the other Intelligence Index tasks. It improves on Terminal-Bench 4.0 (+4.5 percentage points) and GDP.pdf (+3.0 p.p.), with regressions on AA-LCR (-3.7 p.p.) and AutomationBench-AA (-1.1 p.p.). ➤ High token use across tasks: Grok 4.7's gains come with higher token usage. Grok 4.7 (xhigh) uses approximately 81k output tokens per Intelligence Index task, compared with 36k for Grok 4.6 (high) and 27k for GPT-6 Astra (max) - 125% and 196% more, respectively. Other model details: ➤ Context window of 500k tokens, unchanged from Grok 4.6 ➤ Pricing of $2/$6 per 1M input/output tokens with cache hits discounted to $0.50 per 1M tokens, matching Grok 4.6 ➤ Configurable reasoning effort spans low to xhigh. Our evaluation uses xhigh.
推荐理由:引用 Artificial Analysis 的评测数据并结合作者自身判断,给出了 Grok 4.7 在智能体编码中的排名与速度成本权衡视角。
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
xAI 发布 Grok 4.7,称其为目前编码与知识工作能力最强的模型,价格与速度与 Grok 4.6 持平,输入 $2、输出 $6 每百万 token。新模型采用更大的 base model 并加长强化学习训练,CursorBench 4.0 得分 46.3%,HackerBench v0.3 仅放行 3.3% 危险双用途提示。
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
xAI 发布 Grok 4.7,定位为其最强编码与知识工作模型,定价 $2/百万输入 token、$6/百万输出 token,与 Grok 4.6 同价同速,另有速度和价格加倍的快速变体。
推荐理由:官方给出了完整定价、速度和多项基准对比,读者可以据此把 Grok 4.7 放进现有模型选型里横向比较。
当地时间 9 月 18 日,SpaceXAI 发布 Grok Voice Transcribe 2.0 语音转文本模型,在价格不变的前提下错误率降低约一半。该模型基于 Grok Voice 底层音频基础模型构建,在 Artificial Analysis 全部 32 款流式模型中准确率居首;官方四个内部数据集评测均超越 1.0 版本。
xAI 发布 Grok Voice Transcribe 2.0 语音转写模型,官方称在真实场景评测中准确率是 1.0 的两倍且价格不变。在 Artificial Analysis 公开榜单上,它在 32 个 streaming 模型中准确率排名第一;多语言能力提升最大,短短语集合上词错误率从 20.6% 降至 6.8%。
推荐理由:原文给出公开榜单排名、内部错误率数字和与 1.0 同价的定价,读者可据此评估迁移成本与实际收益。
xAI 宣布 Grok Build 上线记忆功能,会在每轮对话结束后于后台记录项目约定、决策及事实,供后续会话读取。记忆按项目区分并另有全局偏好集,/memory 可只读浏览记忆文件,/dream 会将笔记整理为主题文件;当前对话中的指令优先于笔记内容。
推荐理由:原文说明了记忆的记录、整理和优先级机制,读者可以据此评估它对长期项目编码工作流的影响。
微软将 xAI 的 Grok 模型整合进 Microsoft 365 Copilot,Grok 已陆续登陆 Word、Excel 和 PowerPoint,通过 Microsoft Frontier 计划处于有限预览阶段。
据彭博社报道,Anthropic CEO 达里奥·阿莫代伊呼吁全行业放缓最前沿模型研发,将增设独立第三方评估等安全防护措施,OpenAI CEO 奥尔特曼与马斯克表态支持。市场观察人士认为半导体及 AI 相关股票短期或遭抛售,但因芯片、能源与算力需求供不应求,长期影响大概率有限;纳斯达克 100 指数较 6 月高点已跌超 4%,美国芯片指数下跌 14%。
Anthropic CEO 阿莫迪发文呼吁 AI 行业放缓发展速度,提出独立监控、行业统一监管和全球监管机制三项计划,并表示将向第三方评估机构提供员工级别的永久访问权限。
@farzyness Grok 4.7 needs a few more days to cook. We might have penalized response length too much (or something) in RL, as it still gives up on hard tasks (that it can do!) too early and isn’t yet sufficiently rigorous in checking its work.
Elon Musk 转发 Grok Bot 对 SpaceX CFO Bret Johnsen 在 Goldman Sachs Communacopia 演讲的摘要。
Grok Bot Summary of SpaceX CFO Bret Johnsen at Goldman Sachs Communacopia today. Vertical integration Vertical integration is the company’s core operating model, not a side strategy. - Rockets: own metal → engines → avionics → software - Starlink: own launch, satellites, and the end customer - AI: build facilities and power themselves, run their own models, sell to consumer and enterprise, and soon orbital compute Starship and launch Starship is the foundation for every other business. - Flight 13: big learning flight. Delivered demo V3 payloads, relit a Raptor, and got a soft, precise second-stage splashdown. Recovery team towed the stage back so engineers could study the heat shield. - Those learnings feed straight into Flight 14 and beyond. - Flight 14 (later this month): first revenue-generating Starship flight, flying production V3 Starlink satellites. - Later this year: aim to recover both first and second stages. Orbital compute Most of the AI industry agrees orbital compute is the future. Almost everyone else thinks it’s many years away. SpaceX disagrees because they control the stack. - Target: first orbital compute satellites next year - Scale: big compute in space into 2028 - Hardware approach: same V3 bus as Starlink, swap the payload, add larger solar arrays Why orbital can beat terrestrial on cost The crossover is about Starship reusability. - Falcon 9: first-stage reuse since Dec 2015; 500+ booster reflights - Starship: first stage already recovered/reflown; second-stage recovery progressing - Goal: reflight of both stages as soon as next year, which drops deployment cost sharply Terrestrial compute is getting more expensive (power, cooling, buildings, real estate). Orbital rides the opposite curve: cheaper rockets + better/cheaper satellites + scale. Johnsen said cost parity could come as soon as next year. Terrestrial compute and the $100B ARR goal - End of this year: on track for ~$100B ARR (annualizing the December number) - New update: another hosting deal closed earlier this month → about $1.1B/month starting Dec 1 → roughly +$13B ARR - Capacity: end this year well over 2 GW; next year 5–10 GW deployed - Confidence comes from line of sight to power, facilities, and permitting, plus being NVIDIA-exclusive for allocation - They stand compute up fast for themselves and for industry partners, which strengthens the NVIDIA relationship How they monetize compute Most hosting deals are short: ~90 days with a 90-day out (~6-month commits), including the newest deal. Why keep them short? - High conviction in their own products (Grok, Grok Bot, Cursor team after closing that deal) - Don’t want to lock forever capacity they may need internally - Internal bar: don’t let internal monetization fall below external hosting Earnings framing for next year: roughly $30–$50 per watt monetization range; they said they’re at the high end. Hosting customers appear to monetize even higher, which is why demand stays strong. Payback is under one year on new compute capex, so residual GPU value and financing options look attractive. “Not all CapEx is the same” — GPUs with <1-year payback are different from a launch tower built for decades. AI products and M&A Historically SpaceX was almost all organic growth. This year they did M&A because the AI product cycle rewards speed to frontier. - Closed Cursor deal weeks ago; product cycles already accelerating (called out Grok Bot) - Grok 4.6 improved on 4.5; 4.7 coming soon - Pitch: best infrastructure + competitive model + lower token cost = best position for customers - Market mood shift: months ago people bought the infra story but doubted the products; ~90 days later that skepticism is fading Starlink broadband Started as “better than nothing” (~2020–21). Now enterprise-grade with strong uptime/SLAs. - Resiliency pitch: boards will ask why Starlink wasn’t in the network if you go down - Mobility: aircraft backlog is large and production is ramping; cruise ships, yachts, trains too - Awareness, especially outside the US, is still a growth unlock - Longer-term: physical AI (robots, cars, aircraft) will need always-on connectivity terrestrial networks can’t fully cover Mobile / direct-to-cell Not a distraction. Same V3 bus, different payload. - Fly direct-to-device satellites through next year - Target service turn-on: first half of 2028 - V1 today (e.g. T-Mobile / T-SAT): text / light voice, great for emergencies and dead zones - Next gen: full 5G-quality from space - US: mid-band spectrum from EchoStar, FCC path for space + terrestrial - Go-to-market: flexible — own terrestrial build, or partner with carriers - International: same regulator-by-regulator playbook as broadband (Starlink now in 170+ countries) Near-term priorities: 1. Starship (enables everything else) 2. Terrestrial compute (funds growth and teaches them how to do orbital) Bottom line in one line Own the full stack, make Starship reusable at scale, use terrestrial AI compute as a cash engine now, and use the same satellite bus + Starship cadence to win broadband, mobile, and orbital AI.
推荐理由:原文以 Grok Bot 摘要形式整理 SpaceX CFO 演讲要点,覆盖 Starship 复用、地面算力营收和轨道计算时间表等关键信息。
马斯克引用 Polymarket 消息称,萨尔瓦多在 171 所公立学校的 AI 辅导试点中,阅读、数学和科学成绩高于全国平均水平,与德国和瑞典相当,并配文 Grok。
JUST IN: El Salvador’s AI tutoring pilot in 171 public schools produced reading, math, & science results above the national average & comparable to Germany & Sweden.
Grok Imagine Video 1.5 agent is now available. Powered by our newest Image 2.0 model, it delivers higher quality, better storytelling from a smarter agent and excels at connecting multiple shots together with greater continuity.
DogeDesigner 发推分享一张原始人钻木取火的图片,配文“没有 Grok Bot 的生活”,以玩梗方式调侃没有 Grok Bot 的状态。
Grok Bot 现已支持 iPad,Elon Musk 转发确认这一上线消息。
Grok Bot is now available for iPad
You can now add Bot templates from our marketplace. We’re sharing Haggle Bot, our in-house procurement specialist. It negotiates vendor contracts, finds unused SaaS seats, and price-checks recurring purchases. One week in, it's saved us over $100K.
Grok bot 现已支持开启自动更新,作者称过去 7 天发布了 49 个新版本,该功能可帮助用户保持最新。需将客户端升级到 0.40 版本才能启用此设置。
Grok Bot for Enterprise is available today. It’s free for all Grok and Cursor enterprise customers for the next two weeks. https://x.ai/news/grok-bot-for-enterprise