Back to News

xAI and SpaceXAI release Grok 4.6, matching GPT-5.6 Sol on AI Index

#grok-4.6#xai#spacexai#agent

xAI and SpaceXAI released Grok 4.6, a frontier model focused on long-running agents, available in Cursor, Grok Build, Grok Bot, and API. It scores 61 on the Artificial Analysis Intelligence Index, matching GPT-5.6 Sol, and tops GDPVal-AA v2 with 1753 points, priced at $2/1M input and $6/1M output tokens.

Coverage timeline

  1. Cursor Blog

    Today we are releasing Grok 4.6 together with SpaceXAI. Grok 4.6 builds on Grok 4.5 with a particular focus on long-running agents and more ambitious interactive and visual work. It stays with complex tasks across many steps, whether researching a topic, analyzing information, working across a codebase, or turning an idea into a polished application or work artifact. Grok 4.6 achieves frontier intelligence across several agentic coding and knowledge work benchmarks. It matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index, which is a composite score of nine benchmarks. Grok 4.6 is available today in Cursor and Grok Build. We’re offering 2x included usage inside Cursor and Grok Build for the first week. ## Training Grok 4.6 Grok 4.6 underwent a longer supplemental training run than Grok 4.5, with curated model-generated data for reasoning and advanced technical concepts, high-quality engineering data, and an improved optimizer and training recipe. This produced a stronger fo

  2. Techmeme

    xAI : SpaceXAI releases Grok 4.6, saying it matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index, and prices it at $2/1M input and $6/1M output tokens — Grok 4.6 builds on Grok 4.5 with a particular focus on long-running agents and more ambitious interactive and visual work. — Try for free

  3. 量子位量子位

    马斯克Grok 4.6重回一梯队!更低价格反超Fable 5,这Cursor是真没白收购 克雷西 2026-08-13 20:04:29 来源: 量子位 马斯克版Workbuddy也来了 克雷西 发自 凹非寺 量子位 | 公众号 QbitAI 马斯克带着Grok 4.6,让SpaceXAI重新回到了大模型牌桌。 它跑分反超了GPT-5.6 Sol和Fable 5 Max,价格却只要每百万token输入2美元、输出6美元,比两个对手便宜不少。 衡量真实工作能力的GDPVal-AA v2上,Grok 4.6拿到了全场最高分1753,把它们都甩在了后面;AA-Briefcase和Harvey LAB两项测试,它同样排第一。 目前,新模型已经同步接入Grok Build、Cursor、Grok Bot和API,Cursor里首周还有双倍用量。 这次升级定的重点,是长程agent任务,要求模型在没人盯着的时候也能连续干很久不掉线。 就在前一天,SpaceXAI还发布了另一款产品Grok Bot,一队能自己登录各种工具、24小时连轴转的智能体。 现在,Grok Bot已经能直接调用Grok 4.6了。 前脚发Harness,后脚模型也跟着出炉,老马这一波,是真的不想在AI上掉队。 更低价格反超Sol和Fable SpaceXAI这次拿出的官方跑分表,摆了10项基准,对比对象是Grok 4.5 High、GPT-5.6 Sol Max和Fable 5 Max。 综合智力指数AA Intelligence Index上,Grok 4.6拿到61分,跟GPT-5.6 Sol打平,比Grok 4.5的56分高了5分,只比排名最高的Fable 5 Max少1分。 真正把两个对手都甩在后面的,是另外三项。 GDPVal-AA v2上,Grok 4.6拿到1753分,反超GPT-5.6 Sol的1728分和Fable 5 Max的1741分。 AA-Briefcase和Harvey LAB这两项,Grok 4.6同样双双超过,分别是1577分和15.8%。 但Terminal-Bench v3.0上,它只拿到26%,GPT-5.6 Sol是34.6%。 Terminal-Bench测的是纯命令行环境下的agent操作能力,模型要在没有图形界面的情况下连续执行多步指令,出一次错就可能中断整条链