随手先关注,不错过每日AI热讯速递
①
9 月 22 日,Anthropic 发布Claude Opus 5.5;约 90 分钟后,OpenAI 端出GPT-6 SolGPT-6 Luna[1][2]。
两家都没打「最强」。按 Anthropic 的说法,Opus 5.5 是Claude 5.5 家族第一款,多数工作达到自家最强模型Fable 5.1的水平,而运行成本比 Opus 5 低 40%[1][6]。
OpenAI 更干脆:把 6 系列价格砍到 5.6 系列的一半[2][4]。
![]()
图:Anthropic 官方 Opus 5.5 发布图,标注 40% 成本降幅、Terminal-Bench 4.0 横向对比与 API 单价(来源:MarkTechPost[7])
Opus 5.5输入4 美元、输出20 美元每百万 token,比 Opus 5低 20%;输出生成也比 Opus 5快 30% 以上[1][3][6]。
缓存读取(智能体反复读同一份代码库或长文档时,模型对已处理内容的复用,它通常占编码类任务成本的大头)从 0.50 美元降到0.20 美元,降幅60%[1][6]。
GPT-6 Sol输入2 美元、输出10 美元GPT-6 Luna更低,输入0.10 美元、输出0.50 美元[5]。
降法并不相同。Anthropic 把账拆成三份——单价降 20%、缓存读取降 60%、完成同一任务用更少 token,三者叠加才得出「典型工作负载降 40%」[1][6]。
OpenAI 则把缓存命中率的提升与推理效率改进直接折进标价:GPT-6 的缓存输入读取享受 90% 折扣[2]。
②
调价的关键落点在缓存读取上:写代码的智能体(能自己拆解目标、调用工具,并根据中间结果决定下一步的 AI 程序)要反复读代码库、跑测试,缓存命中率直接决定账单[6]。
Terminal-Bench 4.0给出 Opus 5.566.4%、Fable 5.155.8%、Opus 552.3%、GPT-6 Astra57.9%[1][7]。
GDPval-AA v2.1覆盖44 个职业的真实工作,Opus 5.5 的 Elo 得分是1846,高于 Fable 5.1 的1735与 Opus 5 的1708[7]。
一名测试者用 Opus 5.5 在不到一天内完成 68 万行代码迁移;同一桩虚构并购案分析中,它用时63 分钟,成本比 Opus 5 低 50%[1]。
Anthropic 自己也收了一句:内部使用中,Opus 5.5 与 Fable 5.1 的差距比这些分数显示得更小[1]。
![]()
图来源:Anthropic 官方发布页[1]
Sonnet 5.5 与 Haiku 5.5 将在未来几周发布,沿用同批性能与安全改进[1][7]。
③
OpenAI 这一侧,Sol面向复杂工作与编码,Luna面向高体量、目标明确的任务[2]。
在 OpenAI 公开的DeepSWE v1.1上,Sol 的最高分是68.8%,距Claude Fable 5该评测最高分69.9%只差1.1 个百分点,而每任务成本约低80%[2]。
Luna 在同一评测拿66.6%,官方称与 Opus 5、Fable 5 的中等强度水平相当,每任务成本比 Opus 5 低 93%[2]。
AutomationBench上,Sol 在 xhigh 强度下得分33.2%、每任务0.27 美元,官方称其在相同评测上胜过 Opus 5 的最高强度,每任务成本只有后者的 9%[2]。
分级也很明确:Astra 仍是处理高难度项目时能力最强的模型,Sol 与 Luna 是成本更低的选项[2][4]。
上线是分级的:Sol 与 Luna 即日起在ChatGPT WorkCodex面向 Plus、Pro、Business、Enterprise、Edu 用户提供,Free 与 Go 用户只能在桌面端用到 Luna[2]。
时间点值得留意:Opus 5.5 是 Anthropic 呼吁「放慢前沿」之后发布的第一款模型,距那次倡议只过了 10 天[1][7]。
Anthropic 的解释是,减速针对的是能自动做 AI 研究、跑在人类审核能力之前的那类系统[1]。
同一天,Perplexity 把 Opus 5.5 设为 Computer 的 Standard 档、GPT-6 Sol 设为 Light 档默认项[2]。
微软则把 Astra、Sol、Luna 一起上架 Microsoft Foundry[2]。
衡量单位正从「每百万 token 多少钱」滑向「完成一件事要多少钱」[1][2]。
欢迎在评论区聊聊:你更看重模型的单价,还是完成一件事的总成本?
参考来源:
[1] Introducing Claude Opus 5.5https://www.anthropic.com/claude-opus-5-5
[2] Introducing GPT-6 Sol and Lunahttps://openai.com/index/introducing-gpt-6-sol-and-luna
[3] Anthropic releases Opus 5.5 with lower prices and Fable-level performancehttps://techcrunch.com/2026/09/22/anthropic-releases-opus-5-5-with-lower-prices-and-fable-level-performance/
[4] OpenAI launches GPT-6 Sol and Luna, boasting lower cost and fewer mistakeshttps://techcrunch.com/2026/09/22/openai-launches-gpt-6-sol-and-luna/
[5] OpenAI's GPT-6 Sol and Luna cut prices in half but barely move the needle on performancehttps://the-decoder.com/openais-gpt-6-sol-and-luna-cut-prices-in-half-but-barely-move-the-needle-on-performance/
[6] Claude Opus 5.5 matches Fable 5.1 at 40 percent lower cost, as Anthropic promises to fix “Claudish” writinghttps://the-decoder.com/claude-opus-5-5-matches-fable-5-1-at-40-percent-lower-cost-as-anthropic-promises-to-fix-claudish-writing/
[7] Anthropic Releases Claude Opus 5.5: Fable 5.1-Level Performance at 40% Lower Running Cost Than Opus 5https://www.marktechpost.com/2026/09/22/anthropic-claude-opus-5-5-release/
欢迎点赞、分享、推荐
一起拥抱AI
特别声明:以上内容(如有图片或视频亦包括在内)为自媒体平台“网易号”用户上传并发布,本平台仅提供信息存储服务。
Notice: The content above (including the pictures and videos if any) is uploaded and posted by a user of NetEase Hao, which is a social media platform and only provides information storage services.