我们已推出 Claude Opus 4.8(claude-opus-4-8),这是我们目前能力最强且广泛可用的模型。Claude Opus 4.8 在 Claude API、Amazon Bedrock、Google Cloud 和 Microsoft Foundry 上默认支持 100 万 token 的上下文窗口、12.8 万 token 的最大输出长度,以及与 Claude Opus 4.7 相同的工具集和平台功能。请参阅《Claude Opus 4.8 新特性》了解能力改进、新功能和迁移指南。
我们推出了对话中途系统消息功能。在 Claude Opus 4.8 上,你可以在用户轮次之后(需遵守放置规则)在消息数组中发送 role: "system" 消息,从而在长时间运行的会话中指令发生变化时保留提示缓存命中。无需 beta 标头。
拒绝响应中的 stop_details 字段现已公开文档化;它会返回一个类别(cyber、bio 或 null)以及人类可读的解释,以便你的应用程序将不同类别的拒绝路由到正确的下一步。无需 beta 标头。
在 Claude Opus 4.8 上,effort 参数在所有界面(包括 Claude Code 和 Messages API)中默认设为 high。
在 Claude Opus 4.8 上,提示缓存的最小可缓存提示长度为 1024 个 token,低于 Claude Opus 4.7。
启用自适应思考后,Claude Opus 4.8 仅在需要时才触发推理,相比相同 effort 级别的 Claude Opus 4.7,减少了浪费的思考 token。
Claude Opus 4.8 支持高分辨率图像输入(长边最多 2576 像素),与 Claude Opus 4.7 相同。
任务预算现在支持 Claude Opus 4.8。
顾问工具现在支持 Claude Opus 4.8。
计算机使用现在支持 Claude Opus 4.8。
Claude Opus 4.8 的快速模式仅作为研究预览在 Claude API 上提供。
将采样参数 temperature、top_p 或 top_k 设置为非默认值会在 Claude Opus 4.8 上返回 400 错误,与 Claude Opus 4.7 相同。详情请参阅迁移指南。
在 Claude Code 中,我们已将自动模式扩展到更多用户,用于长时间运行的任务。请参阅 Claude Code 文档。
在 Claude Code 中,Max 计划用户现在默认在 Claude Opus 4.8 上使用快速模式。请参阅 Claude Code 文档。
在 Claude Code 中,工作流作为研究预览提供,让你能够定义和运行多步骤的代理计划。请参阅 Claude Code 文档。
我们已弃用 Claude Opus 4.6 的快速模式,将在发布后约 30 天移除。请迁移到 Claude Opus 4.8 或 Claude Opus 4.7 的快速模式。更多信息请阅读《快速模式》。
关于本版本中 claude.ai、Cowork、Claude for Microsoft 365 及其他 Claude 应用的更新,请参阅《Claude 应用发布说明》。
We've launched Claude Opus 4.8 (claude-opus-4-8), our most capable generally available model. Claude Opus 4.8 supports a 1M token context window by default on the Claude API, Amazon Bedrock, Google Cloud, and Microsoft Foundry, 128k max output tokens, and the same set of tools and platform features as Claude Opus 4.7. See What's new in Claude Opus 4.8 for capability improvements, new features, and migration guidance. We've launched mid-conversation system messages. On Claude Opus 4.8, you can send role: "system" messages after a user turn (subject to placement rules) in the messages array, preserving prompt cache hits when instructions change during a long-running session. No beta header is required. The stop_details field on refusal responses is now publicly documented; it returns a category (cyber, bio, or null) and a human-readable explanation, so your application can route different classes of refusal to the right next step. No beta header is required. On Claude Opus 4.8, the effort parameter defaults to high across all surfaces, including Claude Code and the Messages API. On Claude Opus 4.8, the minimum cacheable prompt length for prompt caching is 1,024 tokens, lower than on Claude Opus 4.7. With adaptive thinking enabled, Claude Opus 4.8 triggers reasoning only when a turn needs it, reducing wasted thinking tokens compared to Claude Opus 4.7 at the same effort level. Claude Opus 4.8 supports high-resolution image input (up to 2576 pixels on the long edge), same as Claude Opus 4.7. Task budgets now support Claude Opus 4.8. The advisor tool now supports Claude Opus 4.8. Computer use now supports Claude Opus 4.8. Fast mode for Claude Opus 4.8 is available as a research preview on the Claude API only. Setting the sampling parameters temperature, top_p, or top_k to a non-default value returns a 400 error on Claude Opus 4.8, same as on Claude Opus 4.7. See the migration guide for details. In Claude Code, we've expanded Auto mode to more users for long-running tasks. See the Claude Code documentation. In Claude Code, Max plan users now default to fast mode on Claude Opus 4.8. See the Claude Code documentation. In Claude Code, Workflows are available as a research preview, letting you define and run multistep agentic plans. See the Claude Code documentation. We've deprecated fast mode for Claude Opus 4.6, with removal approximately 30 days after launch. Migrate to fast mode for Claude Opus 4.8 or Claude Opus 4.7. Read more in Fast mode. For updates to claude.ai, Cowork, Claude for Microsoft 365, and other Claude apps in this release, see the release notes for Claude Apps.
本文内容采集自官方网站,排版和翻译可能与原页面存在差异。
阅读官方全文