
claude-sonnet-5-5Claude Sonnet 5.5 是 Anthropic Claude 5.5 家族的第二款模型,2026 年 9 月 28 日发布。在 APIMODELS 上输入 $1.20、输出 $6.00 每 1M token,比 Anthropic 官方 $2 / $10 低 40%,缓存读 $0.12、缓存写 $1.50。Anthropic 把它称为速度与智能的最佳组合,定位是 Opus 5.5 更快、更省钱的搭档:最擅长边界清楚的日常任务、修 bug,以及做成品级的文档、幻灯片和表格,设计眼光也不错。相对 Sonnet 5 的新变化(均为 Anthropic 公布):输出快 30% 以上,是迄今最快的 Sonnet;同样的活用的 token 和工具调用少得多,单任务成本最多低 30%;Terminal-Bench 4.0 从 10.3% 升到 70.6%,CursorBench 4.0 从 34.1% 升到 55.5%;GDPval-AA 知识工作 1844 分,只比 Opus 5.5 低 2 分;OSWorld 2.1 操作电脑 80.1%(Sonnet 5 为 57.0%);写作更清楚;也是第一个只看截图就通关《宝可梦 红》的 Sonnet。Anthropic 同时直说:需要长时间判断的复杂开放式工作,Opus 5.5 仍明显更强。规格:100 万 token 上下文、128K 最大输出、文本与图像输入,知识截止 2026 年 6 月;分词器与 Sonnet 5 相同,最小可缓存提示词降到 512 token。自适应思考默认开启,effort 默认 high;thinking 不能再传 disabled,最低一档是 between_tools(关闭前置思考)。从 Sonnet 5 迁过来还要注意:强制工具调用(tool_choice 为 any 或指定某个工具)会报错,temperature、top_p、top_k 传非默认值会报错,thinking 块绑定产生它的模型和对话,旧的 computer_20251124 工具不再接受。通过 /v1/messages 或 OpenAI 兼容的 /v1/chat/completions 一把 key 调用,失败不收费。
View complete API reference with all parameters and examples.
Enable real-time streaming responses with Server-Sent Events.
{
"model": "claude-sonnet-5-5",
"stream": true,
"max_tokens": 1024,
"messages": [...]
}Enable Claude to use tools and call functions.
{
"model": "claude-sonnet-5-5",
"max_tokens": 1024,
"tools": [{
"name": "get_weather",
"description": "Get current weather for a location",
"input_schema": {
"type": "object",
"properties": {
"location": {"type": "string", "description": "City name"}
},
"required": ["location"]
}
}],
"tool_choice": {"type": "auto"},
"messages": [{"role": "user", "content": "What's the weather in Tokyo?"}]
}Analyze PDF documents by sending them as base64 encoded content.
{
"model": "claude-sonnet-5-5",
"max_tokens": 1024,
"messages": [{
"role": "user",
"content": [{
"type": "document",
"source": {
"type": "base64",
"media_type": "application/pdf",
"data": "<base64_encoded_pdf>"
}
}, {
"type": "text",
"text": "Summarize this document."
}]
}]
}Get structured JSON responses that match your schema.
{
"model": "claude-sonnet-5-5",
"max_tokens": 1024,
"output_format": {
"type": "json_schema",
"schema": {
"type": "object",
"properties": {
"name": {"type": "string"},
"age": {"type": "integer"}
},
"required": ["name", "age"]
}
},
"messages": [{"role": "user", "content": "Extract info: John is 30 years old"}]
}Enable Claude to search the web for up-to-date information.
{
"model": "claude-sonnet-5-5",
"max_tokens": 1024,
"tools": [{
"type": "web_search_20250305",
"name": "web_search",
"max_uses": 5
}],
"messages": [{"role": "user", "content": "What's the latest news about AI?"}]
}| Parameter | Type | Required | Description |
|---|---|---|---|
| model | string | Yes | Model identifier (e.g., claude-sonnet-5-5) |
| messages | array | Yes | Array of message objects with role and content |
| max_tokens | integer | Yes | Maximum tokens in the response (1 - 128000) |
| system | string | No | System prompt to set context |
| stream | boolean | No | Enable streaming responses (SSE) |
| temperature | number | No | Sampling temperature (0.0 - 1.0) |
| top_p | number | No | Nucleus sampling threshold (0.0 - 1.0) |
| top_k | integer | No | Top-k sampling (0 - infinity) |
| stop_sequences | array | No | Sequences that stop generation |
| tools | array | No | Function calling tools definition |
| tool_choice | object | No | Tool selection strategy (auto/any/tool) |
| thinking | object | No | Enable extended thinking mode |
| output_format | object | No | Structured output with JSON schema |
View complete API reference with streaming, thinking, and more.
Billing: Cost = (input_tokens * input_price + output_tokens * output_price) / 1,000,000
$1.20 / $6.00 每 1M token,Anthropic 官方 $2 / $10;缓存读 $0.12、写 $1.50
Anthropic:迄今最快的 Sonnet,token 和工具调用更少,单任务成本最多低 30%
Anthropic:Terminal-Bench 4.0 70.6%(Sonnet 5 为 10.3%),CursorBench 4.0 55.5%(Sonnet 5 为 34.1%)
GDPval-AA 1844,Opus 5.5 为 1846;OSWorld 2.1 操作电脑 80.1%(Sonnet 5 为 57.0%)
文本与图像输入;知识截止 2026 年 6 月;分词器与 Sonnet 5 相同
最低一档是 between_tools(传 disabled 会报错);强制工具调用会报错
Claude Sonnet 5.5 是由 Anthropic 提供的大语言模型 API。Claude Sonnet 5.5 是 Anthropic Claude 5.5 家族的第二款模型,2026 年 9 月 28 日发布。在 APIMODELS 上输入 $1.20、输出 $6.00 每 1M token,比 Anthropic 官方 $2 / $10 低 40%,缓存读 $0.12、缓存写 $1.50。Anthropic 把它称为速度与智能的最佳组合,定位是 Opus 5.5 更快、更省钱的搭档:最擅长边界清楚的日常任务、修 bug,以及做成品级的文档、幻灯片和表格,设计眼光也不错。相对 Sonnet 5 的新变化(均为 Anthropic 公布):输出快 30% 以上,是迄今最快的 Sonnet;同样的活用的 token 和工具调用少得多,单任务成本最多低 30%;Terminal-Bench 4.0 从 10.3% 升到 70.6%,CursorBench 4.0 从 34.1% 升到 55.5%;GDPval-AA 知识工作 1844 分,只比 Opus 5.5 低 2 分;OSWorld 2.1 操作电脑 80.1%(Sonnet 5 为 57.0%);写作更清楚;也是第一个只看截图就通关《宝可梦 红》的 Sonnet。Anthropic 同时直说:需要长时间判断的复杂开放式工作,Opus 5.5 仍明显更强。规格:100 万 token 上下文、128K 最大输出、文本与图像输入,知识截止 2026 年 6 月;分词器与 Sonnet 5 相同,最小可缓存提示词降到 512 token。自适应思考默认开启,effort 默认 high;thinking 不能再传 disabled,最低一档是 between_tools(关闭前置思考)。从 Sonnet 5 迁过来还要注意:强制工具调用(tool_choice 为 any 或指定某个工具)会报错,temperature、top_p、top_k 传非默认值会报错,thinking 块绑定产生它的模型和对话,旧的 computer_20251124 工具不再接受。通过 /v1/messages 或 OpenAI 兼容的 /v1/chat/completions 一把 key 调用,失败不收费。 通过 APIMODELS 平台,您可以使用统一的 API 接口调用该模型,按量计费、价格透明。当前定价:Input: $1.20, Output: $6.00 per 1M tokens。
构建智能对话系统,自动回答用户问题,提升客户服务效率。
自动撰写文章、邮件、广告文案等文本内容,提高内容产出效率。
辅助代码编写、调试和代码审查,加速软件开发流程。
理解和分析非结构化数据,提取关键信息并生成报告摘要。
Claude Sonnet 5.5 通过 APIMODELS 平台调用,当前定价:Input: $1.20, Output: $6.00 per 1M tokens。按量计费,用多少付多少。
在 APIMODELS 注册账号并获取 API Key,然后通过我们的统一 API 端点调用即可。我们提供详细的 API 文档和 cURL、Python、Node.js 代码示例。
APIMODELS 通过聚合平台提供与官方相同的 Claude Sonnet 5.5 模型。我们提供统一的 API 接口,无需分别注册各平台账号,一个 API Key 即可调用所有模型。
输入 $1.20、输出 $6.00 每百万 token,缓存读 $0.12、缓存写 $1.50。Anthropic 官方是 $2 / $10、缓存读 $0.20、5 分钟缓存写 $2.50,所以四项都低 40%。按实际 token 扣费,失败的请求不收钱,也没有月费门槛。我们不提供 Batch 接口。
按 Anthropic 公布的数据:输出快 30% 以上,是迄今最快的 Sonnet;同样的活用的 token 和工具调用少得多,单任务成本最多低 30%;智能体编码大幅提升,Terminal-Bench 4.0 从 10.3% 到 70.6%、CursorBench 4.0 从 34.1% 到 55.5%;GDPval-AA 知识工作 1844 分,只比 Opus 5.5 低 2 分;OSWorld 2.1 操作电脑从 57.0% 到 80.1%;写作更清楚,做文档、幻灯片和表格更拿手。规格上上下文 100 万、输出 128K 不变,分词器相同,最小可缓存提示词从 1,024 降到 512 token。
两种都行:Anthropic 原生 POST https://api.apimodels.app/v1/messages,或 OpenAI 兼容的 /v1/chat/completions,model 填 claude-sonnet-5-5,Authorization: Bearer 你的 apimodels key。流式、工具调用、图片输入都支持;Claude Code、Anthropic SDK 只需把 base URL 换成 https://api.apimodels.app、key 换成 apimodels 的。从 Sonnet 5 迁过来,模型名改成 claude-sonnet-5-5 即可。
在我们这里 Sonnet 5.5 是 $1.20 / $6.00,Sonnet 5 是 $1.60 / $8.00,Opus 5.5 是 $2.40 / $12.00。Sonnet 5.5 的输入输出价都比 Sonnet 5 低,又更快、单任务用的 token 更少,所以新项目没有理由再选 Sonnet 5;唯一的例外是缓存很重的负载,Sonnet 5 的缓存读 $0.10、写 $1.40 略低于 Sonnet 5.5 的 $0.12 / $1.50。Opus 5.5 贵一倍,Anthropic 自己的说法是需要长时间判断的复杂开放式工作它仍明显更强;边界清楚的日常编码、修 bug、办公文档用 Sonnet 5.5 就够。
按 Anthropic 的 cache_control 用法打标记即可。写入缓存按 $1.50 每百万 token 计,之后命中按 $0.12 计,只有输入价 $1.20 的十分之一;没命中的部分照常按输入价。Sonnet 5.5 的最小可缓存提示词是 512 token,比这短的前缀不会被缓存。我们实测连发三次同一前缀,第一次写入、后两次命中。Claude Code 会自动设置缓存标记。
上下文 100 万 token、单次最多输出 128K token,输入支持文字和图片,知识截止 2026 年 6 月。Anthropic 对 Sonnet 5.5 的规定:自适应思考默认开、effort 默认 high,thinking 不能再传 disabled(会报 400),最低一档是 between_tools;tool_choice 为 any 或指定某个工具会报错,只能用 auto 或 none;temperature、top_p、top_k 传非默认值会报错,不传即可;thinking 块绑定产生它的模型和对话,改动历史消息后重放旧 thinking 块可能报错,对话请只追加不修改。
在 APIMODELS,Claude Sonnet 5.5 与 60+ 个模型共用一个 API Key、一个余额,所以选型只看合不合适,不存在锁定。它支持 40% Below Official、30%+ Faster、1M Context、Agentic Coding,你可以在价格和能力上把它和其它大语言模型模型对比,换模型只需改一个模型名字符串——无需新账号、无需重新对接。所有大语言模型模型与实时价格见 apimodels.app/models。
Claude Sonnet 5.5 支持:40% Below Official、30%+ Faster、1M Context、Agentic Coding。完整参数与调用方式见 APIMODELS 的 API 文档。
可以。APIMODELS 提供可直接访问的统一 API,一个 API Key 即可调用 Claude Sonnet 5.5,无需分别注册官方账号、也无需自行处理官方接口的网络访问。
我们支持 Stripe(Visa、Mastercard 等国际信用卡)和支付宝付款。充值后积分即时到账。