GPT-5.6 Sol Pricing Cut by 50%
摘要
OpenRouter 上 OpenAI GPT-5.6 Sol 模型的产品页,显示该模型定价下调 50%,输入降至 $2.50/1M tokens、输出降至 $15/1M tokens,上下文 1M,发布于 2026 年 7 月 9 日,知识截止 2026 年 2 月。页面列出多家提供商的性能数据(吞吐最高 57 tok/s、延迟最低 2.87s)、3 天可用性 99.80%、GPQA Diamond 与 TAU-Bench 基准分数,以及流量最大的应用(Codex、Hermes Agent 等)和调用示例。
荐读理由
GPT-5.6 Sol 输入价砍半至每百万 2.5 美元,上下文 1M,你接 OpenRouter 换下 slug 就能用,做长任务编码代理的成本直接降一截
原文
OpenAI: GPT-5.6 Sol
openai/gpt-5.6-sol
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks and long-horizon problem solving.
Modalities
In / Out Price
50% off
$2.50 / $15per 1M
Context
1M
Released
Jul 9, 2026
Knowledge Cutoff
Feb 2026
Providers
Different companies host the same model. OpenRouter routes your request to one of them based on the routing mode you pick — Balanced (price + speed), Nitro (fastest), or Exacto (highest tool-calling accuracy).
Pricing
The average price customers actually pay for this model, next to the prices providers post. Caching and discounts mean the price actually paid is often well below the listed one.
Performance
Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better).
Uptime
Uptime is the percentage of the past 3 days that at least one provider was responding to requests. Availability is the percentage of time that inference was successfully served. OpenRouter continuously monitors and uses the next-best provider when one returns an error.
Benchmarks
Scores on standardized evaluations. Higher percentages are better — and rank percentile shows where this model lands among all models on OpenRouter.
Apps
Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.
Activity
Token volume and request traffic to this model over time.
Quick Start
Drop-in code to call this model. OpenRouter's API is OpenAI-compatible — most SDKs work by just swapping the base URL. The only thing that changes between models is the model slug below.
Frequently asked questions
50% off
99.00%
99.86%
93.70%
--
--
| $5.00$2.50 | $30.00$15.00 | $0.50$0.25 | 3.12s | 39 tps | ||
| $5.00 | $30.00 | $0.50 | 5.13s | 31 tps | ||
| $5.50 | $33.00 | $0.55 | 5.92s | 14 tps | ||
| $5.50 | $33.00 | $0.55 | 4.88s | 57 tps | ||
| $5.50 | $33.00 | $0.55 | 2.87s | 46 tps |
Throughput
57tok/s
P50, best across providers
Latency
2.87s
P50, best provider
AutoExacto Benchmarks
GPQA DiamondTAU-Bench
Azure (EU)
90.6%80.0%
auto-routing
91.4%77.3%
Azure (US)
92.1%74.0%
OpenAI
91.4%74.0%
Amazon Bedrock (US)
92.3%70.4%
Azure
--77.3%
Uptime (3d)
100.00%
Availability (3d)
99.80%
Availability over the last 3 days
Last 72 hours
Availability 99.80%
3 Days Ago2 Days AgoYesterdayNow
Availability over the last 24 hours
OpenRouter Availability
99.68%
Without Routing
97.58%
When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.
A coding agent that helps you build and ship with AI
368Btokens
Hermes Agent is an open-source, self-improving AI agent by Nous Research that runs persistently with memory across sessions, and builds reusable skills from experience. It comes with 40+ built-in tools, including web search, browser automation, and vision, plus scheduled automations and subagents.
201Btokens
There are many coding agents, but this one is yours.
161Btokens
Claude Code is Anthropic's agentic coding tool that reads your entire codebase, plans and executes changes across files, runs tests, and iterates on failures, all from natural language prompts.
61.4Btokens
OpenClaw is an open-source AI agent that connects to your messaging apps and takes real actions on your behalf, from running commands and browsing the web to managing files and sending emails.
52.7Btokens
这条对你有帮助吗?