MODEL PROFILE · Anthropic
Claude Sonnet 5.5
Released September 28, with a documented Google Cloud endpoint and separate provider token pricing.
Source reviewed: · Language model
Provider documentation summary. No hands-on results or independent ranking are claimed.
Documented facts
The following fields come from the official provider source, reviewed on the date above.
- Release date
- September 28, 2026
- API identifier
- claude-sonnet-5-5 (Google Cloud documentation)
- Availability
- Google Cloud lists general availability; account access and quotas still apply.
- Standard provider rates
- USD $2 input and $10 output per million tokens; cache reads $0.20 per million tokens. These launch prices are not a Google Cloud contract quote.
- Inputs and output
- Google Cloud: text, image and PDF inputs; text output.
- Token limits
- Google Cloud lists maximum input 1,000,000 tokens and maximum output 128,000 tokens. Check request budgeting in your chosen endpoint.
- Weights and license
- Hosted access is documented. No downloadable weights or open-weight license is established by the reviewed sources; service terms govern access.
Documented endpoint, not a leaked registry
The Google Cloud reference supplies the model identifier, GA status, modalities and limits above. Earlier reports discussed different client context figures. The documentation establishes the production endpoint; it does not authenticate a private checkpoint or explain every client reservation. Google Cloud Sonnet 5.5 reference.
The practical decision
Evaluate Sonnet on a repeatable task before changing a working deployment. For a bug fix, preserve the starting repository, failing test and allowed tools. For document analysis, retain the same source file and expected citations. Count retries, review time and tool costs alongside model tokens. A published token limit tells you what the endpoint accepts; it does not establish answer quality across the entire input. Anthropic positions Sonnet for bounded everyday work and Opus for more complex judgment. That is provider positioning, not a measured DualView ranking.
What to record in your own comparison
- Verify account access, exact identifier and endpoint-specific limits before routing production traffic.
- Keep input, output and cache-read tokens separate in the budget; confirm additional billing categories with the platform.
- Compare accepted outputs under the same task constraints and retain failed attempts.
Limitations and unknowns
DualView has not run this model or independently measured quality, latency or cost per completed task. Provider benchmarks and speed claims are not reproduced as our results. Cloud quotas, enabled tools and request limits can differ from a consumer chat product. Do not infer API limits from a model-picker label.
Compare your saved results
Use the DualView editor to inspect files you already have. The accepted-output calculator uses your own rates and counts. Neither link runs a model or purchases generation.