DualView

MODEL PROFILE · Anthropic

Claude Sonnet 5.5

Released September 28, with a documented Google Cloud endpoint and separate provider token pricing.

Source reviewed: · Language model

Provider documentation summary. No hands-on results or independent ranking are claimed.

Documented facts

The following fields come from the official provider source, reviewed on the date above.

Release date
September 28, 2026
API identifier
claude-sonnet-5-5 (Google Cloud documentation)
Availability
Google Cloud lists general availability; account access and quotas still apply.
Standard provider rates
USD $2 input and $10 output per million tokens; cache reads $0.20 per million tokens. These launch prices are not a Google Cloud contract quote.
Inputs and output
Google Cloud: text, image and PDF inputs; text output.
Token limits
Google Cloud lists maximum input 1,000,000 tokens and maximum output 128,000 tokens. Check request budgeting in your chosen endpoint.
Weights and license
Hosted access is documented. No downloadable weights or open-weight license is established by the reviewed sources; service terms govern access.

Documented endpoint, not a leaked registry

The Google Cloud reference supplies the model identifier, GA status, modalities and limits above. Earlier reports discussed different client context figures. The documentation establishes the production endpoint; it does not authenticate a private checkpoint or explain every client reservation. Google Cloud Sonnet 5.5 reference.

The practical decision

Evaluate Sonnet on a repeatable task before changing a working deployment. For a bug fix, preserve the starting repository, failing test and allowed tools. For document analysis, retain the same source file and expected citations. Count retries, review time and tool costs alongside model tokens. A published token limit tells you what the endpoint accepts; it does not establish answer quality across the entire input. Anthropic positions Sonnet for bounded everyday work and Opus for more complex judgment. That is provider positioning, not a measured DualView ranking.

What to record in your own comparison

Limitations and unknowns

DualView has not run this model or independently measured quality, latency or cost per completed task. Provider benchmarks and speed claims are not reproduced as our results. Cloud quotas, enabled tools and request limits can differ from a consumer chat product. Do not infer API limits from a model-picker label.

Compare your saved results

Use the DualView editor to inspect files you already have. The accepted-output calculator uses your own rates and counts. Neither link runs a model or purchases generation.

Related reporting

Related model profiles

All model profiles and dated changes · Subscribe via RSS