Anthropic launches Claude Sonnet 5.5 with faster, lower-cost everyday performance

Anthropic has launched Claude Sonnet 5.5, the second model in its Claude 5.5 family, positioning it as the faster and lower-cost option for well-scoped everyday work. The company says the model runs more than 30% faster than Claude Sonnet 5 and can cost up to 30% less for most work. GitHub also made Sonnet 5.5 generally available in Copilot on the same day, giving developers an immediate path to use it across coding environments.

The change is about throughput as much as capability

Anthropic is not presenting Sonnet 5.5 simply as a larger reasoning model. Its pitch combines capability, latency and economics. The company describes the model as suited to fixing bugs and producing polished documents, slides and spreadsheets, while reserving Claude Opus 5.5 for work that needs more complex judgment.

Anthropic reports a 70.6% score on Terminal-Bench 4.0 and says Sonnet 5.5 is two points behind Opus 5.5 on GDPval-AA. These are vendor-reported benchmark results, not independent evidence that the model will outperform alternatives on every production workload. They are still useful for understanding the intended product tier: Anthropic is trying to make a high-capability model practical for frequent, bounded tasks rather than only exceptional requests.

Lower cost can change agent design

For agentic systems, a faster response is not only a user-experience improvement. Agents can make many sequential model calls, invoke tools repeatedly and revise their own work. Latency and token cost therefore compound across a workflow.

Aipolix's analysis is that Sonnet 5.5's most important operational claim may be the combination of lower latency and lower cost. If Anthropic's reported gains hold for a team's workload, developers may be able to use a stronger model for more steps without increasing the total workflow budget. Alternatively, teams could keep the same task design and reduce completion time or spend.

The qualification matters: Anthropic says costs can be up to 30% lower for most work. Actual savings depend on prompt size, output length, caching, tool use and the number of model turns. A headline percentage should not be treated as a universal reduction.

GitHub Copilot makes the launch immediately actionable

GitHub announced general availability of Claude Sonnet 5.5 in Copilot on September 28. The model can be selected in Visual Studio Code, Visual Studio, Copilot CLI, the Copilot coding agent, the Copilot app, github.com, GitHub Mobile, JetBrains IDEs, Xcode and Eclipse.

GitHub says its early testing found Sonnet 5.5 matched Sonnet 5 on coding tasks while using fewer steps, tokens and tool calls, and completed tasks faster. Those observations come from GitHub rather than an independent benchmark and should be read accordingly.

Availability is also gradual. GitHub says eligible users may not see the model immediately. Access covers Copilot Pro, Pro+, Max, Business and Enterprise, with administrators able to manage model access through Copilot settings.

A model tier optimized for repeated work

The broader product signal is a shift toward optimizing the full cost of useful work rather than benchmark capability alone. In production agents, a model that is slightly less capable on the hardest task can still be the better component when it is substantially faster or cheaper across thousands of routine steps.

That creates a routing question for developers. Complex, ambiguous tasks may still justify Opus-class models, while bounded coding, editing and document-generation steps can be routed to Sonnet 5.5. The practical architecture is therefore less about choosing one model and more about assigning model tiers to different stages of a workflow.

Limitations

The main performance and cost claims currently come from Anthropic. GitHub's efficiency observations are also based on its own early testing. Neither source establishes that every workload will see the same improvements. Teams should measure end-to-end task success, latency, token consumption and tool-call count on their own workloads before changing routing or budgets.

The launch nevertheless creates a concrete deployment option today: a new Sonnet-tier model with explicit speed and cost goals and immediate integration into a major developer platform.

Sources
https://www.anthropic.com/claude-sonnet-5-5
https://github.blog/changelog/2026-09-28-claude-sonnet-5-5-in-github-copilot/