Anthropic launches Claude Haiku 5.5 with up to 75% cost reduction
Read more
Olhar Digital
olhardigital.com.br

Anthropic launches Claude Haiku 5.5 with up to 75% cost reduction

Anthropic introduced Claude Haiku 5.5 this Wednesday (7th), a new artificial intelligence (AI) model designed to execute tasks requiring high speed, large processing volume, and low operational cost.

This launch represents an update to the Haiku 4.5 version and consolidates the new Claude 5.5 family, which also includes the Opus 5.5 and Sonnet 5.5 models. Anthropic claims that Haiku 5.5 is currently its smallest, fastest, most affordable, and most capable model.

One of the highlights is the significant cost reduction; the company reports that the new model operates on average 75% cheaper than Haiku 4.5. For requests limited to 100 thousand tokens, this decrease reaches an impressive 90% per token.

In terms of pricing, for prompts up to 100 thousand tokens, Haiku 5.5 costs US$ 0.10 per million input tokens (approximately R$ 0.53) and US$ 0.50 per million output tokens (about R$ 2.65). When the limit exceeds 100 thousand tokens, the values rise to US$ 0.50 (R$ 2.65) per million input tokens and US$ 2.50 (R$ 13.25) per million output tokens.

Anthropic observes that about 90% of calls made to Haiku 4.5 were within the 100 thousand token limit, which justifies the greater price savings for most uses of the previous model.

Furthermore, there was a halving of the cache reading cost for Sonnet 5.5, dropping from US$ 0.20 to US$ 0.10 per million tokens. According to Anthropic, this results in an approximate 20% reduction in costs for agent tasks.

Claude Haiku 5.5 also stands out as the first Haiku category model to incorporate an adjustable effort level system. This feature allows developers to define the degree of reasoning the model should employ before generating a response, enabling prioritization of speed and cost in simple activities or increasing effort for more complex demands, thus balancing cost and intelligence according to the application.

Additionally, the platform documentation indicates that the model has a context window of one million tokens.

Anthropic also released the results of its own comparative tests against Haiku 4.5. In the OSWorld 2.1 benchmark, which evaluates the ability of agents to operate computers for complex tasks, Haiku 5.5 achieved 72.4%, surpassing Haiku 4.5's 15.7%. In Terminal-Bench 4.0, focused on terminal programming tasks, the new model reached 39.2%, while Haiku 4.5 registered 0%. In the Humanity’s Last Exam test, Haiku 5.5 scored 45.9% without tools and 57.4% with tools, contrasting with the previous model's 10.2% and 18.7%. It is important to note that this data is disclosed by Anthropic itself and does not constitute an independent evaluation.

Comparison with OpenAI Models

Anthropic also positioned Haiku 5.5 on performance charts alongside OpenAI models. In OSWorld 2.1, for example, Haiku 5.5 achieved 72.4%, while GPT-6 Luna recorded 48.9%. In Terminal-Bench 4.0, the percentages were 39.2% and 16.4%, respectively. Since these results were selected and published by Anthropic itself, they should be analyzed within the company's methodology and not as an impartial comparison between the models.

The company also reported improvements in the safety behavior of Haiku 5.5 compared to Haiku 4.5. According to Anthropic, the new model demonstrated fewer instances of misaligned behavior and less propensity to assist in inappropriate uses. Cybersecurity protections are stricter than in Haiku 4.5 but less restrictive than those applied to the company's latest models. The system allows for a wider variety of defensive tasks but prohibits penetration testing and other attack-related practices. Biological safeguards maintain the same standard established in Sonnet 5, Sonnet 5.5, and Opus 5.

Along with the launch, Anthropic offered monthly API credits for subscribers to the Max and Team plans. Max 5x users will receive US$ 100 (about R$ 530) monthly, and Max 20x subscribers will receive US$ 200 (about R$ 1,060). For Team clients, the benefit totals US$ 500 (about R$ 2,650), distributed among organization members. These credits can be used on any Anthropic model on the platform and aim to encourage the development of agents and applications that use the API.

The company also updated its SDKs for Python and TypeScript, which now support in beta mode usage on computer and browser.

Haiku 5.5 Availability

Claude Haiku 5.5 is already accessible on Anthropic's platforms, including Amazon Web Services (AWS), Google Cloud, and Microsoft Azure. On the Claude Platform, developers can access the model using the identifier claude-haiku-5-5. With this launch, Anthropic concludes, in just over two weeks, the update of its main Claude 5.5 line. Haiku 5.5 establishes itself as the most economical alternative in the family, while Sonnet and Opus remain focused on more sophisticated tasks.

Similar stories

Anthropic launches Claude Sonnet 5.5 with greater speed and cost reduction for everyday tasks
Read more
olhardigital.com.br

Anthropic launches Claude Sonnet 5.5 with greater speed and cost reduction for everyday tasks

Anthropic has introduced Claude Sonnet 5.5, a new intermediate model within its Claude 5.5 family. According to the company, this model offers faster performance, exceeding the speed of Sonnet 5 by over 30%, and can provide savings of up to 30% per task.

The goal of this launch is to optimize the execution of well-defined activities, such as code correction, document drafting, presentation creation, and spreadsheet generation. Furthermore, Sonnet 5.5 has received enhancements in areas like design, collaboration, and security, while the Opus 5.5 model remains focused on more intricate demands.

In terms of pricing, Sonnet 5.5 maintained the values of the previous version: US$ 2 per million input tokens (equivalent to about R$ 10.42), US$ 10 per million output tokens (R$ 52.10), and US$ 0.20 per million tokens in cache reads (R$ 1.04).

Tokens are the minimum textual units that AI models use to interpret a request and generate a response. Since Sonnet 5.5 generally requires fewer of these units to complete a task, Anthropic projects a cost reduction of up to 30% per job. Additionally, the model generates responses with over 30% more speed.

This performance gain was notable in programming tests. In Terminal-Bench 4.0, an evaluation that measures an AI's ability to execute complex tasks via command line—a textual environment used by developers—Sonnet 5.5 achieved 70.6%, compared to 10.3% for Sonnet 5.

Anthropic also shared other comparative results between Sonnet 5.5 and Sonnet 5. Despite the progress, Anthropic emphasizes that Opus 5.5 still demonstrates greater strength in complex and open-ended tasks requiring extended reasoning. However, in certain evaluations, Sonnet 5.5 reached results close to those of the most advanced model when operating at maximum load.

The evolution is particularly evident in the field of programming, where Sonnet 5.5 can quickly assimilate codebases and complete tasks in fewer steps. Testers also pointed out the model's efficiency in tool usage.

In preliminary tests conducted by Epic, Claude Sonnet 5.5 met the expected quality standard for a top-tier model. Daniel Vogel, Director of Operations at Epic Games, commented on this in a note.

The improvements are not limited to code; Anthropic also emphasizes performance in creation and knowledge tasks, covering interfaces and presentations. In an internal test, the model received quarterly materials, transcripts from a public corporation, and a presentation template. The result was an operational review of ten slides that two specialists deemed ready to be sent.

Curtis Allen, Principal Engineer at Slack, stated: 'Without changing our prompts, Claude Sonnet 5.5 performed better than Sonnet 5 in almost all our internal Slackbot evaluations.'

The increase in capability was accompanied by new layers of protection. Sonnet 5.5 is the first model in the Sonnet line to incorporate cybersecurity safeguards and reserve mechanisms similar to those employed by Anthropic in Opus 5.5.

The safeguards function as devices designed to restrict the use of AI in scenarios classified as high risk. In the context of cybersecurity, routine development activities remain accessible, while more dangerous requests may be redirected to Sonnet 5.

In the biological sphere, protections remain identical to those of Sonnet 5, focusing on requests considered high risk. Sonnet 5.5 is already available on various platforms, including Amazon Web Services, Google Cloud, and Microsoft Azure. Claude Haiku 5.5 will be launched in the coming weeks to complete the new model range. The original article was initially published in Olhar Digital.

Popular