CompaniesInvestorsPeople
Home
Loading

aVenture is in Beta: research coverage is expanding as we build, so please independently verify key details before making investment decisions.

aVenture is in Beta: research coverage is expanding as we build, so please independently verify key details before making investment decisions.

Get in Touch

  • Contact

  • Request a Demo

  • Request Data Updates

  • Add a Company

Research

  • Companies

  • Investors

  • People

aVenture

  • Download App

  • Pricing

Download the aVenture Research beta for iOS and iPadOSDownload aVenture Research on the Mac App Store

Resources

  • Documentation

  • Use Cases

  • CLI

  • MCP

  • Feature Requests

  • Sitemap

Member

Backed by

Ask AI about aVenture

© aVenture Investment Company, 2026. All rights reserved.

San Francisco, CA, USA

Privacy · Terms of Service

aVenture Investment Company ("aVenture") is an independent research platform providing detailed analysis and data on startups, venture capital investments, and key industry individuals. It is not a registered investment adviser, broker-dealer, or investment advisor and does not provide investment advice or recommendations. The data provided by aVenture does not constitute recommendations or advice, whether by methodology, analysis, AI-generated content, or a statement written by a staff member of aVenture.

aVenture is not affiliated with any of the people, companies, organizations, government agencies, regulatory bodies, or investment funds we provide coverage for on this site unless explicitly stated otherwise. Users assume full responsibility for decisions made based on information obtained from this platform. Links to external websites do not imply endorsement or affiliation with aVenture. Any links that provide the ability to invest in a primary or secondary transaction in a company are for convenience only and do not constitute solicitations or offers to buy or sell an investment. Investors should exercise heightened precaution and due diligence when investing in private companies, especially those not independently audited.

While we strive to provide valuable insights with objectivity and professional diligence, we cannot guarantee the accuracy of the information provided on our platform. Before making any investment decisions, you should verify the accuracy of all pertinent details for your decision. To the fullest extent permitted by law, aVenture shall not be liable for any direct, indirect, incidental, consequential, or financial damages arising from use of this site, whether by consumers of its contents directly or by persons or organizations covered by our research, even if we are advised of the possibility. Our best-efforts processes and correction request forms do not create a warranty or duty of care.

Profiles on this platform may include content generated in part by large language models (LLMs, artificial intelligence) that aggregate publicly available sources (e.g., SEC EDGAR, public filings, press releases). Source attribution is provided where known; always verify statements and claims here against original sources before relying on any data. Content on our site may contain inaccuracies, omissions, or what are commonly called 'hallucinations' if generated in part or in full by AI / LLMs. The risk can also exist even when content is written by a human, as internal and third-party sources may also have inaccuracies for the same or different reasons. While we randomly audit a proportion of content, this is not exhaustive.

We recommend that an independent auditor be hired to verify the accuracy of the information before relying on it for any sensitive decisions. By accessing this platform, you agree not to rely solely on any information generated by AI, aggregated, or sourced or written otherwise on this site, for investment, financial, or other decisions. aVenture assumes no responsibility for inaccuracies, omissions, or hallucinations. You must independently verify all data from primary sources. Use of this platform constitutes your waiver of claims for reliance-based damages, including negligent misrepresentation. To report an error, request a correction, or dispute information about a company or individual, contact us via our request data updates form.

Loading
Loading
Home
News
Anthropic launches Haiku 5.5 at a much lower price

From The New Stack

By Frederic Lardinois

October 7, 2026

Anthropic launches Haiku 5.5 at a much lower price

Anthropic launches Haiku 5.5 at a much lower price

Anthropic on Wednesday launched Claude Haiku 5.5, the first new version of its smallest and most affordable model in nearly a year.

That’s Anthropic’s third 5.5 model in a month, though what’s still missing is Fable 5.5. That model, however, will likely go through a much longer review process, so it’s no surprise its launch is taking a bit longer.

How Haiku 5.5 pricing works

Like with previous iterations, the company describes Haiku 5.5 as its “fastest and most efficient model,” but this time, it is also much cheaper. While Haiku 4.5 cost $1/$5 per million input/output tokens, Anthropic is bringing the price down to $0.10/$0.50 for requests under 100,000 tokens.

Anthropic describes Haiku 5.5 as its “fastest and most efficient model,” but this time, it is also much cheaper.

For larger requests, the new price is $0.50/$2.50 (Haiku 4.5 costs the same, no matter the number of tokens in a request), but Anthropic notes that about 90% of requests to Haiku 4.5 fell into the lower-priced category (which is likely a function of the kind of work that developers have sent to Haiku in the past).

Those are 90% and 50% cuts to the per-token prices, respectively. Anthropic puts the average savings at around 75%. That accounts for a mix of requests and an updated tokenizer that uses slightly more tokens per task.

This new version is the first Haiku model with effort controls (the default is medium), giving developers control over how many tokens the model uses on a given task.

Haiku 5.5 Haiku 4.5 GPT-6 Luna Sonnet 5.5
Knowledge workGDPval-AA v2.1 1620 735 1437 1840
Knowledge workAA-Briefcase v1.1 1578 614 1336 1824
Computer useOSWorld 2.1 (Offline subset) 72.4% 15.7% 48.9% 83.9%
Multidisciplinary reasoningHumanity’s Last Exam No tools 45.9% 10.2% – 56.9%
With tools 57.4% 18.7% – 64.5%
Agentic codingTerminal-Bench 4.0 39.2% 0.0% 16.4% 70.6%
Agentic codingFrontierCode 1.1 (Main) 46.4% – 42.4% 52.1%(Xhigh)
Visual reasoningChartography (no tools) 46.4% 6.4% 29.1% 61.6%

Small models, new jobs

The use case for these small models has always been to handle high-volume tasks like summarization, classification, and routing.

That, of course, is also where decision models like Jev are currently making a splash — and at an even lower price.

Going forward, that may not be where these small models like Haiku or OpenAI’s GPT-6 Luna will be most useful, so it’s probably no surprise that Anthropic also notes that the model can handle tasks like compaction and database queries, as well as agentic workloads where speed matters, including live customer support and browser use.

Haiku 5.5 benchmarks

Anthropic’s own evaluations show a clear improvement over Haiku 4.5. On the offline subset of OSWorld 2.1, which tests computer use, Haiku 5.5 scored 72.4%, up from 15.7% for its predecessor and ahead of GPT-6 Luna’s 48.9%.

On the GDPval-AA v2.1 knowledge-work benchmark, Haiku 5.5 scored 1,620, compared with 735 for Haiku 4.5 and 1,437 for GPT-6 Luna. Unsurprisingly, Sonnet 5.5 remains ahead on both benchmarks.

On Terminal-Bench 4.0, which tests complex, multi-step tasks in a command-line environment, Haiku 5.5 scored 39.2%, compared with 0% for Haiku 4.5, 16.4% for GPT-6 Luna and 70.6% for Sonnet 5.5.

How Chinese rivals compare

Anthropic is only directly comparing Haiku 5.5 to its own models and OpenAI’s GPT-6 Luna. But when it comes to small models, many developers are also looking at competitors form Z.ai, Alibaba, and others.

Artificial Analysis currently gives Z.ai’s GLM-5.3-Flash a score of 1,647 on GDPval-AA v2.1 and 1,454 on AA-Briefcase v1.1. Anthropic, in comparison, reports 1,620 and 1,578 for Haiku 5.5, respectively.

There are still cheaper options, too. On Alibaba’s international service, Qwen3.7 Flash costs $0.03 per million input tokens and $0.13 per million output tokens for inputs of up to 32,000 tokens. For inputs above 32,000 and up to 256,000 tokens, those rates rise to $0.10 and $0.40, respectively.

Haiku 5.5 is available on the Claude Platform, Amazon Web Services, Google Cloud and Microsoft Azure. Developers using Anthropic’s platform can access it as claude-haiku-5-5. The company is also adding beta support for computer use and browser use to its Python and TypeScript SDKs.

Also new: Sonnet 5.5 cache price drop, API credits for Max and Team subscriptions

Alongside the launch, Anthropic is cutting Sonnet 5.5’s cache-read price in half, from $0.20 to $0.10 per million tokens. The company says this should make most agentic tasks about 20% cheaper. The cut is rolling out across platforms on Wednesday, though some existing Azure and Google Cloud customers will have to wait a few days.

Anthropic is cutting Sonnet 5.5’s cache-read price in half

Anthropic is also adding monthly API credits to its Max and Team subscriptions this week. Max 5x subscribers will get $100 per month and Max 20x subscribers will get $200. Team subscribers will receive up to $500, pooled across their users. You can use the credits with any model on the Claude Platform.

Haiku 5.5 also comes with tighter cybersecurity safeguards than Haiku 4.5, although Anthropic says these allow a wider range of defensive work than Sonnet 5.5’s safeguards. They still block penetration testing. Organizations that need broader access for cybersecurity or biology work can apply to Anthropic’s verification programs.

TRENDING STORIES
Before joining The New Stack as its senior editor for AI, Frederic was the enterprise editor at TechCrunch, where he covered everything from the rise of the cloud and the earliest days of Kubernetes to the advent of quantum computing....
Read more from Frederic Lardinois

View original article on thenewstack.io

Most Recent

Ex-Trump AI Adviser Sriram Krishnan Targets Raising a $500 Million Venture Fund

Sriram Krishnan, a key AI adviser to President Donald Trump who left the White House in June, is raising an AI-focused venture fund …

Oct 7, 2026

GPT-6 and Intelligent UI for everyone

GPT‑6 is rolling out globally in ChatGPT with Intelligent UI, delivering faster responses with visuals and interactive experiences you can explore and use directly.

Oct 7, 2026

Surface RTX Spark Dev Box is available for preorder for $5,999

Microsoft's high-end AI mini PC for developers starts shipping in November. — Microsoft's Nvidia-powered Surface RTX Spark Dev Box …

Oct 7, 2026

Meta rolls out new AI tools to detect ads that secretly lead to child sexual abuse material

Meta launches new AI tools after discovering ads on its platforms that may look normal but direct users to harmful content elsewhere online.

Oct 7, 2026

Similar Posts

Introducing Claude Sonnet 5.5

Claude Sonnet 5.5 is a clear upgrade over Claude Sonnet 5, runs 30%+ faster, and costs up to 30% less for most work.

Sep 28, 2026

Anthropic launches Claude Haiku 5.5 with 90% API price reduction, matching GPT-6 Luna

Anthropic calls Haiku its fastest model at standard speeds, while acknowledging that Opus in Fast Mode runs faster. It does not supply a tokens-per-second figure in the materials reviewed for this article.

Oct 7, 2026

Anthropic's Claude Fable 5.1 and Mythos 5.1 arrive with a 75% cost reduction for Fable cache reads

It's only the first day of September 2026, but the month and fall season are already off to the races in AI land, as Anthropic has just released its latest and most powerful large language models yet — Claude Fable 5.1 and Claude Mythos 5.1. The two names refer to the same underlying model. Fable 5.

Sep 1, 2026

Anthropic Subscriptions Offer 5x+ More Value Than OpenAI

Limit testing every AI subscription plan from Anthropic, OpenAI, Meta, SpaceXSI, MiniMax, Moonshot, Z.ai, Cursor, and Cognition

Oct 5, 2026