CompaniesInvestorsPeople
Home
Loading

aVenture is in Beta: research coverage is expanding as we build, so please independently verify key details before making investment decisions.

aVenture is in Beta: research coverage is expanding as we build, so please independently verify key details before making investment decisions.

Get in Touch

  • Contact

  • Request a Demo

  • Request Data Updates

  • Add a Company

Research

  • Companies

  • Investors

  • People

aVenture

  • Download App

  • Pricing

Download the aVenture Research beta for iOS and iPadOSDownload aVenture Research on the Mac App Store

Resources

  • Documentation

  • Use Cases

  • CLI

  • MCP

  • Feature Requests

  • Sitemap

Member

Backed by

Ask AI about aVenture

© aVenture Investment Company, 2026. All rights reserved.

San Francisco, CA, USA

Privacy · Terms of Service

aVenture Investment Company ("aVenture") is an independent research platform providing detailed analysis and data on startups, venture capital investments, and key industry individuals. It is not a registered investment adviser, broker-dealer, or investment advisor and does not provide investment advice or recommendations. The data provided by aVenture does not constitute recommendations or advice, whether by methodology, analysis, AI-generated content, or a statement written by a staff member of aVenture.

aVenture is not affiliated with any of the people, companies, organizations, government agencies, regulatory bodies, or investment funds we provide coverage for on this site unless explicitly stated otherwise. Users assume full responsibility for decisions made based on information obtained from this platform. Links to external websites do not imply endorsement or affiliation with aVenture. Any links that provide the ability to invest in a primary or secondary transaction in a company are for convenience only and do not constitute solicitations or offers to buy or sell an investment. Investors should exercise heightened precaution and due diligence when investing in private companies, especially those not independently audited.

While we strive to provide valuable insights with objectivity and professional diligence, we cannot guarantee the accuracy of the information provided on our platform. Before making any investment decisions, you should verify the accuracy of all pertinent details for your decision. To the fullest extent permitted by law, aVenture shall not be liable for any direct, indirect, incidental, consequential, or financial damages arising from use of this site, whether by consumers of its contents directly or by persons or organizations covered by our research, even if we are advised of the possibility. Our best-efforts processes and correction request forms do not create a warranty or duty of care.

Profiles on this platform may include content generated in part by large language models (LLMs, artificial intelligence) that aggregate publicly available sources (e.g., SEC EDGAR, public filings, press releases). Source attribution is provided where known; always verify statements and claims here against original sources before relying on any data. Content on our site may contain inaccuracies, omissions, or what are commonly called 'hallucinations' if generated in part or in full by AI / LLMs. The risk can also exist even when content is written by a human, as internal and third-party sources may also have inaccuracies for the same or different reasons. While we randomly audit a proportion of content, this is not exhaustive.

We recommend that an independent auditor be hired to verify the accuracy of the information before relying on it for any sensitive decisions. By accessing this platform, you agree not to rely solely on any information generated by AI, aggregated, or sourced or written otherwise on this site, for investment, financial, or other decisions. aVenture assumes no responsibility for inaccuracies, omissions, or hallucinations. You must independently verify all data from primary sources. Use of this platform constitutes your waiver of claims for reliance-based damages, including negligent misrepresentation. To report an error, request a correction, or dispute information about a company or individual, contact us via our request data updates form.

Loading
Loading
Home
News
Anthropic launches Claude Haiku 5.5 with 90% API price reduction, matching GPT-6 Luna

From VentureBeat

By Carl Franzen

October 7, 2026

Anthropic launches Claude Haiku 5.5 with 90% API price reduction, matching GPT-6 Luna

Anthropic launches Claude Haiku 5.5 with 90% API price reduction, matching GPT-6 Luna
Featured
Carl Franzen

Anthropic today released Claude Haiku 5.5, cutting token prices by 90% for requests below 100,000 tokens and targeting the repetitive work that can make enterprise AI expensive at scale: summarizing documents, classifying information, querying databases and handling smaller assignments for more capable agents.

The model starts at $0.10 per million input tokens and $0.50 per million output tokens, matching OpenAI’s GPT-6 Luna. Anthropic estimates that workloads cost approximately 75% less to run than on Haiku 4.5, once request sizes and changes in token consumption are considered.

The launch also brings lower Sonnet 5.5 caching charges and monthly API credits for Max and Team subscribers, broadening the announcement beyond a single inexpensive model.

A support role for enterprise agents

Anthropic positions Haiku as a supporting worker for Opus and Sonnet, which remain its recommended choices for demanding coding assignments. A larger model might assemble a financial presentation while Haiku retrieves the revenue figure needed for one slide, according to an example supplied by financial AI company Rogo.

“It's accurate enough that we'd trust it there and fast and cheap enough that we can run it a lot,” Alex Wang, who works in applied AI at Rogo, said in a statement provided by Anthropic.

The largest savings apply to the shortest requests

The pricing has an important boundary: above 100,000 tokens, Haiku 5.5 costs $0.50 per million input tokens and $2.50 per million output tokens. Those rates represent a 50% reduction from Haiku 4.5, compared with the 90% reduction below that threshold.

Charge, per million tokens

Haiku 5.5: below 100,000 tokens

Haiku 5.5: above 100,000 tokens

Haiku 4.5

Input

$0.10

$0.50

$1.00

Output

$0.50

$2.50

$5.00

Cache reads

$0.01

$0.05

$0.10

Cache writes

$0.125

$0.625

$1.25

Source: Anthropic’s launch materials. The materials do not specify how requests of exactly 100,000 tokens are billed or precisely which tokens determine the threshold.

Anthropic says approximately 90% of Haiku 4.5 requests fall into the shorter category. Its estimated 75% average workload savings also accounts for an updated tokenizer—the mechanism that divides content into billable units—which uses somewhat more tokens for equivalent work.

How Haiku compares with OpenAI, Google and Grok

The lower rates make Haiku competitive with other proprietary models, but do not establish it as the cheapest option across the market. OpenAI’s Luna matches all four of Haiku’s lower-tier rates, including caching.

Model

Input per million tokens

Output per million tokens

Pricing qualification

Anthropic Claude Haiku 5.5

$0.10

$0.50

Requests below 100,000 tokens

OpenAI GPT-6 Luna

$0.10

$0.50

Base rates; higher rates above 272,000 input tokens

Google Gemini 3.5 Flash-Lite

$0.30

$2.50

Standard pricing

Anthropic Claude Haiku 5.5

$0.50

$2.50

Requests above 100,000 tokens

Google Gemini 3.8 Flash

$0.75

$3.75

Promotional standard rates through Dec. 31, 2026

Grok 4.3

$1.25

$2.50

Below 200,000 prompt tokens

Grok 4.7

$2.00

$6.00

Below 200,000 prompt tokens

OpenAI GPT-6.1 Sol

$2.00

$10.00

Base standard rates

Anthropic Claude Sonnet 5.5

$2.00

$10.00

Listed launch-table rates

Sources: Anthropic’s supplied materials; official pricing for GPT-6 Luna, GPT-6.1 Sol, Gemini and Grok, checked October 7. Figures exclude caching, batch discounts, tools, regional premiums and negotiated rates. This compares prices, not equivalent capabilities.

The context thresholds matter: Haiku’s higher tier is five times its lower rate, while Luna’s surcharge begins above 272,000 input tokens. Token prices alone also cannot establish the cost of completing a job; token consumption, retries and accuracy affect that calculation.

Benchmark gains come with an effort-setting caveat

Anthropic’s benchmark table shows gains over Haiku 4.5 and leads over GPT-6 Luna on the selected evaluations below, while Sonnet 5.5 remains ahead. These are vendor-reported results, not independent verification.

Evaluation

Haiku 5.5

Haiku 4.5

GPT-6 Luna

Sonnet 5.5

GDPval-AA v2.1: knowledge work

1,620

735

1,437

1,840

AA-Briefcase v1.1: knowledge work

1,578

614

1,336

1,824

OSWorld 2.1, offline subset: computer operation

72.4%

15.7%

48.9%

83.9%

Terminal-Bench 4.0: agent-based coding

39.2%

0.0%

16.4%

70.6%

FrontierCode 1.1, main evaluation

46.4%

—

42.4%

52.1%*

Source: Anthropic. Sonnet’s FrontierCode result is labeled High effort; the dash indicates no result supplied. The first two rows are scores, not percentages.

Haiku 5.5 introduces adjustable effort levels, with medium as the default. That distinction matters for interpreting its results: Anthropic’s Terminal-Bench chart places the approximately 39% score at maximum effort, while medium scores approximately 20%.

Customers report faster results, but throughput remains unspecified

Anthropic calls Haiku its fastest model at standard speeds, while acknowledging that Opus in Fast Mode runs faster. It does not supply a tokens-per-second figure in the materials reviewed for this article.

Customer reports offer narrower evidence. Box’s VP of AI Products, Yashodha Bhavnani, reports an 11-point improvement over Haiku 4.5 with approximately half the latency, without identifying the scoring scale. Asana’s Aaron Vinh, a staff software engineer, reports task-completion latency falling more than 30% and inference per agent turn accelerating by as much as 2.5 times, compared with an unnamed model.

HubSpot’s Ze’ev Klapow, a distinguished software engineer, reports a 92.8% average across three runs of its CRM evaluation, the strongest result among the models it tested. These company-supplied testimonials do not establish comparable throughput across providers.

Sonnet price cuts and new subscription API credits

Alongside Haiku, Sonnet 5.5 cache reads are getting a cost cut from $0.20 to $0.10 per million tokens.

Anthropic estimates approximately 20% savings on typical agent workloads; actual savings depend on how much stored context an application reuses.

Monthly API credits roll out this week: $100 for Max 5x subscribers, $200 for Max 20x and up to $500 shared among a Team subscription’s users. They can be spent on any model through Anthropic’s platform. The supplied draft does not explain Team allocation or rollover conditions.

Cloud availability and deployment considerations

Anthropic says Haiku launches through its own platform, AWS, Google Cloud and Microsoft Azure, with the API identifier claude-haiku-5-5. Python and TypeScript SDK updates add beta capabilities for operating browsers and computers. Some existing Azure and Google Cloud customers receive Sonnet’s caching reduction over the following days.

The model also tightens cybersecurity restrictions compared with Haiku 4.5, including blocking penetration testing under its standard safeguards. Organizations seeking broader cybersecurity or biology access can apply through Anthropic’s verification programs.

For enterprise developers, the practical test is whether Haiku can complete a narrowly defined assignment reliably enough to justify repeated use. Its lower prices and reported gains make that test more attractive; its own benchmark results still support reserving more complex work for larger models.

View original article on venturebeat.com

Most Recent

Ex-Trump AI Adviser Sriram Krishnan Targets Raising a $500 Million Venture Fund

Sriram Krishnan, a key AI adviser to President Donald Trump who left the White House in June, is raising an AI-focused venture fund …

Oct 7, 2026

GPT-6 and Intelligent UI for everyone

GPT‑6 is rolling out globally in ChatGPT with Intelligent UI, delivering faster responses with visuals and interactive experiences you can explore and use directly.

Oct 7, 2026

Surface RTX Spark Dev Box is available for preorder for $5,999

Microsoft's high-end AI mini PC for developers starts shipping in November. — Microsoft's Nvidia-powered Surface RTX Spark Dev Box …

Oct 7, 2026

Meta rolls out new AI tools to detect ads that secretly lead to child sexual abuse material

Meta launches new AI tools after discovering ads on its platforms that may look normal but direct users to harmful content elsewhere online.

Oct 7, 2026

Similar Posts

Anthropic Subscriptions Offer 5x+ More Value Than OpenAI

Limit testing every AI subscription plan from Anthropic, OpenAI, Meta, SpaceXSI, MiniMax, Moonshot, Z.ai, Cursor, and Cognition

Oct 5, 2026

Anthropic's Claude Fable 5.1 and Mythos 5.1 arrive with a 75% cost reduction for Fable cache reads

It's only the first day of September 2026, but the month and fall season are already off to the races in AI land, as Anthropic has just released its latest and most powerful large language models yet — Claude Fable 5.1 and Claude Mythos 5.1. The two names refer to the same underlying model. Fable 5.

Sep 1, 2026

Introducing Claude Sonnet 5.5

Claude Sonnet 5.5 is a clear upgrade over Claude Sonnet 5, runs 30%+ faster, and costs up to 30% less for most work.

Sep 28, 2026

Google Gemini 4 Argon closes the gap with OpenAI and Anthropic but doesn't take a clear lead

Gemini 4 Argon is Google's first frontier model in over seven months. It matches GPT-6 Astra in independent testing but can't keep up with Anthropic's Claude Opus 5.5. The per-token price is low, but Argon burns through more than twice as many tokens per task as Astra. Select testers get access firs

Sep 30, 2026