Home
Loading

aVenture is in Alpha: During this preview period, you should expect the research data to be limited and may not yet meet our exacting standards. We've made the decision to provide early access to our data to showcase the product as we build, but you should not yet rely upon it alone for your investment decisions.

aVenture is in Alpha: During this preview period, you should expect the research data to be limited and may not yet meet our exacting standards. We've made the decision to provide early access to our data to showcase the product as we build, but you should not yet rely upon it alone for your investment decisions.

Get in touch

  • Contact

  • Request a demo

  • Request data updates

  • Add a company

Research

  • Companies

  • Investors

  • People

aVenture

  • Sitemap

  • Feature requests

Member

Backed by

© aVenture Investment Company, 2026. All rights reserved.

San Francisco, CA, USA

Privacy Policy

aVenture Investment Company ("aVenture") is an independent research platform providing detailed analysis and data on startups, venture capital investments, and key industry individuals. It is not a registered investment adviser, broker-dealer, or investment advisor and does not provide investment advice or recommendations. The data provided by aVenture does not constitute recommendations or advice, whether by methodology, analysis, AI-generated content, or a statement written by a staff member of aVenture.

aVenture is not affiliated with any of the people, companies, organizations, government agencies, regulatory bodies, or investment funds we provide coverage for on this site unless explicitly stated otherwise. Users assume full responsibility for decisions made based on information obtained from this platform. Links to external websites do not imply endorsement or affiliation with aVenture. Any links that provide the ability to invest in a primary or secondary transaction in a company are for convenience only and do not constitute solicitations or offers to buy or sell an investment. Investors should exercise heightened precaution and due diligence when investing in private companies, especially those not independently audited.

While we strive to provide valuable insights with objectivity and professional diligence, we cannot guarantee the accuracy of the information provided on our platform. Before making any investment decisions, you should verify the accuracy of all pertinent details for your decision. To the fullest extent permitted by law, aVenture shall not be liable for any direct, indirect, incidental, consequential, or financial damages arising from use of this site, whether by consumers of its contents directly or by persons or organizations covered by our research, even if we are advised of the possibility. Our best-efforts processes and correction request forms do not create a warranty or duty of care.

Profiles on this platform may include content generated in part by large language models (LLMs, artificial intelligence) that aggregate publicly available sources (e.g., SEC EDGAR, public filings, press releases). Source attribution is provided where known; always verify statements and claims here against original sources before relying on any data. Content on our site may contain inaccuracies, omissions, or what are commonly called 'hallucinations' if generated in part or in full by AI / LLMs. The risk can also exist even when content is written by a human, as internal and third-party sources may also have inaccuracies for the same or different reasons. While we randomly audit a proportion of content, this is not exhaustive.

We recommend that an independent auditor be hired to verify the accuracy of the information before relying on it for any sensitive decisions. By accessing this platform, you agree not to rely solely on any information generated by AI, aggregated, or sourced or written otherwise on this site, for investment, financial, or other decisions. aVenture assumes no responsibility for inaccuracies, omissions, or hallucinations. You must independently verify all data from primary sources. Use of this platform constitutes your waiver of claims for reliance-based damages, including negligent misrepresentation. To report an error, request a correction, or dispute information about a company or individual, contact us via our request data updates form.

Loading
Loading
Home
News
Runware uses custom hardware and advanced orchestration for fast AI inference

From TechCrunch

By Romain Dillet

October 1, 2024

Runware uses custom hardware and advanced orchestration for fast AI inference

Runware uses custom hardware and advanced orchestration for fast AI inference

Sometimes, a demo is all you need to understand a product. And that’s the case with Runware. If you head over to Runware’s website, enter a prompt and hit enter to generate an image, you’ll be surprised by how quickly Runware generates the image for you — it takes less than a second.

Runware is a newcomer in the AI inference, or generative AI, startup landscape. The company is building its own servers and optimizing the software layer on those servers to remove bottlenecks and improve inference speeds for image generation models. The startup has already secured $3 million in funding from Andreessen Horowitz’s Speedrun, LakeStar’s Halo II and Lunar Ventures.

The company doesn’t want to reinvent the wheel. It just wants to make it spin faster. Behind the scenes, Runware manufactures its own servers with as many GPUs as possible on the same motherboard. It has its own custom-made cooling system and manages its own data centers.

When it comes to running AI models on its servers, Runware has optimized the orchestration layer with BIOS and operating system optimizations to improve cold start times. It has developed its own algorithms that allocate interference workloads.

The demo is impressive by itself. Now, the company wants to use all this work in research and development and turn it into a business.

Unlike many GPU hosting companies, Runware isn’t going to rent its GPUs based on GPU time. Instead, it believes companies should be encouraged to speed up workloads. That’s why Runware is offering an image generation API with a traditional cost-per-API-call fee structure. It’s based on popular AI models from Flux and Stable Diffusion.

“If you look at Together AI, Replicate, Hugging Face — all of them — they are selling compute based on GPU time,” co-founder and CEO Flaviu Radulescu told TechCrunch. “If you compare the amount of time it takes for us to make an image versus them. And then you compare the pricing, you will see that we are so much cheaper, so much faster.”

“It’s going to be impossible for them to match this performance,” he added. “Especially in a cloud provider, you have to run on a virtualized environment, which adds additional delays.”

As Runware is looking at the entire inference pipeline, and optimizing hardware and software, the company hopes that it will be able to use GPUs from multiple vendors in the near future. This has been an important endeavor for several startups as Nvidia is the clear leader in the GPU space, which means that Nvidia GPUs tend to be quite expensive.

“Right now, we use just Nvidia GPUs. But this should be an abstraction of the software layer,” Radulescu said. “We can switch a model from GPU memory in and out very, very fast, which allow us to put multiple customers on the same GPUs.

“So we are not like our competitors. They just load a model into the GPU and then the GPU does a very specific type of task. In our case, we’ve developed this software solution, which allow us to switch a model in the GPU memory as we do inference.“

If AMD and other GPU vendors can create compatibility layers that work with typical AI workloads, Runware is well positioned to build a hybrid cloud that would rely on GPUs from multiple vendors. And that will certainly help if it wants to remain cheaper than competitors at AI inference.

View original article on techcrunch.com

Most Recent

Neil Rimer thinks the AI money is coming back out

Neil Rimer thinks the AI money is coming back out

Neil Rimer, the venture capitalist who co-founded Index Ventures, predicts the historic wealth AI is generating in Silicon Valley will have to be redistributed, voluntarily or involuntarily.

Jul 17, 2026

Databricks hits $188B valuation, extending its run as AI’s favorite second act

Databricks hits $188B valuation, extending its run as AI’s favorite second act

Databricks has remade its image into an AI company and has published research on the cost savings of open weight AI models for coding.

Jul 17, 2026

Nuclear startup Valar Atomics in talks to raise new funding at $6B valuation

Nuclear startup Valar Atomics in talks to raise new funding at $6B valuation

The potential deal highlights a growing trend of complex, multi-stage funding rounds that mask true entry prices.

Jul 17, 2026

AI-powered travel agency Fora hits unicorn status, raises $60M

AI-powered travel agency Fora hits unicorn status, raises $60M

Travel agency Fora announced a $60 million Series D round led by Forerunner and Tactile Ventures, valuing the company at $1 billion.

Jul 16, 2026

Similar Posts

Inference.ai matches AI workloads with cloud GPU compute

Inference.ai matches AI workloads with cloud GPU compute

GPUs’ ability to perform many computations in parallel make them well-suited tonrunning today’s most capable AI. But GPUs are becoming tougher to procure, asncompanies of all sizes increase their investments in AI-powered products.nNvidia’s best-performing AI cards sold out last year, and the CEO of chipmakernTSMC suggested that general supply could be […]

Jan 30, 2024

Inflection CEO says it’s done trying to make next generation AI models

Inflection CEO says it’s done trying to make next generation AI models

Just last year, Inflection AI was as hot as a startup could be, releasing best-in-class AI models it claimed could outperform technology from OpenAI, Meta, and Google. That’s a stark contrast compared to today, as Inflection’s new CEO tells TechCrunch that his startup is simply no longer trying to compete on that front. Between then and now, there’s of course been a major change at Inflection. Microsoft hired then-CEO Mustafa Suleyman to run its own AI business, and paid the startup $650 millio

Nov 26, 2024

GMI Cloud secures $82M in Series A for its GPU cloud infrastructure

GMI Cloud secures $82M in Series A for its GPU cloud infrastructure

The AI boom has spurred massive demand for graphics processing units (GPUs). As many enterprises seek to integrate AI technology into their systems, providers of GPU infrastructure are helping businesses get access to the chips they need. In the latest development, GMI Cloud, a San Jose-based startup that provides GPU cloud infrastructure, has raised an $82 million Series A led by Headline Asia with participation from strategic investors such as Banpu, a Thailand energy firm, and Wistron, a Ta

Oct 29, 2024

Andreessen Horowitz helps founders meet compute needs with ‘Oxygen’ private GPU cluster

Andreessen Horowitz helps founders meet compute needs with ‘Oxygen’ private GPU cluster

Andreessen Horowitz has a massive cluster of Nvidia H100 GPUs to help its portfolio of AI startups meet their compute needs, the venture capital firm confirmed for the first time on Wednesday. The program, called “Oxygen”, allows their portfolio companies to train or operate their AI models without negotiating market rates. A16Z’s Oxygen cluster gives startups some breathing room, so to speak, to compete against larger tech companies – such as Google, Meta, and Microsoft – in building large AI

Oct 23, 2024

Most Recent

Neil Rimer thinks the AI money is coming back out

Neil Rimer thinks the AI money is coming back out

Neil Rimer, the venture capitalist who co-founded Index Ventures, predicts the historic wealth AI is generating in Silicon Valley will have to be redistributed, voluntarily or involuntarily.

Jul 17, 2026

Databricks hits $188B valuation, extending its run as AI’s favorite second act

Databricks hits $188B valuation, extending its run as AI’s favorite second act

Databricks has remade its image into an AI company and has published research on the cost savings of open weight AI models for coding.

Jul 17, 2026

Nuclear startup Valar Atomics in talks to raise new funding at $6B valuation

Nuclear startup Valar Atomics in talks to raise new funding at $6B valuation

The potential deal highlights a growing trend of complex, multi-stage funding rounds that mask true entry prices.

Jul 17, 2026

AI-powered travel agency Fora hits unicorn status, raises $60M

AI-powered travel agency Fora hits unicorn status, raises $60M

Travel agency Fora announced a $60 million Series D round led by Forerunner and Tactile Ventures, valuing the company at $1 billion.

Jul 16, 2026

Similar Posts

Inference.ai matches AI workloads with cloud GPU compute

Inference.ai matches AI workloads with cloud GPU compute

GPUs’ ability to perform many computations in parallel make them well-suited tonrunning today’s most capable AI. But GPUs are becoming tougher to procure, asncompanies of all sizes increase their investments in AI-powered products.nNvidia’s best-performing AI cards sold out last year, and the CEO of chipmakernTSMC suggested that general supply could be […]

Jan 30, 2024

Inflection CEO says it’s done trying to make next generation AI models

Inflection CEO says it’s done trying to make next generation AI models

Just last year, Inflection AI was as hot as a startup could be, releasing best-in-class AI models it claimed could outperform technology from OpenAI, Meta, and Google. That’s a stark contrast compared to today, as Inflection’s new CEO tells TechCrunch that his startup is simply no longer trying to compete on that front. Between then and now, there’s of course been a major change at Inflection. Microsoft hired then-CEO Mustafa Suleyman to run its own AI business, and paid the startup $650 millio

Nov 26, 2024

GMI Cloud secures $82M in Series A for its GPU cloud infrastructure

GMI Cloud secures $82M in Series A for its GPU cloud infrastructure

The AI boom has spurred massive demand for graphics processing units (GPUs). As many enterprises seek to integrate AI technology into their systems, providers of GPU infrastructure are helping businesses get access to the chips they need. In the latest development, GMI Cloud, a San Jose-based startup that provides GPU cloud infrastructure, has raised an $82 million Series A led by Headline Asia with participation from strategic investors such as Banpu, a Thailand energy firm, and Wistron, a Ta

Oct 29, 2024

Andreessen Horowitz helps founders meet compute needs with ‘Oxygen’ private GPU cluster

Andreessen Horowitz helps founders meet compute needs with ‘Oxygen’ private GPU cluster

Andreessen Horowitz has a massive cluster of Nvidia H100 GPUs to help its portfolio of AI startups meet their compute needs, the venture capital firm confirmed for the first time on Wednesday. The program, called “Oxygen”, allows their portfolio companies to train or operate their AI models without negotiating market rates. A16Z’s Oxygen cluster gives startups some breathing room, so to speak, to compete against larger tech companies – such as Google, Meta, and Microsoft – in building large AI

Oct 23, 2024