EN
Back to the archive

The archive · Developer & Business Tools · Product decision · 2023–2026

Runware bets one API can run all AI; 10B creations, $50M Series A from Dawn Capital

Runware's 'one API for all AI' powered 10B+ generations for 200K developers; the London/SF startup raised a $50M Series A led by Dawn Capital.

Runware

The betThat every AI model could run through one API: custom hardware plus a unified endpoint would undercut fragmented providers and become the default for generative media.Scaling

What the business is

London/San Francisco AI inference platform: developers call one API to generate images, video, audio, 3D and LLM outputs, powered by Runware's Sonic Inference Engine and custom AI hardware.

Starting capital$3M seed (Oct 2024) with a16z Speedrun, Lunar Ventures and others; $50M Series A led by Dawn Capital with Comcast Ventures, Speedinvest, Insight Partners and a16z Speedrun, bringing total to $66M

How it started

Flaviu Radulescu started Runware in 2023 after testing a text-to-image company and finding generative AI too slow for real products. He teamed up with Ioana Hreninciuc to build a dev-tool platform for real-time image, video and audio generation, launching with offices in London and San Francisco.

What happened

Runware grew from a $3M seed to supporting over 400,000 models with day-zero access to new releases, aggregating ~300 model classes and hundreds of thousands of variants, and claims up to 10x better price/performance for open-source models. In December 2025 it announced a $50M Series A led by Dawn Capital (total $66M), with a ~25-person team planning to expand, competitors cited as fal.ai and Replicate, and the goal of making every model available through one API.

How it ended up

Still running and scaling: Series A closed Dec 2025, Sonic Inference Pods being deployed for on-prem-style compute, and a stated target to put all 2M+ Hugging Face models on Runware by end of 2026; no exit.

Background

Runware is a London/San Francisco AI inference startup founded in 2023 by Flaviu Radulescu and Ioana Hreninciuc. Radulescu started it after testing a text-to-image company and finding that powerful generative AI was too slow for real-world products. The company's thesis: 'one API for all AI' — a single endpoint that lets developers generate images, video, audio, 3D and LLM outputs without integrating dozens of providers.

The wedge is vertical integration. Runware built the Sonic Inference Engine on custom AI hardware, with heavily optimized model loading and offloading, supporting 400,000+ models with day-zero access to new releases and claiming up to 10x better price/performance for open-source models. It sells cost-per-generation rather than blocks of GPU compute time, positioning against fal.ai and Replicate.

Traction scaled quickly: by December 2025 Runware reported more than 10 billion generations for 200K+ developers and 300M+ end-users, with customers including Wix, Quora, Together.ai, ImagineArt, Freepik, OpenArt and Higgsfield. The same month it announced a $50M Series A led by Dawn Capital, with Comcast Ventures, Speedinvest, Insight Partners and a16z Speedrun, bringing total funding to $66M.

With a ~25-person team, Runware plans to expand the Sonic Inference Engine and deploy modular 'inference PODs' near where power is cheap, targeting all 2M+ Hugging Face models on the platform by end of 2026, as the AI inference market is projected to approach $70B by 2028.

What has to be true

  • Fragmentation was the pain: teams integrating image, video and audio had to stitch together multiple providers, RPMs and pricing models before shipping anything.
  • Owning hardware plus software let Runware bend unit costs — up to 10x for open-source models — instead of reselling GPU time like fal.ai and Replicate.
  • Day-zero model support and a consistent schema made the API sticky for developers building media features at scale.
  • The model-agnostic bet de-risked the platform: whichever model won, Runware's aggregate-and-optimize layer still captured the inference spend.

What can be applied

Own the layer everyone shares: instead of betting on one model, Runware aggregated them all behind one API and built hardware to bend the cost curve — the platform wins either way.

Aftermath

As of 2026-09-02, Runware is live and scaling: 10B+ generations served across 200K+ developers, $66M raised, and a Series A-led push into Sonic Inference Pods for distributed compute. The company aims to make every Hugging Face model (2M+) available through its single API by end of 2026 and to keep expanding modalities beyond image, video and audio; no revenue or valuation figures have been disclosed.

Sources

spotted an error? The archive wants to know.

Your turn

You just read one. Describe what you are building, and see who is betting on the same thing.

Free account · 3 free questions · no card

Related cases