Guozhen AIGlobal AI field notes and model intelligence

Realtime AI News

Runware launches portable modular data center Sonic Inference Pod for edge inference

AI infrastructure company Runware launched its Sonic Inference Pod, a transportable modular data center designed to deliver faster, cheaper inference than traditional GPU clouds. The company says 10 pods are already deployed across the U.S., Europe, and Asia-Pacific, serving customers such as Higgsfield AI and Wix.

Published
Runware推出模块化数据中心Sonic Inference Pod,主打靠近用户的分布式推理
Image source: techcrunch.com

On Tuesday, AI infrastructure company Runware announced the launch of its own modular data center called Sonic Inference Pod. Designed as a single transportable unit, the Pod represents a more flexible kind of compute that can sit alongside hyperscalers' massive data center projects.

Runware says the Pod can deliver inference at higher quality and lower cost than other serverless inference platforms and GPU clouds. Its modular design makes it easy to add capacity quickly by standing up new pods rather than expanding a fixed data center.

"We believe distributed compute, positioned closer to end users for faster inference, is what will win in the long term," co-founder and CEO Flaviu Radulescu told TechCrunch. Beyond price, the pods can scale fast, deploy anywhere there is power, adapt quickly to new hardware, and use a closed-loop, water-free cooling system that can be built in days, versus the months or even years traditional data centers require.

"Demand for inference is growing faster than facilities can be built," Radulescu said. "What we want is to power the world's intelligence, to be the backbone every AI model runs on with capacity that keeps up with demand instead of throttling it."

Runware currently has 10 pods deployed across the U.S., Europe, and Asia-Pacific, and already provides inference to companies including Higgsfield AI and Wix, with 160 sites available to power its pods. The company raised a $50 million Series A in December, led by Dawn Capital and Comcast Ventures, to fund infrastructure for image generation; it sees the expansion into pods as part of its core mission of providing inference to companies rather than a single product.

Meanwhile, AI labs like OpenAI and SpaceX are still racing to build data centers across the U.S. — OpenAI is reportedly close to a $500 billion deal to build a data center in Ohio. Radulescu doesn't see those projects as a threat, describing the pods' flexibility as the key differentiator.

"Every pod runs as part of a single network, so requests go wherever there's capacity, closer to the users, and if one pod goes offline, traffic moves to another," he said, adding that a system failure means one pod is down rather than a whole fixed facility. Customers who want dedicated hardware get whole pods to themselves.

The watch point: AI data centers are controversial largely because of the resources they consume. Whether Runware's small, fast-to-deploy pods can thread the needle on cost, efficiency, and public scrutiny will determine whether its promise of keeping capacity ahead of inference demand holds up.

Why it matters

Runware's portable pod model challenges the hyperscaler playbook for inference capacity, potentially reshaping how AI compute is deployed and scaled closer to end users.

RunwareAI InfrastructureData Center
Back to realtime news

Nearby Updates

All

08/04, 21:00

Linux Foundation and Open Secure AI Alliance Propose SAFE Guidelines for Sharing AI Incident Intelligence

On the opening day of Black Hat in Las Vegas, the Linux Foundation published a request for comments on the Shared AI Findings Exchange (SAFE), a proposed framework for turning agentic AI security incidents into shared, ecosystem-wide protection. The guidelines were drafted by a working group of the Open Secure AI Alliance, which now counts more than 120 member organizations.

08/04, 21:14

Open-source 'Claude Science' arrives: zero dependencies, MIT license, 30+ built-in research skills

On August 4, Chinese tech outlet QbitAI reported that a joint laboratory of Peking University and YuanKong AI released an open-source version of "Claude Science," built with zero dependencies, an MIT license, and more than 30 built-in research skills. The project offers a ready-to-use open-source option for AI agents in scientific research.

08/04, 21:52

HappyRobot Raises $150M Series C to Expand Its AI-Agent Platform for Logistics and Supply Chains

HappyRobot, which builds AI agents for logistics and supply chain operations, has raised $150 million in Series C funding to expand its platform, as reported by AI Insider. The round underscores growing investor appetite for AI agents built for specific industries rather than general-purpose assistants.

08/04, 21:58

Liquid AI Releases LFM2.5-2.6B: A 2.6B-Parameter Agent Model Built for On-Device Deployment

Liquid AI has released LFM2.5-2.6B, a 2.6B-parameter model designed to run capable AI agents entirely on-device, with support for tool calling and multi-step workflows. The company says it hits up to 220 tokens per second on an Apple M5 Max in under 2.5GB of memory, making local agent deployment viable on laptops and phones.