The Sonic Inference Pod is a fully self-contained inference compute unit. Each pod packs approximately 1,200 GPUs—primarily NVIDIA RTX PRO 6000 chips, with B200 and B300 series GPUs for larger workloads—arranged in nodes of 2–8 GPUs each . The unit sits inside a 20-foot shipping container with a chiller stacked on top, and all components are custom-built from the ground up: case-less servers, custom racks, proprietary PCIe switching, high-frequency CPUs, and local NVMe storage .
The pod runs a sealed closed-loop liquid cooling system that recirculates approximately 1.5 cubic meters of water continuously, with zero water consumption, no wastewater, and no pollution . The system maintains temperatures stable within 2°C under sustained load .
As of the August 2026 launch, Runware reported 10 pods already deployed across the United States, Europe, and Asia-Pacific . The company has identified 160 sites for rollout during the second half of 2026, with Europe now live and US West Coast deployment underway .
Before the Sonic Pod launch, Runware had already built significant customer traction. The platform has powered over 10 billion AI generations for 200,000+ developers and 300 million+ end users . Named customers include Wix, Together.ai, ImagineArt, Quora, OpenArt, Freepik, and Higgsfield AI, plus dozens of private enterprise deployments . By mid-2026, Runware's website reported 20 billion+ requests served, 500 million+ end users, 1 million+ developers, and 400,000+ models on the platform .
Runware has raised two significant rounds:
The company has raised a total of $63 million .
Runware emphasizes several environmental advantages for the Sonic Inference Pod:
The Sonic Inference Pod represents a fundamentally different approach to AI infrastructure: bringing compute to where the power and users are, rather than forcing traffic to distant centralized data centers . With 10 pods already live, 160 sites queued, and a goal of 1 GW of capacity by 2027, Runware is betting that modular, portable data centers will define the next phase of AI inference deployment.