70%
“70% cost reduction for customers switching to open-source models on Together GPUs”
As published on together.ai. Captured by usedby on Oct 3, 2026.
What happened
Runware, a generative image and video API platform, pairs its own Sonic Inference Engine with on-demand H100, H200 and B200 GPUs from Together AI. This lets it deploy new models quickly, test across GPU generations, and handle demand spikes without long-term hardware commitments.
Summary written by usedby from the source page, in English. The figures are those of Together AI and Runware, not ours.
- 5-10x“5-10x lower pricing than competitors by blending custom and cloud infrastructure”
- 4+ billion“all while serving 4+ billion generated assets for 100,000+ developers”
When Runware needed H200s urgently for a major model launch, Together's team responded same-day, enabling rapid deployment without hardware procurement delays.
Most GPU providers just want to sell you what they have. Together AI actually listens to what we need, which is critical when you're navigating the chaos of constant model launches. This flexibility is what enables us to maintain the lowest pricing in the industry.
What the story claims, and what we checked
We compared the story with its live page on Oct 3, 2026.
- The figure: 70%CheckedPrinted word for word on the page, near the name of Runware.
- The passage quoted aboveCheckedCopied word for word from the page, near the name of Runware.
- The numbers in our summaryCheckedEach one is printed on the page.
- Runware uses Together AICheckedConfirmed line. Latest check across sources: Oct 3, 2026.
- The result itselfNot checkedWe quote it; we did not measure it.




