Skip to main content
usedby

70%

“70% cost reduction for customers switching to open-source models on Together GPUs”

As published on together.ai. Captured by usedby on Oct 3, 2026.

What happened

Runware, a generative image and video API platform, pairs its own Sonic Inference Engine with on-demand H100, H200 and B200 GPUs from Together AI. This lets it deploy new models quickly, test across GPU generations, and handle demand spikes without long-term hardware commitments.

Summary written by usedby from the source page, in English. The figures are those of Together AI and Runware, not ours.

  • 5-10x“5-10x lower pricing than competitors by blending custom and cloud infrastructure”
  • 4+ billion“all while serving 4+ billion generated assets for 100,000+ developers”

When Runware needed H200s urgently for a major model launch, Together's team responded same-day, enabling rapid deployment without hardware procurement delays.

From the page. together.ai, captured Oct 3, 2026

Most GPU providers just want to sell you what they have. Together AI actually listens to what we need, which is critical when you're navigating the chaos of constant model launches. This flexibility is what enables us to maintain the lowest pricing in the industry.

Ioana Hreninciuc, Co-Founder, Runware. Source, captured Oct 3, 2026

What the story claims, and what we checked

We compared the story with its live page on Oct 3, 2026.

  • The figure: 70%CheckedPrinted word for word on the page, near the name of Runware.
  • The passage quoted aboveCheckedCopied word for word from the page, near the name of Runware.
  • The numbers in our summaryCheckedEach one is printed on the page.
  • Runware uses Together AICheckedConfirmed line. Latest check across sources: Oct 3, 2026.
  • The result itselfNot checkedWe quote it; we did not measure it.

Same company, same tool or same industry.

6דDecagon achieved nearly 6× cost reduction per turn compared to closed models like GPT-5 mini.”Decagon uses Together AI. Another customer of Together AIunder two weeks“Together’s team shipped FP8, FP4, and INT4 quantized variants of the Cogito models in under two weeks”Deep Cogito uses Together AI. Another customer of Together AICursor uses Together AIAnother customer of Together AI