As published on deepgram.com. Checked by usedby on Oct 9, 2026.
What happened
Telnyx runs Deepgram Flux for real-time speech-to-text on its own GPUs at its telephony points of presence, so call audio is transcribed inside its network. The transcripts feed its call control and agent orchestration layer for turn-taking, LLM reasoning and text-to-speech in voice AI calls.
Summary written by usedby from the source page, in English. The figures are those of Deepgram and Telnyx, not ours.
Because Deepgram Flux is running at the edge, inside the media plane, Telnyx can deliver sub-second end-to-end latency under load while preserving the reliability and observability of its carrier network.
We chose Deepgram Flux for its high accuracy and consistent real-time performance, running directly on Telnyx-managed GPUs at the edge so our voice experiences feel instant; latency is constrained by physics, not external APIs.
What the story claims, and what we checked
What we compared with the page.
- The passage quoted aboveCheckedCopied word for word from the page, near the name of Telnyx.
- Telnyx uses DeepgramCheckedConfirmed line.
- The result itselfNot checkedWe quote it; we did not measure it.




