# How SGLang uses Nebius AI Cloud

**2×**: “SGLang achieved a 2× boost in throughput and markedly lower latency on one node.”

As published on [nebius.com](https://nebius.com/customer-stories/sglang). Captured by usedby on 2026-10-07.

- Company: [SGLang](https://www.usedby.ai/companies/sglang.md)
- Tool: [Nebius AI Cloud](https://www.usedby.ai/tools/nebius.md)

## What the story says

SGLang, an open-source serving framework for large language models, ran and benchmarked DeepSeek R1 on Nebius AI Cloud infrastructure. It used on-demand compute clusters to test inference optimizations such as new attention algorithms, FP8 matrix multiplication and kernel fusion.

Summary written by usedby from the source page, in English. The figures are those of Nebius AI Cloud and SGLang, not ours.

> SGLang achieved a 2× boost in throughput and markedly lower latency on one node. In practice, this means faster answers from R1, even on long prompts or with dozens of users at once.

## What usedby checked

We compared the story with its live page on 2026-10-07.

- Checked: the figure 2× is printed word for word on the page, near the name of SGLang.
- Checked: the passage quoted above is copied word for word from the page, near the name of SGLang.
- Checked: each number in our summary is printed on the page.
- Checked: SGLang uses Nebius AI Cloud. Confirmed line. Latest check across sources: 2026-10-07.
- Not checked: the result itself. We quote it; we did not measure it.

---
Source: https://www.usedby.ai/case-studies/sglang-nebius · How we check: https://www.usedby.ai/methodology
