---
title: Welcome to SGLang
description: High-performance serving framework for large language and multimodal models.
keywords:
- sglang
- llm serving
- multimodal
- inference runtime
mode: wide
---
import { popularModels } from "/src/snippets/configs/popular-models.jsx";
import { PopularModels } from "/src/snippets/_popular_models.jsx";
{/* One hero per model, rotating. The Cookbook landing page renders the same
list as a compact strip. Edit the list, not this page. */}
{"Qwen3.8-Flash-Next: Day-0 Support in SGLang"}
{"August 26, 2026"}
{"Fast Engine Recovery: Sub-Second Engine Restart for SGLang via Weight Cache Daemon"}
{"August 21, 2026"}
{"Chasing the Batch-1 Floor: Ling-3.0-flash Speculative Decode on Blackwell"}
{"August 21, 2026"}
{"Mooncake for Miles: From Fragmented Rollout Data to Efficient Bulk I/O"}
{"August 20, 2026"}
{"Pushing the Limits of Serving DeepSeek-V4-Pro"}
{"August 19, 2026"}
{"Miles v0.1: Production-level Post-training"}
{"August 18, 2026"}
Stay connected