---
title: Welcome to SGLang
description: High-performance serving framework for large language and multimodal models.
keywords:
- sglang
- llm serving
- multimodal
- inference runtime
mode: wide
---
import { popularModels } from "/src/snippets/configs/popular-models.jsx";
import { PopularModels } from "/src/snippets/_popular_models.jsx";
{/* One hero per model, rotating. The Cookbook landing page renders the same
list as a compact strip. Edit the list, not this page. */}
{"SGLang and Miles Add Day-0 Support for DeepSeek-V4.1"}
{"September 10, 2026"}
{"Running DeepSeek-V4-Flash and Kimi-K3 on Consumer Hardware with SSD Expert Pack"}
{"August 29, 2026"}
{"Infer-forge: Harness, Loop, and Graph Engineering Around SGLang"}
{"August 28, 2026"}
{"MiniMax-H3 on 8\u00d7H200: 1.95\u00d7 Lossless, Up to 6.24\u00d7 at 0.76\u20130.91 SSIM"}
{"August 27, 2026"}
{"Qwen3.8-Flash-Next: Day-0 Support in SGLang"}
{"August 26, 2026"}
{"Fast Engine Recovery: Sub-Second Engine Restart for SGLang via Weight Cache Daemon"}
{"August 21, 2026"}
Stay connected