š Qwen3.6-27B is here, and we have day 0 support on SGLang ā 27B params, beats Qwen3.5-397B-A17B across major coding benchmarks ā Agentic coding ā Text + multimodal reasoning ā Thinking / non-thinking modes Smaller model. Bigger results. Try it on SGLang now š„ https://t.co/x7tVVdvoFm
LMSYS Org Adds Day Zero SGLang Support for Qwen3.6-27B Reasoning Model
Ā· Updated
LMSYS Org integrated immediate support for Qwen3.6-27B into its SGLang inference framework, enabling high-speed serving of the new 27-billion parameter model. The model outperforms the massive Qwen3.5-397B-A17B on coding benchmarks and introduces native thinking modes for complex reasoning.
- Parameter count
- 27B
- Reasoning modes
- Thinking and non-thinking
- Framework support
- SGLang (day 0)
- Modality
- Text and multimodal
- Coding performance
- Beats Qwen3.5-397B-A17B
- Availability
- Open weights, self-hostable
This release marks an efficiency shift, following the pattern of the Qwen3.5-397B-A17B but surpassing it on coding benchmarks. While the previous generation used a massive Mixture-of-Experts architecture (specialized sub-networks for efficiency), Qwen3.6 achieves superior results with a smaller footprint. It continues the trajectory of the Qwen3-Coder-Next series by prioritizing agentic coding.
You can deploy Qwen3.6-27B immediately via SGLang to build high-throughput coding agents. This integration is part of a broader shift toward day-zero inference support across major serving frameworks. The model is available for self-hosting, providing a cost-effective alternative to larger frontier models for teams requiring local, high-performance environments.
Still wondering? A few quick answers below.
Every HeadsUpAI update is written based on its original source and reviewed before it's published. Read our editorial standards ā





