系列:Community Tracker Notes

vLLM & SGLang Community Tracker · 2026-08-06 — Both Frameworks Land Ling-3.0-flash Same Day; Step Stalls

★ Most Worth Watching Today

On the same day, both frameworks landed the domestic model Ling-3.0-flash; meanwhile StepFun (Step) shows zero pushes for a third straight day and is reported to have split its internal strategy into two lines.

The contrast: StepFun-ai org’s main repos are still frozen at 06-01 / 04-03, and its official vllm fork stopped at 05-28; while on 08-05 media reported Step internally established two independent strategy lines — “large model” and “agent terminal” — with overseas commercialization led by API (especially voice models).

Conclusion: Step’s stagnation in the open-source inference ecosystem is not an ad-hoc scheduling issue; it aligns with where the org is moving resources. Treat Step as an observation benchmark for domestic inference enablement, but down-weight the expectation.

1. Baseline (no new tag from any of the three)

ObjectLatest stableReleasedThis window
vLLMv0.26.02026-07-27main: 52 commits
SGLangv0.5.162026-07-25main: 61 commits (sgl-kernel → 0.4.6)
StepFunStep-3.7-Flash / 3.5-Flashpush frozen 06-01 / 04-03zero GitHub progress; strategy news only

Note: sgl-kernel bumped to 0.4.6 via #33678 — a sub-package version; the SGLang main tag is unchanged.

2. The “Same-Frame” Signal of Domestic Model Enablement

Ling-3.0-flash entering both vLLM and SGLang trunks the same day, both explicitly with speculative decoding (MTP drafter / cookbook throughput), tells us:

  1. “Day-one dual-framework” support for new domestic models is now normal; the ecosystem moved from “asking frameworks to adopt” to “frameworks competing to adopt”.
  2. Speculative decoding (see this blog’s “Speculative Decoding Notes” ep1–ep6) is now a standard launch feature, not optional.
  3. GB300 measured GSM8K 96%+ confirms domestic models “can run” on inference cards; competition shifts to “run cheaper, run faster”.

3. How to Read Step’s “Absence”

Three days of zero commits + main repos frozen for months + strategy-split news all point the same way:

4. Positioning of This “Tracker” Series

This series (tr) turns daily vLLM / SGLang upstream commits, model cookbooks, and domestic-model enablement progress into readable notes with a “same-day same-frame / who is falling behind” comparative lens — filling the timeliness gap left by the principle-heavy “vLLM & SGLang Framework Notes (fw)”.

5. Today’s Watchlist (to verify)

Sources: GitHub vLLM / SGLang commit streams (2026-08-05 01:00Z → 08-06 01:00Z); Sina/NetEase/Toutiao 08-05 strategy-line reporting.

觉得有用?欢迎点赞、收藏,或请作者喝咖啡 ☕️

支付宝收款码

支付宝

微信收款码

微信

💬 留言

评论由 Giscus 驱动(基于 GitHub Discussions)。 当前仓库 NaphJohn/LLM-blog 尚未启用 Discussions:请在 GitHub 仓库 Settings → General → Features 勾选 Discussions 后刷新本页,评论区即自动显示。