Shipped · AI and LLM tooling

What vllm-project/vllm shipped

The public repository of vLLM · github.com/vllm-project/vllm

Written by FoxPlug from public releases; not affiliated with vLLM. An automatic summary of the public release, pull request and commit data of github.com/vllm-project/vllm. vLLM did not write it and does not use or endorse FoxPlug. Every line links to the public change it describes.

Get a weekly update like this for your product, free

Or Use it as a GitHub Action

Follow vllm's weekly shipped digest

Week of September 21, 2026

What shipped

Changelog entry

Example posts FoxPlug drafted from these changes. Not written or posted by the project.

Post for X

vLLM updates: MoE routing fusion, DSv4.1 kernel optimization, FlashKDA fp32 state preservation, expert mapping indexing, sharding-aware weight transfer, and multimodal security hardening.

Post for LinkedIn

vLLM's latest updates focus on performance and reliability: fused MoE routing for non-unit scaling factors, optimized DSv4.1 O-projection kernels, fixed accumulated rounding errors in long prefills, indexed expert mapping to fix quadratic scaling, enabled mixed DCP/non-DCP pipelines, and hardened multimodal processor security. CI improvements reduce overhead on AMD systems.

Weeks with too little public activity are left out rather than filled in. Last updated 2026-09-30.

Is this your repo? Ask us to remove this page.