COMMITS
August 26, 2026
T
[Agents] Add CUDA IMA debugging skill (#53702)
Thien Tran committed
August 25, 2026
C
[Kimi K3][Kernel] Enable low-latency decode GEMM dispatch on SM100 (#53534)
Canlin Guo committed
A
[Bugfix][KV Offload] Defer request-level cascade of in-flight primary keys (#53329)
Almog Tavor committed
D
W
W
N
[Bugfix][MRV2] Run cudagraph memory profiling in a throwaway graph pool (#53682)
Nick Hill committed
M
[CI] Preserve Rust Docker cache across commits (#53290)
Misha Goin committed
M
[Bugfix][Multimodal] Reject malformed base64 audio with 400 instead of 500 (#53744)
Moe Huzaifa committed
M
[Spec Decode] Enable adaptive DSpark on SM100 sparse MLA (#52783)
Misha Goin committed
W
[Mypy Fix] Mypy fix for "vllm/model_executor/models/[gG]" (#53616)
Wentao Ye committed
M
[Bugfix][Frontend] Apply the stop string limit to Cohere requests (#53750)
Moe Huzaifa committed
M
[Bugfix][Frontend] Keep credentials out of the Rust frontend launch log (#53738)
Moe Huzaifa committed
M
[Bugfix][Tokenizer] Replace bare asserts in the DeepSeek V4 encoder (#53747)
Moe Huzaifa committed
L
[Bugfix] Resolve B12X modules before Dynamo tracing (#53326)
Luke Alonso committed
C
[CI] Add GSM8K accuracy test for amd/DeepSeek-V4-Flash-MXFP4 (#50632)
Colin Z committed
Y
[XPU] update key supported models (#53494)
Yan Ma committed
W
N
[Feature][DSpark]: Logprobs adaptive verification (#52242)
Naveenraj Kamalakannan committed
E
[Quantization][Humming] Support MXFP4 weight + block-FP8 activation for MoE (#51332)
Elvir Crnčević committed
T
[Bugfix] Handle parenthesized Gemma4 tool calls (#53657)
Taneem Ibrahim committed
T
[Pooling UX] Improve serve --task error guidance (#53467)
Taneem Ibrahim committed
E
[Profiler] Fix start_profile permanently no-op after max_iterations auto-stop (#51839)
Elvir Crnčević committed
D
T
[Config][EC] Normalize producer-only encoder config (#53656)
Tianyu Guo committed
J
[Bugfix][DeepSeek V4] Handle trailing system messages in prompt rendering (#51262)
Jiahao Liang committed
N
[sleep functionality] code refactor about sleep/wake_up (#50431)
Ning Xie committed
T
[Agents] Add kernel microbenchmark skill (#53688)
Thien Tran committed
J
[Bugfix][MM] Fix JinaVL processing cache order (#53553)
Jiatai Wang committed