COMMITS
October 22, 2025
J
[Core] Handle MoE LoRA edge cases (#27335)
Jee Jee Li committed
G
W
L
[Bugfix][CPU] Disable dual stream execution for experts on CPU (#27320)
Li, Jiang committed
W
E
fixed reasoning streaming with tool_choice="required" (#24108)
ExtReMLapin committed
H
Remove last `level` references not removed in #26355 (#27260)
Harry Mellor committed
H
Update release pipeline for PyTorch 2.9.0 (#27303)
Huy Do committed
W
[1/N][Platform] Cleanup useless function (#26982)
wangxiyuan committed
J
[torch.compile] Enable silu_mul_fp8_quant fusion without custom ops enabled (#27146)
Jiangyun Zhu committed
C
[Benchmark] Add plot utility for parameter sweep (#27168)
Cyrus Leung committed
N
[CI] Nixl integration tests DP-EP (#27199)
Nicolò Lucchesi committed
V
[DOC] [ROCm] Add ROCm quickstart guide (#26505)
vllmellm committed
October 21, 2025
L
[Deepseek v3.2] Remove extra logics in indexer (#26465)
Lain committed
T
[P/D] KVConnector for decode benchmarking (#25986)
Tyler Michael Smith committed
B
[Bugfix] skip cuda graph for drafter when running with eager (#26821)
Benjamin Chislett committed
E
Updated xgrammar backend to not deny supported string formats (#27253)
ExtReMLapin committed
A
[Performance] Dual stream execution of "shared_experts" and "selected_experts" inside FusedMoE (#26440)
Alexander Matveev committed
H
Update PyTorch to 2.9.0+cu129 (#24994)
Huy Do committed
T
N
[V0 Deprecation] Remove V0 executors (#27142)
Nick Hill committed
D
[Bugfix][P/D] Reduce num_threads used by nixl ucx backend (#27196)
David Whyte-Gray committed
W
[Feature] Batch Invariant for R1 TP 8 on Blackwell (#27229)
Wentao Ye committed
M
[ROCm] Update Triton, Torch, and AITER branches for ROCm base Dockerfile (#27206)
Micah Williamson committed
P
Add @pavanimajety to .github/codeowners for Flashinfer, ModelOpt related code (#27213)
Pavani Majety committed
J
[ROCM] Enable CompressedTensorsWNA16 (#27187)
JartX committed
H
[CI] Install pre-release version of `apache-tvm-ffi` for `flashinfer` (#27262)
Harry Mellor committed
D
[Chore] Separate out NCCL utilities from vllm.utils (#27197)
dongbo910220 committed
D
[Deepseek v3.2] Optimize top_k_per_row (#26763)
Daniel Cámpora committed
R
[MM][Core] Decouple ViT backend from LM backend (#27061)
Roger Wang committed