COMMITS
August 14, 2026
Q
fix(retrieval): support canonical Sentence Transformers metadata (#3546)
Qiaochu Zhu US committed
H
docs(models): rename Qwen3.8-Max to Qwen3.8-2.4T-A95B (#3549)
Huiying committed
S
feat(loss): return per-depth MTP losses (#3525)
Slyne Deng committed
August 13, 2026
S
ci: AUT-1501 bump SSO preflight to v1.8.10 (#3537)
svcnemo-autobot committed
August 12, 2026
H
D
fix(deps): resolve Starlette and GitPython CVEs (#3523)
Dong Hyuk Chang committed
H
feat(moe): integrate Mixture-of-Kittens backend (#3422)
Huiying committed
A
feat(examples): add Qwen3-32B Tulu-3 finetune recipes (#3104)
Abhishree Thittenamane committed
A
fix(test): use the recipe's attention backend for the HF reload parity check (#3500)
Abhishree Thittenamane committed
Z
feat(models): add cohere micro vision (#3518)
zeryx committed
H
docs(models): add Qwen3.8-Max coverage (#3517)
Huiying committed
J
fix(checkpoint): preserve FSDP2 mixed precision during recompute (#3513)
jQizhang committed
H
H
fix(deepseek_v4): avoid TileLang boolx8 backward codegen (#3467)
Huiying committed
A
ci(convergence): fix gemma4 eval setup and re-baseline Qwen3-MoE (#3493)
Abhishree Thittenamane committed
A
fix(checkpoint): validate PEFT adapter-only state (#3501)
Alexandros Koumparoulis committed
J
fix(models): preserve lm-head dtype boundaries (#3491)
jQizhang committed
August 11, 2026
A
docs: add Nemotron 3.5 Lightning coverage (#3499)
Alexandros Koumparoulis committed
A
perf(recipes): run GPT-OSS 120B with EP64 and no activation checkpointing (#3497)
Alexandros Koumparoulis committed
S
ci: AUT-1408 replace cicd PAT with GitHub App tokens (#3463)
svcnemo-autobot committed
A
perf(benchmarks): recompute deterministic MoE routers under AC (#3474)
Alexandros Koumparoulis committed
H
fix(recipe): shard GPT-OSS 120B across 64 experts (#3483)
Huiying committed
C
chore: Bump gitpython to >= 3.1.59 (#3482)
Charlie Truong committed
P
fix(ci): register flux2/wan2.2/qwen-image-edit diffusion recipes in CI (#3485)
Pranav Thombre committed
K
K
feat(speculative): add Kimi K3 EAGLE-3 training (#3286)
khazzz1c committed
S
H
fix(datasets): reopen IndexedDataset readers in spawned workers (#3382)
Huiying committed
H
feat(models): add MuseGlimmer training support (#3476)
Huiying committed
H
feat(vlm): support packed THD vlm context parallelism (#3322)
Huiying committed