Skip to content

Pull requests: NVIDIA-NeMo/Automodel

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

fix(peft): support mixed-dtype memory-efficient LoRA backward r0.6.0 Auto-cherrypick to release branch. Apply before merge; cherrypick happens after merge.
#3675 opened Aug 25, 2026 by akoumpa Contributor Loading…
3 tasks done
feat(dflash): add fused CE and total anchor budget
#3673 opened Aug 25, 2026 by Slyne Contributor Draft
docs(examples): add Qwen3-32B SWE-bench Verified eval (Phase 3) docs-only With great power comes great responsibility.
#3671 opened Aug 25, 2026 by athitten Contributor Loading…
feat(recipes): add model-ready hooks
#3661 opened Aug 25, 2026 by pstjohn Loading…
Enable activation checkpointing for repeated dense MTP blocks
#3660 opened Aug 25, 2026 by pzelasko Contributor Loading…
fix(test): unblock the Kimi-Linear vanilla-HF parity reference
#3659 opened Aug 25, 2026 by yuhezhang-ai Contributor Loading…
fix(nemotron-v3): align latent projection input dtype r0.6.0 Auto-cherrypick to release branch. Apply before merge; cherrypick happens after merge.
#3657 opened Aug 25, 2026 by HuiyingLi Contributor Draft
3 tasks done
feat(laguna): support packed THD context parallelism
#3640 opened Aug 24, 2026 by akoumpa Contributor Loading…
3 tasks done
fix: Canonicalize LoRA compute dtype community-request
#3638 opened Aug 24, 2026 by benthecarman Loading…
3 tasks done
fix(glm): align and diagnose cross-framework router parity
#3635 opened Aug 22, 2026 by yuhezhang-ai Contributor 3/3 Draft
fix(kimi_k25): honour the quantization flag community-request waiting-on-customer Waiting on the original author to respond
#3633 opened Aug 22, 2026 by akx Contributor Loading…
3 tasks done
ci: Update transformers to latest version 5.15.1
#3628 opened Aug 22, 2026 by svcnvidia-nemo-ci Contributor Loading…
feat(dflash): support top-p and top-k in speculative decoding community-request waiting-on-customer Waiting on the original author to respond
#3627 opened Aug 22, 2026 by kashif Contributor Loading…
3 tasks done
fix(moe): export fused expert lora in the corrected peft >= 0.19.1 layout community-request waiting-on-customer Waiting on the original author to respond
#3626 opened Aug 22, 2026 by stanley1208 Contributor Loading…
ProTip! Type g p on any issue or pull request to go back to the pull request listing page.