Skip to content

Pull requests: NVIDIA/Megatron-LM

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

chore(ci): bump community workflow to v1.8.8 Approved All necessary approvals have been made complexity: low
#6322 opened Aug 6, 2026 by ko3n1g Contributor Loading…
[Dev] Skip broadcast for empty cu_seqlens tensors
#6320 opened Aug 6, 2026 by lhb8125 Contributor Draft
[Dev] Preserve explicit NCCL NVLS configuration
#6319 opened Aug 6, 2026 by lhb8125 Contributor Draft
Refactor hybrid layers to use per-layer configs
#6313 opened Aug 6, 2026 by Phlip79 Member Draft
6 tasks
Add Nemotron 3.5 Lightning GB200 nightly test complexity: low
#6312 opened Aug 6, 2026 by Phlip79 Member Loading…
1 task done
Add HybridModel Python recipe API and launcher support
#6310 opened Aug 6, 2026 by Phlip79 Member Draft
3 of 5 tasks
Support reentrant MFSDP activation recomputation
#6309 opened Aug 6, 2026 by wujingyue Contributor Draft
[DEV] Support energon dataloader in Qwen35 community-request
#6308 opened Aug 6, 2026 by VVsssssk Loading…
6 tasks
tests: Update cerno.yml with linear issue status Approved All necessary approvals have been made complexity: low
#6304 opened Aug 5, 2026 by balasaajay Contributor Loading…
1 of 6 tasks
Require an explicit process-group collection in LanguageModule
#6303 opened Aug 5, 2026 by Connor-XY Contributor Draft
5 of 6 tasks
Support direct hybrid MTP layer specs
#6298 opened Aug 5, 2026 by Phlip79 Member Draft
Add direct hybrid architecture descriptors
#6295 opened Aug 5, 2026 by Phlip79 Member Draft
Remove global process-group reads from megatron/core (207 -> 155)
#6293 opened Aug 5, 2026 by Connor-XY Contributor Draft
5 of 6 tasks
ProTip! Add no:assignee to see everything that’s not assigned.