-
Notifications
You must be signed in to change notification settings - Fork 4.4k
Pull requests: NVIDIA/Megatron-LM
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
ci(auth): treat svcnemo-autobot as internal
complexity: low
#6437
opened Aug 11, 2026 by
ko3n1g
Contributor
Loading…
ci(auth): treat svcnemo-autobot as internal
complexity: low
#6436
opened Aug 11, 2026 by
ko3n1g
Contributor
Loading…
ci: AUT-1408 backport GitHub App authentication to dev
#6435
opened Aug 11, 2026 by
svcnemo-autobot
Collaborator
Loading…
cp: build: bump transformer-engine to 2.17.1
Run tests
#6433
opened Aug 11, 2026 by
svcnemo-autobot
Collaborator
•
Draft
Avoid unnecessary MoE router host synchronization
complexity: low
nemotron
#6432
opened Aug 11, 2026 by
JF-D
Contributor
Loading…
6 tasks done
Add MFSDP v2 design document
docs-only
documentation only (docs or docstrings)
Final Review
PR is in the "final review" stage
#6431
opened Aug 11, 2026 by
wujingyue
Contributor
Loading…
Add async scheduling pairwise unit coverage
#6430
opened Aug 11, 2026 by
lmcafee-nvidia
Contributor
•
Draft
Test GPTOSS SSE with async scheduling
Run functional tests
#6428
opened Aug 11, 2026 by
lmcafee-nvidia
Contributor
•
Draft
Kimi K3 VPP + cross-stage caching
community-request
#6427
opened Aug 11, 2026 by
wuweiqiang24
•
Draft
6 tasks
Add DSA over GQA with the streamed min-memory kernels
complexity: high
#6426
opened Aug 11, 2026 by
alokpathy
Loading…
3 of 6 tasks
Add multimodal tokenizer prompt formats
complexity: high
Run tests
#6424
opened Aug 10, 2026 by
matthieule
Contributor
Loading…
4 of 6 tasks
Propagate RL runtime config updates to module configs
complexity: low
#6423
opened Aug 10, 2026 by
Phlip79
Member
Loading…
1 task done
Add tensor-parallel Muon Hyperball optimizer
#6422
opened Aug 10, 2026 by
mkhona-nvidia
Contributor
•
Draft
Paged stashing no longer assumes single layer config
complexity: low
#6419
opened Aug 10, 2026 by
Phlip79
Member
Loading…
1 task done
Add prefix cache stress test foundation
#6418
opened Aug 10, 2026 by
lmcafee-nvidia
Contributor
•
Draft
Preserve greedy temperature in RL inference
#6417
opened Aug 10, 2026 by
lmcafee-nvidia
Contributor
•
Draft
2 of 6 tasks
Defer prefix-cache hit accounting until admission succeeds
#6416
opened Aug 10, 2026 by
lmcafee-nvidia
Contributor
•
Draft
2 of 6 tasks
Add opt-in CuTeDSL copy kernels for MoE paged stash
#6414
opened Aug 10, 2026 by
sraman-rgb
Contributor
•
Draft
6 tasks
Fix tool call reasoning boundary
complexity: low
#6411
opened Aug 10, 2026 by
nvcsathe
Contributor
Loading…
6 tasks
Add hybrid layer config definitions
complexity: low
Final Review
PR is in the "final review" stage
#6410
opened Aug 10, 2026 by
Phlip79
Member
Loading…
1 task done
Previous Next
ProTip!
Type g i on any issue or pull request to go back to the issue listing page.