Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
NVIDIA
/
Megatron-LM
Public
Notifications
You must be signed in to change notification settings
Fork
4.4k
Star
17.7k
Code
Issues
400
Pull requests
865
Discussions
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Security and quality
Insights
Actions: NVIDIA/Megatron-LM
Actions
All workflows
Workflows
CICD Megatron-LM
CICD Megatron-LM
Trigger MBridge Tests
Trigger MBridge Tests
.github/workflows/_build_test_publish_wheel.yml
.github/workflows/_build_test_publish_wheel.yml
API Compatibility Check
API Compatibility Check
Approve Test Queue
Approve Test Queue
Auto Assign Review Labels
Auto Assign Review Labels
Auto Reminder Bot
Auto Reminder Bot
Auto Swap Labels
Auto Swap Labels
Auto Update Copy PR Bot
Auto Update Copy PR Bot
Build docs
Build docs
Show more workflows...
Management
Caches
Deployments
Auto Swap Labels
Auto Swap Labels
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
auto-swap-labels.yml
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
Refactor hybrid layers to use per-layer configs
Auto Swap Labels
#27665:
Pull request
#6313
synchronize by
Phlip79
22s
22s
View #6313
View workflow file
test: add GDP hybrid dynamic-inference functional tests (GB200 + H100)
Auto Swap Labels
#27664:
Pull request
#6526
synchronize by
shanmugamr1992
21s
21s
View #6526
View workflow file
Auto Swap Labels
Auto Swap Labels
#27663:
completed by
prajwal1210
8s
8s
View workflow file
fix: synchronize grouped MLP params before norm
Auto Swap Labels
#27662:
Pull request
#6654
synchronize by
nvegesna-netizen
24s
24s
View #6654
View workflow file
Add configurable server evaluation defaults
Auto Swap Labels
#27661:
Pull request
#6195
synchronize by
nvcsathe
25s
25s
View #6195
View workflow file
Fix full-iteration CUDA graph capture for PP and grouped hybrid models
Auto Swap Labels
#27660:
Pull request
#6958
synchronize by
Connor-XY
2s
2s
View #6958
View workflow file
Integrate sequence-packing scheduler into training loop and wire varlen dataset into GPT pretraining
Auto Swap Labels
#27659:
Pull request
#6742
synchronize by
ilml
27s
27s
View #6742
View workflow file
Shortcut MoE
Auto Swap Labels
#27658:
Pull request
#6541
synchronize by
jiemingz
9s
9s
View #6541
View workflow file
Abort inference requests when the HTTP client disconnects
Auto Swap Labels
#27657:
Pull request
#6916
synchronize by
santhnm2
18s
18s
View #6916
View workflow file
Shortcut MoE
Auto Swap Labels
#27656:
Pull request
#6541
synchronize by
jiemingz
1s
1s
View #6541
View workflow file
Abort inference requests when the HTTP client disconnects
Auto Swap Labels
#27655:
Pull request
#6916
synchronize by
santhnm2
24s
24s
View #6916
View workflow file
[GTP] Symmetric memory registration to use custom VMM allocator
Auto Swap Labels
#27654:
Pull request
#6956
synchronize by
prajwal1210
19s
19s
View #6956
View workflow file
[GTP] Symmetric memory registration to use custom VMM allocator
Auto Swap Labels
#27653:
Pull request
#6956
synchronize by
prajwal1210
26s
26s
View #6956
View workflow file
Auto Swap Labels
Auto Swap Labels
#27652:
completed by
jaredcasper
24s
24s
View workflow file
Auto Swap Labels
Auto Swap Labels
#27651:
completed by
prajwal1210
1s
1s
View workflow file
Auto Swap Labels
Auto Swap Labels
#27650:
completed by
prajwal1210
10s
10s
View workflow file
feat(inference): add HYBRID cuda-graph sizing and make it the default
Auto Swap Labels
#27649:
Pull request
#6667
synchronize by
santhnm2
30s
30s
View #6667
View workflow file
Auto Swap Labels
Auto Swap Labels
#27648:
completed by
yaoyu-33
21s
21s
View workflow file
fix(optimizer): support MFSDP v2 gradient clipping across dense and expert meshes
Auto Swap Labels
#27647:
Pull request
#6489
synchronize by
wujingyue
21s
21s
View #6489
View workflow file
ci: Testmon integration for selective unit testing
Auto Swap Labels
#27646:
Pull request
#6934
synchronize by
balasaajay
7s
7s
View #6934
View workflow file
Auto Swap Labels
Auto Swap Labels
#27645:
completed by
ericharper
24s
24s
View workflow file
Auto Swap Labels
Auto Swap Labels
#27644:
completed by
shanmugamr1992
21s
21s
View workflow file
Exclude cuBLAS workspaces from the MFSDP peak-memory bound
Auto Swap Labels
#27643:
Pull request
#6936
ready_for_review by
wujingyue
19s
19s
View #6936
View workflow file
Exclude cuBLAS workspaces from the MFSDP peak-memory bound
Auto Swap Labels
#27642:
Pull request
#6936
synchronize by
wujingyue
1s
1s
View #6936
View workflow file
Exclude cuBLAS workspaces from the MFSDP peak-memory bound
Auto Swap Labels
#27641:
Pull request
#6936
synchronize by
wujingyue
1s
1s
View #6936
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.