Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
deepspeedai
/
DeepSpeed
Public
Notifications
You must be signed in to change notification settings
Fork
5k
Star
43.1k
Code
Issues
1.2k
Pull requests
214
Discussions
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Security and quality
Insights
Actions: deepspeedai/DeepSpeed
Actions
All workflows
Workflows
cpu-torch-latest
cpu-torch-latest
aws-torch-latest-full
aws-torch-latest-full
Build and publish DeepSpeed release
Build and publish DeepSpeed release
Copilot
Copilot
Copilot cloud agent
Copilot cloud agent
Copilot code review
Copilot code review
DCO / required
DCO / required
Dependabot Updates
Dependabot Updates
Dependency Graph
Dependency Graph
Formatting
Formatting
Show more workflows...
Management
Caches
Deployments
cpu-torch-latest
cpu-torch-latest
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
cpu-torch-latest.yml
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
Give Muon's momentum the dtype of the gradient it is combined with
cpu-torch-latest
#9671:
Pull request
#8483
opened by
alanhuangyoo
Action required
alanhuangyoo:fix/muon-momentum-dtype-both-paths
alanhuangyoo:fix/muon-momentum-dtype-both-paths
Action required
View #8483
View workflow file
Fixed memory leak in loss backward with AC & zero3 & single-model-multi-branch network
cpu-torch-latest
#9670:
Pull request
#8482
synchronize by
supermeng
Action required
supermeng:fix/zero3_ac_mul_branck_mem_leak
supermeng:fix/zero3_ac_mul_branck_mem_leak
Action required
View #8482
View workflow file
Fixed memory leak in loss backward with AC & zero3 & single-model-multi-branch network
cpu-torch-latest
#9669:
Pull request
#8482
opened by
supermeng
Action required
supermeng:fix/zero3_ac_mul_branck_mem_leak
supermeng:fix/zero3_ac_mul_branck_mem_leak
Action required
View #8482
View workflow file
cpu-torch-latest
cpu-torch-latest
#9668:
Merge group checks requested
38m 56s
38m 56s
View workflow file
[muon] Reconcile the momentum dtype when a checkpoint is restored
cpu-torch-latest
#9667:
Pull request
#8433
synchronize by
delock
40m 31s
alanhuangyoo:fix/muon-momentum-dtype-on-resume
alanhuangyoo:fix/muon-momentum-dtype-on-resume
40m 31s
View #8433
View workflow file
cpu-torch-latest
cpu-torch-latest
#9666:
Merge group checks requested
34m 2s
34m 2s
View workflow file
cpu-torch-latest
cpu-torch-latest
#9665:
Merge group checks requested
33m 15s
33m 15s
View workflow file
cpu-torch-latest
cpu-torch-latest
#9664:
Merge group checks requested
38m 48s
38m 48s
View workflow file
fix(inference): reject non-positive max_out_tokens at config validation
cpu-torch-latest
#9663:
Pull request
#8454
synchronize by
chakshu-dhannawat
Action required
chakshu-dhannawat:fix/max-out-tokens-positive-validation
chakshu-dhannawat:fix/max-out-tokens-positive-validation
Action required
View #8454
View workflow file
Fix ZeRO parameter alignment for grouped_mm
cpu-torch-latest
#9662:
Pull request
#8277
synchronize by
tohtana
38m 24s
fwerkor:fix-8276-grouped-mm-alignment
fwerkor:fix-8276-grouped-mm-alignment
38m 24s
View #8277
View workflow file
Count each expert once in the fp32 gradient-clipping norm
cpu-torch-latest
#9661:
Pull request
#8478
synchronize by
alanhuangyoo
Action required
alanhuangyoo:fix/fp32-clip-moe-norm
alanhuangyoo:fix/fp32-clip-moe-norm
Action required
View #8478
View workflow file
cpu-torch-latest
cpu-torch-latest
#9660:
Merge group checks requested
38m 44s
38m 44s
View workflow file
cpu-torch-latest
cpu-torch-latest
#9659:
Merge group checks requested
37m 27s
37m 27s
View workflow file
cpu-torch-latest
cpu-torch-latest
#9658:
Merge group checks requested
38m 4s
38m 4s
View workflow file
cpu-torch-latest
cpu-torch-latest
#9657:
Merge group checks requested
35m 58s
35m 58s
View workflow file
feat(rollout): add continuous batching generation prototype
cpu-torch-latest
#9656:
Pull request
#8368
synchronize by
delock
31m 21s
nathon-lee:perf/opsd-continuous-batching-prototype
nathon-lee:perf/opsd-continuous-batching-prototype
31m 21s
View #8368
View workflow file
fix(inference): break circular import in ops.transformer.inference
cpu-torch-latest
#9655:
Pull request
#8465
synchronize by
chakshu-dhannawat
Action required
chakshu-dhannawat:fix/inference-circular-import-7159
chakshu-dhannawat:fix/inference-circular-import-7159
Action required
View #8465
View workflow file
cpu-torch-latest
cpu-torch-latest
#9654:
Scheduled
38m 1s
master
master
38m 1s
View workflow file
Carry the affine scale on the replicated map, not the split
cpu-torch-latest
#9653:
Pull request
#8477
synchronize by
delock
37m 22s
Achyuthan-S:affine-scale-on-replicated
Achyuthan-S:affine-scale-on-replicated
37m 22s
View #8477
View workflow file
cpu-torch-latest
cpu-torch-latest
#9652:
Merge group checks requested
37m 4s
37m 4s
View workflow file
Fix comms logger KeyError when log_name is omitted
cpu-torch-latest
#9651:
Pull request
#8267
synchronize by
tohtana
39m 1s
jinyouzhi:comms_logger
jinyouzhi:comms_logger
39m 1s
View #8267
View workflow file
Recognize a norm by its shape, not by whether its class name was listed
cpu-torch-latest
#9650:
Pull request
#8479
opened by
alanhuangyoo
Action required
alanhuangyoo:fix/autotp-norm-by-shape
alanhuangyoo:fix/autotp-norm-by-shape
Action required
View #8479
View workflow file
Count each expert once in the fp32 gradient-clipping norm
cpu-torch-latest
#9649:
Pull request
#8478
opened by
alanhuangyoo
Action required
alanhuangyoo:fix/fp32-clip-moe-norm
alanhuangyoo:fix/fp32-clip-moe-norm
Action required
View #8478
View workflow file
Fix AutoEP global L2 clipping for sharded expert ownership
cpu-torch-latest
#9648:
Pull request
#8476
synchronize by
gss10282023
Action required
gss10282023:fix-autoep-ownership-aware-clipping
gss10282023:fix-autoep-ownership-aware-clipping
Action required
View #8476
View workflow file
Fix Muon optimizer under ZeRO CPU offload and bound gather buffers
cpu-torch-latest
#9647:
Pull request
#8464
synchronize by
jinyouzhi
Action required
jinyouzhi:muon-cpu-offload-fix
jinyouzhi:muon-cpu-offload-fix
Action required
View #8464
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.