Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
NVIDIA
/
TransformerEngine
Public
Notifications
You must be signed in to change notification settings
Fork
825
Star
3.5k
Code
Issues
162
Pull requests
192
Discussions
Actions
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Security and quality
Insights
Actions: NVIDIA/TransformerEngine
Actions
All workflows
Workflows
Attach wheels to release
Attach wheels to release
Blossom-CI
Blossom-CI
Copilot
Copilot
Copilot code review
Copilot code review
Dependabot Updates
Dependabot Updates
Dependency Graph
Dependency Graph
Deploy nightly docs
Deploy nightly docs
Documentation
Documentation
Label community contributions
Label community contributions
License
License
Show more workflows...
Management
Caches
Deployments
Documentation
Documentation
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
docs.yml
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20539:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 21s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 21s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20538:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 22s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 22s
View #3527
View workflow file
Fix THD P2P pad detection and tail-zero gating
Documentation
#20537:
Pull request
#3506
synchronize by
RPalmr
Action required
RPalmr:fix/thd-p2p-pad-detect
RPalmr:fix/thd-p2p-pad-detect
Action required
View #3506
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20536:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 48s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 48s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20535:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 55s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 55s
View #3527
View workflow file
[Common] Experimental CuTeDSL MXFP8 backends in C++ via TVM-FFI
Documentation
#20534:
Pull request
#3137
synchronize by
pre-commit-ci
Bot
1m 25s
kainzhong:cutedsl_mxfp8_common
kainzhong:cutedsl_mxfp8_common
1m 25s
View #3137
View workflow file
[Common] Experimental CuTeDSL MXFP8 backends in C++ via TVM-FFI
Documentation
#20533:
Pull request
#3137
synchronize by
kainzhong
1m 12s
kainzhong:cutedsl_mxfp8_common
kainzhong:cutedsl_mxfp8_common
1m 12s
View #3137
View workflow file
[PyTorch] Remove stale xfail for selective attention recompute
Documentation
#20532:
Pull request
#3531
opened by
negvet
1m 31s
negvet:fix_attn_ci_hybrid
negvet:fix_attn_ci_hybrid
1m 31s
View #3531
View workflow file
[PyTorch] GDN2 support and linear attention refactor
Documentation
#20531:
Pull request
#3521
synchronize by
ksivaman
1m 25s
ksivaman:gdn2_attention
ksivaman:gdn2_attention
1m 25s
View #3521
View workflow file
[PyTorch] GDN2 support and linear attention refactor
Documentation
#20530:
Pull request
#3521
synchronize by
ksivaman
1m 30s
ksivaman:gdn2_attention
ksivaman:gdn2_attention
1m 30s
View #3521
View workflow file
[JAX] WAR for XLA associative scan rewriter crash
Documentation
#20529:
Pull request
#3525
synchronize by
KshitijLakhani
1m 35s
KshitijLakhani:klakhani/fix/associative-scan-rewriter-war
KshitijLakhani:klakhani/fix/associative-scan-rewriter-war
1m 35s
View #3525
View workflow file
[JAX] Remove duplicate score-mod test run
Documentation
#20528:
Pull request
#3524
synchronize by
KshitijLakhani
1m 36s
KshitijLakhani:klakhani/fix/duplicate-score-mod-tests-in-ci
KshitijLakhani:klakhani/fix/duplicate-score-mod-tests-in-ci
1m 36s
View #3524
View workflow file
[All] Bump minimum supported cuDNN version to 9.12
Documentation
#20527:
Pull request
#3236
reopened by
cyanguwa
1m 12s
cyanguwa:update-min-cudnn-9.11
cyanguwa:update-min-cudnn-9.11
1m 12s
View #3236
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20526:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 31s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 31s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20525:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 36s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 36s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20524:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 23s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 23s
View #3527
View workflow file
[PyTorch] Extend no-load-balance CP to the a2a comm type
Documentation
#20523:
Pull request
#3530
opened by
Rudin6
Action required
Rudin6:cp-a2a-no-load-balance
Rudin6:cp-a2a-no-load-balance
Action required
View #3530
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20522:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 26s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 26s
View #3527
View workflow file
[Common] Add non-TMA MXFP8 cast-only kernels for specialized rowwise-only and row+colwise
Documentation
#20521:
Pull request
#3459
synchronize by
tdophung
1m 16s
tdophung:mxfp8-register-cast
tdophung:mxfp8-register-cast
1m 16s
View #3459
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20520:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 23s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 23s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20519:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 30s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 30s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20518:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 21s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 21s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20517:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 18s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 18s
View #3527
View workflow file
feat(attention): cuDNN FROST attention backend for head_dim in (256, 512], with context parallelism
Documentation
#20516:
Pull request
#3527
synchronize by
nvegesna-netizen
1m 46s
nvegesna-netizen:nvegesna/te-frost-d512-cp
nvegesna-netizen:nvegesna/te-frost-d512-cp
1m 46s
View #3527
View workflow file
[PyTorch] Allow F16 GDN inside FP8 autocast
Documentation
#20515:
Pull request
#3529
opened by
layalir
1m 28s
layalir:layali/gdn-f16-recipe-override
layalir:layali/gdn-f16-recipe-override
1m 28s
View #3529
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.