Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
NVIDIA
Model-Optimizer
Repository navigation
Code
Issues
101
(101)
Pull requests
365
(365)
Actions
Security and quality
Insights
More
items
Actions: NVIDIA/Model-Optimizer
Actions
All workflows
Workflows
Example tests
Example tests
GPU tests
GPU tests
Regression tests
Regression tests
Unit tests
Unit tests
.github/workflows/build_puzzletron.yml
.github/workflows/build_puzzletron.yml
Bump uv.lock
Bump uv.lock
Claude
Claude
Claude Code Review
Claude Code Review
Close inactive issues and PRs
Close inactive issues and PRs
Code Quality
Code Quality
Show more workflows...
Management
Caches
Deployments
GPU tests
GPU tests
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
gpu_tests.yml
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
GPU tests
GPU tests
#9032:
Scheduled
In progress
main
main
In progress
View workflow file
Address remote benchmark review feedback
GPU tests
#9031:
Commit
68bd0ad
pushed by
copy-pr-bot
Bot
1h 0m 37s
pull-request/2467
pull-request/2467
1h 0m 37s
View workflow file
Give the unit tests 120 s per test on every platform
GPU tests
#9030:
Commit
524d784
pushed by
copy-pr-bot
Bot
38m 40s
pull-request/2730
pull-request/2730
38m 40s
View workflow file
test(speculative): run the DFlash2 overfit and DSpark compile checks …
GPU tests
#9029:
Commit
614b66a
pushed by
copy-pr-bot
Bot
1h 14m 46s
pull-request/2279
pull-request/2279
1h 14m 46s
View workflow file
Run the torch_onnx lane's torch_quant_to_onnx.py steps in one worker
GPU tests
#9028:
Commit
5064ef3
pushed by
copy-pr-bot
Bot
42m 48s
pull-request/2730
pull-request/2730
42m 48s
View workflow file
Resolve remote timing server executable path
GPU tests
#9027:
Commit
47cf2bc
pushed by
copy-pr-bot
Bot
1h 2m 55s
pull-request/2467
pull-request/2467
1h 2m 55s
View workflow file
Fix linear attention CI dependency and Megatron compatibility
GPU tests
#9026:
Commit
140735b
pushed by
copy-pr-bot
Bot
59m 42s
pull-request/2519
pull-request/2519
59m 42s
View workflow file
Fix linear-attention state QAT phase handling and serving profiles
GPU tests
#9025:
Commit
fec406f
pushed by
copy-pr-bot
Bot
52m 29s
pull-request/2519
pull-request/2519
52m 29s
View workflow file
GPU tests
GPU tests
#9024:
Scheduled
1h 1m 26s
main
main
1h 1m 26s
View workflow file
Run the diffusers example scripts in the pytest process
GPU tests
#9023:
Commit
fc1b344
pushed by
copy-pr-bot
Bot
59m 29s
pull-request/2728
pull-request/2728
59m 29s
View workflow file
test: require source config byte preservation
GPU tests
#9022:
Commit
d3207e3
pushed by
copy-pr-bot
Bot
13m 20s
pull-request/2630
pull-request/2630
13m 20s
View workflow file
Resolve remote timing server executable path
GPU tests
#9021:
Commit
8138a27
pushed by
copy-pr-bot
Bot
1h 10m 59s
pull-request/2467
pull-request/2467
1h 10m 59s
View workflow file
Document automatic vLLM fakequant reload
GPU tests
#9020:
Commit
aa5df15
pushed by
copy-pr-bot
Bot
52m 49s
pull-request/2629
pull-request/2629
52m 49s
View workflow file
[3/5] Fold grouped expert weights during fakequant export
GPU tests
#9019:
Commit
41b1e68
pushed by
copy-pr-bot
Bot
1h 2m 43s
pull-request/2628
pull-request/2628
1h 2m 43s
View workflow file
docs: add a CHANGELOG entry for the NemotronH aggressive-MSE NVFP4 re…
GPU tests
#9018:
Commit
3e9b36c
pushed by
copy-pr-bot
Bot
1h 3m 15s
pull-request/2703
pull-request/2703
1h 3m 15s
View workflow file
test: release vLLM engines between dynamic module tests
GPU tests
#9017:
Commit
2e4b911
pushed by
copy-pr-bot
Bot
12m 53s
pull-request/2627
pull-request/2627
12m 53s
View workflow file
test: require source config byte preservation
GPU tests
#9016:
Commit
01f6019
pushed by
copy-pr-bot
Bot
49m 55s
pull-request/2630
pull-request/2630
49m 55s
View workflow file
Document automatic vLLM fakequant reload
GPU tests
#9015:
Commit
aa8282c
pushed by
copy-pr-bot
Bot
45m 59s
pull-request/2629
pull-request/2629
45m 59s
View workflow file
[3/5] Fold grouped expert weights during fakequant export
GPU tests
#9014:
Commit
752c114
pushed by
copy-pr-bot
Bot
47m 58s
pull-request/2628
pull-request/2628
47m 58s
View workflow file
Merge origin/main into real-quant weight export
GPU tests
#9013:
Commit
699ddb6
pushed by
copy-pr-bot
Bot
42m 39s
pull-request/2251
pull-request/2251
42m 39s
View workflow file
test: preserve Qwen fixture for upstream MoE communication checks
GPU tests
#9012:
Commit
3e060dd
pushed by
copy-pr-bot
Bot
51m 14s
pull-request/2627
pull-request/2627
51m 14s
View workflow file
Fix the GLM-5.1 alias: it is not the input_scale1 recipe
GPU tests
#9011:
Commit
f442120
pushed by
copy-pr-bot
Bot
53m 11s
pull-request/2711
pull-request/2711
53m 11s
View workflow file
Document automatic vLLM fakequant reload
GPU tests
#9010:
Commit
f055084
pushed by
copy-pr-bot
Bot
35m 15s
pull-request/2629
pull-request/2629
35m 15s
View workflow file
Merge branch 'main' into shengliangx/preprocess-spawn-fix
GPU tests
#9009:
Commit
b8a78dd
pushed by
copy-pr-bot
Bot
1h 15m 50s
pull-request/2726
pull-request/2726
1h 15m 50s
View workflow file
Speed up the hf_ptq example tests
GPU tests
#9008:
Commit
8b96222
pushed by
copy-pr-bot
Bot
1h 1m 26s
pull-request/2727
pull-request/2727
1h 1m 26s
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.