Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
NVIDIA
/
Model-Optimizer
Public
Notifications
You must be signed in to change notification settings
Fork
594
Star
3.8k
Code
Issues
93
Pull requests
286
Actions
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Security and quality
Insights
Actions: NVIDIA/Model-Optimizer
Actions
All workflows
Workflows
Example tests
Example tests
GPU tests
GPU tests
Regression tests
Regression tests
Unit tests
Unit tests
.github/workflows/build_puzzletron.yml
.github/workflows/build_puzzletron.yml
Bump uv.lock
Bump uv.lock
Claude
Claude
Claude Code Review
Claude Code Review
Close inactive issues and PRs
Close inactive issues and PRs
Code Quality
Code Quality
Show more workflows...
Management
Caches
Deployments
Code Quality
Code Quality
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
code_quality.yml
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
Fix AutoCast unknown dimension metadata
Code Quality
#10725:
Pull request
#2420
synchronize by
haoxiz-nvidia
4m 49s
haoxiz/autocast-mismatch
haoxiz/autocast-mismatch
4m 49s
View #2420
View workflow file
Fix AutoCast unknown dimension metadata
Code Quality
#10724:
Pull request
#2420
opened by
haoxiz-nvidia
4m 41s
haoxiz/autocast-mismatch
haoxiz/autocast-mismatch
4m 41s
View #2420
View workflow file
Env: update onnxruntime and cuda flag
Code Quality
#10723:
Pull request
#2419
synchronize by
haoxiz-nvidia
5m 8s
haoxiz/env-fix
haoxiz/env-fix
5m 8s
View #2419
View workflow file
Env: update onnxruntime and cuda flag
Code Quality
#10722:
Pull request
#2419
opened by
haoxiz-nvidia
4m 47s
haoxiz/env-fix
haoxiz/env-fix
4m 47s
View #2419
View workflow file
[6721556] Fix mixed dtypes in NVFP4 FP16 ONNX export
Code Quality
#10721:
Pull request
#2415
synchronize by
ajrasane
4m 33s
ajrasane/nvbug-6721556-swin-mixed-dtype-add
ajrasane/nvbug-6721556-swin-mixed-dtype-add
4m 33s
View #2415
View workflow file
[6466805] Add Dynamo ONNX export support for quantized models
Code Quality
#10720:
Pull request
#2418
opened by
ajrasane
5m 7s
ajrasane/dynamo-onnx-export
ajrasane/dynamo-onnx-export
5m 7s
View #2418
View workflow file
[OMNIML-5570] 2/2 Compose GEMM and KV-cache AutoQuant workflows
Code Quality
#10719:
Pull request
#2273
synchronize by
meenchen
4m 1s
agent/kv-autoquant-composition
agent/kv-autoquant-composition
4m 1s
View #2273
View workflow file
[5565357] Fix SDXL NVFP4 export and performance
Code Quality
#10718:
Pull request
#2336
synchronize by
ajrasane
4m 31s
arasane/fix_sdxl_nvfp4_export
arasane/fix_sdxl_nvfp4_export
4m 31s
View #2336
View workflow file
docs: add Local Hessian NVFP4 weight-scale announcement blog
Code Quality
#10717:
Pull request
#2417
synchronize by
realAsma
5m 14s
asma/local_hessian_blog
asma/local_hessian_blog
5m 14s
View #2417
View workflow file
docs: add Local Hessian NVFP4 weight-scale announcement blog
Code Quality
#10716:
Pull request
#2417
synchronize by
realAsma
41s
asma/local_hessian_blog
asma/local_hessian_blog
41s
View #2417
View workflow file
Fix distributed AutoQuantize scoring and share backward setup
Code Quality
#10715:
Pull request
#2231
synchronize by
realAsma
Action required
joshua-hill:fix/autoquant-scoring-infrastructure
joshua-hill:fix/autoquant-scoring-infrastructure
Action required
View #2231
View workflow file
[OMNIML-5899] Add IQ1_S and IQ2_XS quantization and unified checkpoint export
Code Quality
#10714:
Pull request
#2381
synchronize by
ChenhanYu
4m 13s
iq2xs-unified-export
iq2xs-unified-export
4m 13s
View #2381
View workflow file
[5565357] Fix SDXL NVFP4 export and performance
Code Quality
#10713:
Pull request
#2336
synchronize by
ajrasane
5m 16s
arasane/fix_sdxl_nvfp4_export
arasane/fix_sdxl_nvfp4_export
5m 16s
View #2336
View workflow file
PTQ reliability fixes: checkpoint resume + exclude_modules sentinel corruption
Code Quality
#10712:
Pull request
#2129
synchronize by
wyattearp
Action required
wyattearp:ptq-reliability-fixes
wyattearp:ptq-reliability-fixes
Action required
View #2129
View workflow file
docs: add Local Hessian NVFP4 weight-scale announcement blog
Code Quality
#10711:
Pull request
#2417
synchronize by
realAsma
5m 25s
asma/local_hessian_blog
asma/local_hessian_blog
5m 25s
View #2417
View workflow file
docs: add Local Hessian NVFP4 weight-scale announcement blog
Code Quality
#10710:
Pull request
#2417
synchronize by
realAsma
1m 3s
asma/local_hessian_blog
asma/local_hessian_blog
1m 3s
View #2417
View workflow file
[5565357] Fix SDXL NVFP4 export and performance
Code Quality
#10709:
Pull request
#2336
synchronize by
ajrasane
4m 11s
arasane/fix_sdxl_nvfp4_export
arasane/fix_sdxl_nvfp4_export
4m 11s
View #2336
View workflow file
[5612316][OMNIML-2983] Fix FP8 QDQ placement for delegated diffusion attention
Code Quality
#10708:
Pull request
#2416
synchronize by
ajrasane
4m 18s
ajrasane/nvbug-5612316-fp8-attention-qdq
ajrasane/nvbug-5612316-fp8-attention-qdq
4m 18s
View #2416
View workflow file
Code Quality
Code Quality
#10707:
Scheduled
4m 4s
main
main
4m 4s
View workflow file
Fix vLLM fakequant calibration for hybrid attention models
Code Quality
#10706:
Pull request
#2414
synchronize by
kinjalpatel27
4m 59s
kinjal/fix_vllm_0.28
kinjal/fix_vllm_0.28
4m 59s
View #2414
View workflow file
Fix vLLM fakequant calibration for hybrid attention models
Code Quality
#10705:
Pull request
#2414
synchronize by
kinjalpatel27
4m 2s
kinjal/fix_vllm_0.28
kinjal/fix_vllm_0.28
4m 2s
View #2414
View workflow file
[OMNIML-5570] 2/2 Compose GEMM and KV-cache AutoQuant workflows
Code Quality
#10704:
Pull request
#2273
synchronize by
meenchen
5m 27s
agent/kv-autoquant-composition
agent/kv-autoquant-composition
5m 27s
View #2273
View workflow file
docs: add Local Hessian NVFP4 weight-scale announcement blog
Code Quality
#10703:
Pull request
#2417
synchronize by
realAsma
4m 17s
asma/local_hessian_blog
asma/local_hessian_blog
4m 17s
View #2417
View workflow file
[6463897] Fix narrow FP16 histogram calibration
Code Quality
#10702:
Pull request
#2412
synchronize by
ajrasane
4m 53s
ajrasane/nvbug-6463897-histogram-range
ajrasane/nvbug-6463897-histogram-range
4m 53s
View #2412
View workflow file
[6721556] Fix mixed dtypes in NVFP4 FP16 ONNX export
Code Quality
#10701:
Pull request
#2415
synchronize by
ajrasane
3m 59s
ajrasane/nvbug-6721556-swin-mixed-dtype-add
ajrasane/nvbug-6721556-swin-mixed-dtype-add
3m 59s
View #2415
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.