Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
NVIDIA
/
Model-Optimizer
Public
Notifications
You must be signed in to change notification settings
Fork
663
Star
4.7k
Code
Issues
98
Pull requests
322
Actions
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Security and quality
Insights
Actions: NVIDIA/Model-Optimizer
Actions
All workflows
Workflows
Example tests
Example tests
GPU tests
GPU tests
Regression tests
Regression tests
Unit tests
Unit tests
.github/workflows/build_puzzletron.yml
.github/workflows/build_puzzletron.yml
Bump uv.lock
Bump uv.lock
Claude
Claude
Claude Code Review
Claude Code Review
Close inactive issues and PRs
Close inactive issues and PRs
Code Quality
Code Quality
Show more workflows...
Management
Caches
Deployments
All workflows
All workflows
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Showing runs from all workflows
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Workflow
Filter by Workflow
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching workflows.
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
[Bug] modelopt-mcp: all launcher subprocesses hang when the server runs as a stdio MCP server — missing stdin=DEVNULL
Claude
#23527:
Issue
#2551
opened by
lucky20260806
1s
1s
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Claude
#23526:
Pull request
#2550
submitted by
coderabbitai
Bot
Action required
Action required
View #2550
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Claude
#23525:
Pull request
#2550
created by
coderabbitai
Bot
Action required
Action required
View #2550
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Claude
#23524:
Issue comment
#2550 (comment)
created by
coderabbitai
Bot
2s
2s
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Claude Code Review
#4740:
Issue comment
#2550 (comment)
created by
coderabbitai
Bot
1s
1s
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Claude Code Review
#4739:
Issue comment
#2550 (comment)
created by
copy-pr-bot
Bot
1s
1s
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Claude
#23523:
Issue comment
#2550 (comment)
created by
copy-pr-bot
Bot
9s
9s
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Unit tests
#12387:
Pull request
#2550
opened by
Momoyeyu
Action required
Momoyeyu:fix/nvfp4-gemm-effective-bits
Momoyeyu:fix/nvfp4-gemm-effective-bits
Action required
View #2550
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Docs
#13517:
Pull request
#2550
opened by
Momoyeyu
Action required
Momoyeyu:fix/nvfp4-gemm-effective-bits
Momoyeyu:fix/nvfp4-gemm-effective-bits
Action required
View #2550
View workflow file
Fix NVFP4 real-quant GEMM availability check for effective_bits
Code Quality
#11292:
Pull request
#2550
opened by
Momoyeyu
Action required
Momoyeyu:fix/nvfp4-gemm-effective-bits
Momoyeyu:fix/nvfp4-gemm-effective-bits
Action required
View #2550
View workflow file
After applying NVFP4 quantization followed by compression, RealQuantLinear fails to locate the corresponding GEMM implementation during inference.
Claude Code Review
#4738:
Issue comment
#2330 (comment)
created by
Momoyeyu
1s
1s
View workflow file
After applying NVFP4 quantization followed by compression, RealQuantLinear fails to locate the corresponding GEMM implementation during inference.
Claude
#23522:
Issue comment
#2330 (comment)
created by
Momoyeyu
1s
1s
View workflow file
Fix NVFP4 fake quant zeroing blocks with small scales
Claude
#23521:
Pull request
#2549
created by
coderabbitai
Bot
Skipped
Skipped
View #2549
View workflow file
Fix NVFP4 fake quant zeroing blocks with small scales
Claude
#23520:
Pull request
#2549
submitted by
coderabbitai
Bot
1s
1s
View #2549
View workflow file
Use CUDA's FP8 conversion for conv NVFP4 block scales
Regression tests
#3970:
Commit
40a621d
pushed by
copy-pr-bot
Bot
29m 22s
pull-request/2549
pull-request/2549
29m 22s
View workflow file
Use CUDA's FP8 conversion for conv NVFP4 block scales
GPU tests
#8437:
Commit
40a621d
pushed by
copy-pr-bot
Bot
54m 1s
pull-request/2549
pull-request/2549
54m 1s
View workflow file
Use CUDA's FP8 conversion for conv NVFP4 block scales
Example tests
#7924:
Commit
40a621d
pushed by
copy-pr-bot
Bot
1h 5m 37s
pull-request/2549
pull-request/2549
1h 5m 37s
View workflow file
Fix NVFP4 fake quant zeroing blocks with small scales
Unit tests
#12386:
Pull request
#2549
synchronize by
sychen52
25m 19s
sychen52:fix_nvfp4_scale_floor
sychen52:fix_nvfp4_scale_floor
25m 19s
View #2549
View workflow file
Fix NVFP4 fake quant zeroing blocks with small scales
Docs
#13516:
Pull request
#2549
synchronize by
sychen52
4m 47s
sychen52:fix_nvfp4_scale_floor
sychen52:fix_nvfp4_scale_floor
4m 47s
View #2549
View workflow file
Fix NVFP4 fake quant zeroing blocks with small scales
Code Quality
#11291:
Pull request
#2549
synchronize by
sychen52
5m 22s
sychen52:fix_nvfp4_scale_floor
sychen52:fix_nvfp4_scale_floor
5m 22s
View #2549
View workflow file
Puzzletron Progress 6/8 (calculating one block scores) takes 10 to 20 times more than in tutorial
Claude
#23519:
Issue comment
#1667 (comment)
created by
tejasng053
1s
1s
View workflow file
Puzzletron Progress 6/8 (calculating one block scores) takes 10 to 20 times more than in tutorial
Claude Code Review
#4737:
Issue comment
#1667 (comment)
created by
tejasng053
2s
2s
View workflow file
Close inactive issues and PRs
Close inactive issues and PRs
#332:
Scheduled
11s
main
main
11s
View workflow file
pages build and deployment
pages-build-deployment
#5108:
by
github-pages
Bot
3m 26s
gh-pages
gh-pages
3m 26s
Docs
Docs
#13515:
Scheduled
6m 2s
main
main
6m 2s
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.