-
Notifications
You must be signed in to change notification settings - Fork 508
Pull requests: NVIDIA/Model-Optimizer
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Add Dockerfile to examples/puzzletron
puzzletron_v2
#2009
opened Jul 23, 2026 by
grzegorz-k-karch
Contributor
Loading…
Add sidecar GPU/CPU memory+utilization monitor for HF PTQ
#2000
opened Jul 21, 2026 by
Fridah-nv
Contributor
Loading…
Speed up compressed-tensors load-time matching (for Kimi models)
#1999
opened Jul 21, 2026 by
rohansjoshi
Contributor
Loading…
feat(rocm): Add AMD ROCm/MI300X support — FP8 hipBLASLt, MIGraphX backend, AMD quantization configs
#1990
opened Jul 17, 2026 by
zhihuidu-amd
Loading…
[6425069][ONNX][Autocast] Fix autocast metadata propagation
#1983
opened Jul 16, 2026 by
gcunhase
Contributor
Loading…
[5726458] Add NVFP4 projection-output-quantizer recipe and HF embedding ONNX export example
#1981
opened Jul 16, 2026 by
ajrasane
Contributor
Loading…
Scripts and a skill to do per-layer benchmark using flashinfer
#1980
opened Jul 16, 2026 by
sychen52
Contributor
Loading…
Quality: Insecure subprocess usage in get_system_info.py
#1977
opened Jul 15, 2026 by
tomaioo
Loading…
Previous Next
ProTip!
Type g p on any issue or pull request to go back to the pull request listing page.