Skip to content

Pull requests: NVIDIA/Model-Optimizer

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Speed up compressed-tensors load-time matching (for Kimi models)
#1999 opened Jul 21, 2026 by rohansjoshi Contributor Loading…
Offline-KD QAD example
#1998 opened Jul 20, 2026 by AAnoosheh Contributor Loading…
Add link to puzzletron_v2
#1996 opened Jul 20, 2026 by Separius Contributor Loading…
Document non-Diffusers FP8/NVFP4 ComfyUI export
#1991 opened Jul 17, 2026 by jingyu-ml Contributor Draft
Make group-boundary AutoQuant scoring the default
#1988 opened Jul 17, 2026 by meenchen Contributor Draft
Add MLflow tracking to modelopt MCP
#1986 opened Jul 17, 2026 by ChenhanYu Collaborator Loading…
[6425069][ONNX][Autocast] Fix autocast metadata propagation
#1983 opened Jul 16, 2026 by gcunhase Contributor Loading…
Scripts and a skill to do per-layer benchmark using flashinfer
#1980 opened Jul 16, 2026 by sychen52 Contributor Loading…
docs: add AutoQuantize mixed-precision search blog
#1979 opened Jul 15, 2026 by realAsma Contributor Draft
add Qwen3-VL support for DFlash training
#1975 opened Jul 14, 2026 by skierat Contributor Loading…
Add MoE support to runtime stats
#1973 opened Jul 14, 2026 by grzegorz-k-karch Contributor Draft
docs: add announcements landing page
#1971 opened Jul 13, 2026 by ChenhanYu Collaborator Draft
[Examples]: MiniMax-M3 DSpark
#1965 opened Jul 12, 2026 by h-guo18 Contributor Loading…
ProTip! What’s not been updated in a month: updated:<2026-06-21.