Skip to content

Pull requests: NVIDIA/Model-Optimizer

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

[Example] Qwen3.5 DSpark · E2E Training Example
#2338 opened Sep 4, 2026 by jinzex Loading…
[5565357] Fix SDXL NVFP4 export and performance
#2336 opened Sep 4, 2026 by ajrasane Contributor Draft
fix(quantization): fix nvfp4 availability check
#2331 opened Sep 4, 2026 by Lee-YNU Loading…
Add Parallel Decoding Distillation to FastGen
#2329 opened Sep 4, 2026 by mxinO Contributor Draft
added new workflow
#2323 opened Sep 3, 2026 by grzegorz-k-karch Contributor Loading…
Reject unsupported partial-block INT4/W4A8 AWQ export
#2320 opened Sep 2, 2026 by realAsma Contributor Loading…
Fix TEGroupedMLP quantizer checkpoint resharding
#2319 opened Sep 2, 2026 by jenchen13 Contributor Loading…
[6410139] Fix ONNX AutoCast for large external initializers cherry-pick-0.47.0 Upcoming release
#2317 opened Sep 2, 2026 by ajrasane Contributor Loading…
docs: document speculation profiles and how to produce them
#2316 opened Sep 2, 2026 by yeyu-nvidia Contributor Loading…
[6508436] Fix BF16 FP8 ONNX export cherry-pick-0.47.0 Upcoming release
#2314 opened Sep 2, 2026 by ajrasane Contributor Loading…
Add the NVFP4 PTQ recipe for Qwen/Qwen3.8-2.4T-A95B
#2302 opened Sep 1, 2026 by shengliangxu Collaborator Loading…
ProTip! Exclude everything labeled bug with -label:bug.