-
Notifications
You must be signed in to change notification settings - Fork 581
Pull requests: NVIDIA/Model-Optimizer
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Unpack Qwen3.5 MoE routed experts and keep quantized lm_head on Megatron export
cherry-pick-0.47.0
Upcoming release
#2334
opened Sep 4, 2026 by
kevalmorabia97
Collaborator
Loading…
Rename modelopt_recipes/huggingface to model_type with backward-compat alias
#2328
opened Sep 3, 2026 by
shengliangxu
Collaborator
•
Draft
fp32 master weights for the DFlash draft, and keep them across a resume
#2322
opened Sep 3, 2026 by
h-guo18
Contributor
Loading…
Reject unsupported partial-block INT4/W4A8 AWQ export
#2320
opened Sep 2, 2026 by
realAsma
Contributor
Loading…
Fix TEGroupedMLP quantizer checkpoint resharding
#2319
opened Sep 2, 2026 by
jenchen13
Contributor
Loading…
[6410139] Fix ONNX AutoCast for large external initializers
cherry-pick-0.47.0
Upcoming release
#2317
opened Sep 2, 2026 by
ajrasane
Contributor
Loading…
docs: document speculation profiles and how to produce them
#2316
opened Sep 2, 2026 by
yeyu-nvidia
Contributor
Loading…
ar_validate: emit a speculation profile from in-training AR validation
#2315
opened Sep 2, 2026 by
yeyu-nvidia
Contributor
Loading…
[6508436] Fix BF16 FP8 ONNX export
cherry-pick-0.47.0
Upcoming release
#2314
opened Sep 2, 2026 by
ajrasane
Contributor
Loading…
export: attach speculation_profile.json to exported draft checkpoints
#2313
opened Sep 2, 2026 by
yeyu-nvidia
Contributor
Loading…
Add NVFP4 PTQ recipes for zai-org/GLM-5.3-Flash (experts-only and experts + dense MLP)
#2312
opened Sep 2, 2026 by
shengliangxu
Collaborator
Loading…
feat(export): support multimodal and MTP models in layerwise export
#2303
opened Sep 1, 2026 by
Fridah-nv
Contributor
Loading…
Add the NVFP4 PTQ recipe for Qwen/Qwen3.8-2.4T-A95B
#2302
opened Sep 1, 2026 by
shengliangxu
Collaborator
Loading…
Forward kv_cache_free_gpu_memory_fraction to the TensorRT-LLM engines (NVBug 6701763)
#2300
opened Sep 1, 2026 by
cjluo-nv
Collaborator
Loading…
Fix fsdp2_aware_weight_update masking setup errors with UnboundLocalError
#2295
opened Sep 1, 2026 by
harshal-96
Loading…
2 of 4 tasks
specdec: config_overrides for nested text_config checkpoints + load VLM-capable bases in merge_lora
#2289
opened Aug 31, 2026 by
yeyu-nvidia
Contributor
Loading…
Previous Next
ProTip!
Exclude everything labeled
bug with -label:bug.