-
Notifications
You must be signed in to change notification settings - Fork 554
Pull requests: NVIDIA/Model-Optimizer
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Automate Puzzletron setup and fix unattended runtime bugs
#2245
opened Aug 25, 2026 by
j-rausch
Contributor
Loading…
[ONNX][Autocast] Adds
nodes_to_exclude regex support to the QDQ-aware convert_to_f16() API
#2241
opened Aug 24, 2026 by
jai17
Loading…
[chore]: weekly bump of uv.lock on main (2026-08-24)
#2234
opened Aug 24, 2026 by
github-actions
Bot
Loading…
fix(launcher): use the schema's draft_model global var, and validate global_vars keys
#2232
opened Aug 24, 2026 by
h-guo18
Contributor
Loading…
Fix distributed AutoQuantize scoring and share backward setup
#2231
opened Aug 23, 2026 by
joshua-hill
Loading…
Set NVFP4 MSE preset effective bits to 4.5
#2223
opened Aug 20, 2026 by
realAsma
Contributor
Loading…
fix(speculative): read rope_theta from rope_parameters first
#2221
opened Aug 20, 2026 by
h-guo18
Contributor
Loading…
Restructure recipes: split per-model_type recipes from model-hub checkpoint recipes
#2219
opened Aug 20, 2026 by
shengliangxu
Collaborator
Loading…
[Speculative Decoding] DFlash2 draft variant (sublayer convolution + candidate selector)
#2216
opened Aug 19, 2026 by
h-guo18
Contributor
Loading…
Group MLA q_a_proj/kv_a_proj_with_mqa projections in auto_quantize
#2213
opened Aug 18, 2026 by
joshua-hill
•
Draft
Add Kimi-K3 NVFP4 experts and FP8-PB attention recipe
#2206
opened Aug 17, 2026 by
Edwardf0t1
Contributor
Loading…
Updating Cosmos3 Nano DFlash recipe with optional fixes
#2205
opened Aug 17, 2026 by
adsridhar
Loading…
feat(quantization): fail fast when a quant config matches no weight quantizer
#2203
opened Aug 17, 2026 by
Edwardf0t1
Contributor
Loading…
Previous Next
ProTip!
Find all pull requests that aren't related to any open issues with -linked:issue.