-
Notifications
You must be signed in to change notification settings - Fork 576
Pull requests: NVIDIA/Model-Optimizer
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Speed up megatron_bridge example tests by ~6x on a single GPU
cherry-pick-0.47.0
Upcoming release
#2296
opened Sep 1, 2026 by
kevalmorabia97
Collaborator
Loading…
Fix fsdp2_aware_weight_update masking setup errors with UnboundLocalError
#2295
opened Sep 1, 2026 by
harshal-96
Loading…
2 of 4 tasks
README: link Announcement Blogs and add AutoQuantize blog to Latest News
#2291
opened Aug 31, 2026 by
realAsma
Contributor
Loading…
specdec: config_overrides for nested text_config checkpoints + load VLM-capable bases in merge_lora
#2289
opened Aug 31, 2026 by
yeyu-nvidia
Contributor
Loading…
ar_validate: fail loudly when every sample fails
#2288
opened Aug 31, 2026 by
yeyu-nvidia
Contributor
Loading…
[chore]: weekly bump of uv.lock on main (2026-08-31)
#2285
opened Aug 31, 2026 by
github-actions
Bot
Loading…
Support quantized Qwen3-VL / Qwen3.5-VL (dense + MoE) export from Megatron-Bridge and verify exported checkpoints
cherry-pick-0.47.0
Upcoming release
#2276
opened Aug 28, 2026 by
kevalmorabia97
Collaborator
Loading…
Gkarch/sync main 449a3992
puzzletron_v2
Related to feature/puzzletron_v2 branch
#2266
opened Aug 27, 2026 by
grzegorz-k-karch
Contributor
Loading…
Add a Docker image for Puzzletron v2 workers
puzzletron_v2
Related to feature/puzzletron_v2 branch
#2265
opened Aug 27, 2026 by
j-rausch
Contributor
Loading…
Docs: Add WOA documentation
cherry-pick-0.47.0
Upcoming release
#2264
opened Aug 27, 2026 by
haoxiz-nvidia
Contributor
Loading…
Add TensorRT-RTX ABI EP support for ONNX quantization
cherry-pick-0.47.0
Upcoming release
#2262
opened Aug 27, 2026 by
haoxiz-nvidia
Contributor
Loading…
Docs: Add QAT and QAD guide [OMNIML-4859]
#2255
opened Aug 26, 2026 by
jenchen13
Contributor
Loading…
fix: make huggingface_hub plugin import optional
#2250
opened Aug 26, 2026 by
Y-T-G
Contributor
Loading…
fix(quantization): warn once when calibration runs with a KV cache
#2248
opened Aug 25, 2026 by
Fridah-nv
Contributor
Loading…
specdec_bench: emit speculation_profile.json alongside acceptance metrics
#2247
opened Aug 25, 2026 by
yeyu-nvidia
Contributor
Loading…
Add Aumann-Shapley AutoQuantize recipe integration
#2246
opened Aug 25, 2026 by
joshua-hill
Loading…
[ONNX] Add quantization sensitivity ranking + exclusion picker
#2240
opened Aug 24, 2026 by
gcunhase
Contributor
Loading…
Previous Next
ProTip!
Type g i on any issue or pull request to go back to the issue listing page.