Skip to content

Pull requests: NVIDIA/Model-Optimizer

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

fix(quantization): fix nvfp4 availability check
#2331 opened Sep 4, 2026 by Lee-YNU Loading…
Add Parallel Decoding Distillation to FastGen
#2329 opened Sep 4, 2026 by mxinO Contributor Draft
[skill] evaluation: mandate 8 SciCode runs and report the mean
#2327 opened Sep 3, 2026 by cjluo-nv Collaborator Loading…
added new workflow
#2323 opened Sep 3, 2026 by grzegorz-k-karch Contributor Loading…
Reject unsupported partial-block INT4/W4A8 AWQ export
#2320 opened Sep 2, 2026 by realAsma Contributor Loading…
Fix TEGroupedMLP quantizer checkpoint resharding
#2319 opened Sep 2, 2026 by jenchen13 Contributor Loading…
[6410139] Fix ONNX AutoCast for large external initializers cherry-pick-0.47.0 Upcoming release
#2317 opened Sep 2, 2026 by ajrasane Contributor Loading…
docs: document speculation profiles and how to produce them
#2316 opened Sep 2, 2026 by yeyu-nvidia Contributor Loading…
[6508436] Fix BF16 FP8 ONNX export cherry-pick-0.47.0 Upcoming release
#2314 opened Sep 2, 2026 by ajrasane Contributor Loading…
Add the NVFP4 experts-only PTQ recipe for zai-org/GLM-5.3-Flash
#2312 opened Sep 2, 2026 by shengliangxu Collaborator Loading…
Add the NVFP4 PTQ recipe for Qwen/Qwen3.8-2.4T-A95B
#2302 opened Sep 1, 2026 by shengliangxu Collaborator Loading…
ar_validate: fail loudly when every sample fails
#2288 opened Aug 31, 2026 by yeyu-nvidia Contributor Loading…
[chore]: weekly bump of uv.lock on main (2026-08-31)
#2285 opened Aug 31, 2026 by github-actions Bot Loading…
Add WMSE weight scale search calibration
#2283 opened Aug 29, 2026 by realAsma Contributor Draft
ProTip! Find all pull requests that aren't related to any open issues with -linked:issue.