Skip to content

Pull requests: NVIDIA/Model-Optimizer

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Fix AA-next OpenHands timeout guidance
#2094 opened Aug 6, 2026 by Edwardf0t1 Contributor Draft
[6562078]: fix calibration for vLLM 0.26.0 cherry-pick-0.46.0
#2093 opened Aug 6, 2026 by kinjalpatel27 Contributor Loading…
refactor(export): split unified_export_hf into layered modules
#2088 opened Aug 6, 2026 by Fridah-nv Contributor Loading…
Add FastGen quantization-aware distillation
#2085 opened Aug 5, 2026 by jingyu-ml Contributor Draft
[NVBUG: 6562021] Fix vLLM FlashAttention KV cache layout handling
#2084 opened Aug 5, 2026 by sychen52 Contributor Loading…
Fix/tied weight export identity
#2081 opened Aug 5, 2026 by chadvoegele Contributor Loading…
fix(autotune): pre-check remote board connectivity before benchmark
#2078 opened Aug 5, 2026 by willg-nv Contributor Loading…
Puzletron v2 dockerfile
#2077 opened Aug 5, 2026 by chochowski Contributor Loading…
Add Qwen-Image DMD2 QAT and PEFT-backed SVDQuant
#2069 opened Aug 5, 2026 by jingyu-ml Contributor Draft
Bug fix: 6542481 cherry-pick-0.46.0
#2064 opened Aug 4, 2026 by sugunav14 Contributor Loading…
Fix FSDP2 handling for tied embeddings
#2059 opened Aug 3, 2026 by realAsma Contributor Draft
Add Cosmos3 Nano DFlash multimodal training recipe
#2053 opened Aug 3, 2026 by skierat Contributor Loading…
ProTip! Mix and match filters to narrow down what you’re looking for.