Skip to content

Pull requests: NVIDIA/Model-Optimizer

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Consolidate example CI lanes and bump containers to latest
#2086 opened Aug 5, 2026 by kevalmorabia97 Collaborator Loading…
Add FastGen quantization-aware distillation
#2085 opened Aug 5, 2026 by jingyu-ml Contributor Draft
[NVBUG: 6562021] Fix vLLM FlashAttention KV cache layout handling
#2084 opened Aug 5, 2026 by sychen52 Contributor Loading…
Fix/tied weight export identity
#2081 opened Aug 5, 2026 by chadvoegele Contributor Loading…
fix(autotune): pre-check remote board connectivity before benchmark
#2078 opened Aug 5, 2026 by willg-nv Contributor Loading…
Puzletron v2 dockerfile
#2077 opened Aug 5, 2026 by chochowski Contributor Loading…
Add Qwen-Image DMD2 QAT and PEFT-backed SVDQuant
#2069 opened Aug 5, 2026 by jingyu-ml Contributor Draft
Bug fix: 6542481 cherry-pick-0.46.0
#2064 opened Aug 4, 2026 by sugunav14 Contributor Loading…
Fix FSDP2 handling for tied embeddings
#2059 opened Aug 3, 2026 by realAsma Contributor Draft
Add Cosmos3 Nano DFlash multimodal training recipe
#2053 opened Aug 3, 2026 by skierat Contributor Loading…
Add vLLM skip-softmax and mask-reuse calibration
#2045 opened Aug 1, 2026 by kaix-nv Contributor Draft
ProTip! Type g i on any issue or pull request to go back to the issue listing page.