Skip to content

Pull requests: NVIDIA/Model-Optimizer

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Require timeout and output-limit accounting in eval skills
#2499 opened Sep 22, 2026 by Edwardf0t1 Contributor Loading…
specdec_bench: forward kv_cache_dtype to the vLLM engine
#2496 opened Sep 22, 2026 by yeyu-nvidia Contributor Loading…
Fix GitLab authentication guidance for Enroot
#2493 opened Sep 21, 2026 by chadvoegele Contributor Draft
Add MLflow tracking flags to megatron_bridge quantize.py
#2477 opened Sep 18, 2026 by kevalmorabia97 Collaborator Loading…
Normalize calibrated ONNX request contracts
#2476 opened Sep 18, 2026 by ajrasane Contributor Loading…
Refactor ONNX Runtime patching into capability modules
#2472 opened Sep 18, 2026 by ajrasane Contributor Loading…
[6034518] Fix remote safety benchmark latency
#2467 opened Sep 18, 2026 by ajrasane Contributor Draft
Gdn qad
#2455 opened Sep 17, 2026 by sychen52 Contributor Draft
Add Llama to Puzzletron v2 puzzletron_v2 Related to feature/puzzletron_v2 branch
#2454 opened Sep 17, 2026 by grzegorz-k-karch Contributor Draft
ProTip! Mix and match filters to narrow down what you’re looking for.