Repository Issue Activity (beta)

nvidia/model-optimizer

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
94
New in 7 Days
4
Closed in 7 Days
3
Average Open Age
183 days
Stale 30+ Days
81
Stale 90+ Days
52
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-2000001
2026-09-1910000
2026-09-18020095
2026-09-1710000
2026-09-1600000
2026-09-1511000
2026-09-1400000
2026-09-1310000
2026-09-1200000
2026-09-1100000
2026-09-1000000
2026-09-0900000
2026-09-0800000
2026-09-0700000
This Week

Opened: 3

Closed: 3

Comments: 0

Events: 0

Top Labels
bug (131)
feature request (72)
waiting for feedback (62)
question (59)
torch.quantization (57)
stale (55)
investigating (52)
model support (44)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#2481 Your project is on StackMap — a curated map of the AI stack

Opened 23 hours ago
hoghweed
open
No labels
0023 hours ago

#1871 Clarify canonical documentation between guides and example READMEs (pruning)

Opened 3 months ago
danielkorzekwa
closed - completed
bug
documentation
feature request
702 days ago

#2456 GLM-5.3 nvfp4 checkpoint

Opened 3 days ago
yz342
closed - completed
feature request
102 days ago

#2442 qdq_to_dq fails on MaxViT due to unsupported Transpose after DequantizeLinear

Opened 5 days ago
blake-bodycote-advitech
open
bug
105 days ago

#2424 calling mtq.quantize causes AttributeError: 'tuple' object has no attribute 'dim'

Opened 7 days ago
ccarmatic
open
question
105 days ago

#2209 Per-expert weight_quantizer._amax fails Megatron validate_sharding_integrity on topology reshard (TEGroupedMLP NVFP4)

Opened 1 month ago
kevalmorabia97
closed - completed
No labels
305 days ago

#215 INT8 to FP8 scale conversion

Opened 1 year ago
roamiri
closed - completed
No labels
506 days ago

#2011 NVFP4 support for Qwen3_5MoeExperts (fused MoE): quantizer registration + export serializer

Opened 2 months ago
frischeDaten
open
No labels
8011 days ago

#1699 [RFC] NVIDIA Model Optimizer — Product Roadmap

Opened 3 months ago
Trenton-Starkey
open
roadmap
4011 days ago

#2337 Only 1 sample from AutoCast calibration data set is truly used

Opened 16 days ago
lashgar
open
bug
1013 days ago

#1926 [Bugs] Second batch of verified findings from the unit-coverage initiative (fold_weight crash on Llama-4/GPT-OSS MoE, KD-loss default footgun, and smaller items)

Opened 3 months ago
arham766
open
No labels
4015 days ago

#2330 After applying NVFP4 quantization followed by compression, RealQuantLinear fails to locate the corresponding GEMM implementation during inference.

Opened 17 days ago
Lee-YNU
open
bug
2016 days ago

#2326 CI: onnx example lanes fail with CUDNN_STATUS_SUBLIBRARY_LOADING_FAILED on linux-amd64-gpu-rtxpro6000-latest-1

Opened 17 days ago
yeyu-nvidia
closed - completed
No labels
1017 days ago

#1895 warnings.warn(f"RealQuantLinear: No real-quant GEMM found: {self}.")

Opened 3 months ago
zewenli98
closed - completed
bug
waiting for feedback
torch.quantization
16017 days ago

#2189 megatron_generate drops VLM vision inputs during no-cache decoding

Opened 1 month ago
cuichenx
closed - completed
bug
1017 days ago

#2002 [QAD] Request for exact training configurations used in arXiv:2601.20088

Opened 2 months ago
Sekri0
open
No labels
5018 days ago

#1577 generate_context_logits returns truncated logits when KV prefix cache is enabled, causing silently wrong loglikelihood evaluation

Opened 4 months ago
HouChentripleg
open
bug
1019 days ago

#1974 Nvidia Modelopt structured 2:4 weight sparsity showing no speed improvement compared to dense model

Opened 2 months ago
Pavan6136
open
bug
5021 days ago

#2281 ONNX PTQ: no NVFP4 path for convolutional models

Opened 23 days ago
geoffrey-delhomme
open
No labels
0023 days ago

#2131 fold_weight crashes with AttributeError: 'NoneType' object has no attribute 'data' on Megatron models with tied word embeddings

Opened 1 month ago
babyplutokurt
closed - completed
bug
1023 days ago

#2230 hf_ptq.py: deprecated --auto_quantize_bits CLI flag silently no-ops instead of enabling AutoQuantize

Opened 30 days ago
wyattearp
open
No labels
1027 days ago

#2220 [Feature Request] DeepSeek-V4-Flash-0731 NVFP4 checkpoint

Opened 1 month ago
nicole-lihui
open
feature request
201 month ago

#2204 LUT-B Support

Opened 1 month ago
brian-dellabetta
open
question
101 month ago

#2160 [Bug] init_quantized_weights / --low_memory_mode exports numerically broken NVFP4 (quantizes meta tensors before weights load)

Opened 1 month ago
spped2000
open
No labels
001 month ago

#2158 Can `auto_quant` calculate the score for `kv-cache` separately?

Opened 1 month ago
zcfh
open
question
001 month ago

Rows per page:

1–25 of 405