Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-20 | 0 | 0 | 0 | 0 | 1 |
| 2026-09-19 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-18 | 0 | 2 | 0 | 0 | 95 |
| 2026-09-17 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-16 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-15 | 1 | 1 | 0 | 0 | 0 |
| 2026-09-14 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-13 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-11 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-10 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-09 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-08 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-07 | 0 | 0 | 0 | 0 | 0 |
Opened: 3
Closed: 3
Comments: 0
Events: 0
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#2481 Your project is on StackMap — a curated map of the AI stack Opened 23 hours ago | hoghweed | open | No labels | 0 | 0 | 23 hours ago |
#1871 Clarify canonical documentation between guides and example READMEs (pruning) Opened 3 months ago | danielkorzekwa | closed - completed | bug documentation feature request | 7 | 0 | 2 days ago |
#2456 GLM-5.3 nvfp4 checkpoint Opened 3 days ago | yz342 | closed - completed | feature request | 1 | 0 | 2 days ago |
#2442 qdq_to_dq fails on MaxViT due to unsupported Transpose after DequantizeLinear Opened 5 days ago | blake-bodycote-advitech | open | bug | 1 | 0 | 5 days ago |
#2424 calling mtq.quantize causes AttributeError: 'tuple' object has no attribute 'dim' Opened 7 days ago | ccarmatic | open | question | 1 | 0 | 5 days ago |
#2209 Per-expert weight_quantizer._amax fails Megatron validate_sharding_integrity on topology reshard (TEGroupedMLP NVFP4) Opened 1 month ago | kevalmorabia97 | closed - completed | No labels | 3 | 0 | 5 days ago |
#215 INT8 to FP8 scale conversion Opened 1 year ago | roamiri | closed - completed | No labels | 5 | 0 | 6 days ago |
#2011 NVFP4 support for Qwen3_5MoeExperts (fused MoE): quantizer registration + export serializer Opened 2 months ago | frischeDaten | open | No labels | 8 | 0 | 11 days ago |
#1699 [RFC] NVIDIA Model Optimizer — Product Roadmap Opened 3 months ago | Trenton-Starkey | open | roadmap | 4 | 0 | 11 days ago |
#2337 Only 1 sample from AutoCast calibration data set is truly used Opened 16 days ago | lashgar | open | bug | 1 | 0 | 13 days ago |
#1926 [Bugs] Second batch of verified findings from the unit-coverage initiative (fold_weight crash on Llama-4/GPT-OSS MoE, KD-loss default footgun, and smaller items) Opened 3 months ago | arham766 | open | No labels | 4 | 0 | 15 days ago |
#2330 After applying NVFP4 quantization followed by compression, RealQuantLinear fails to locate the corresponding GEMM implementation during inference. Opened 17 days ago | Lee-YNU | open | bug | 2 | 0 | 16 days ago |
#2326 CI: onnx example lanes fail with CUDNN_STATUS_SUBLIBRARY_LOADING_FAILED on linux-amd64-gpu-rtxpro6000-latest-1 Opened 17 days ago | yeyu-nvidia | closed - completed | No labels | 1 | 0 | 17 days ago |
#1895 warnings.warn(f"RealQuantLinear: No real-quant GEMM found: {self}.") Opened 3 months ago | zewenli98 | closed - completed | bug waiting for feedback torch.quantization | 16 | 0 | 17 days ago |
#2189 megatron_generate drops VLM vision inputs during no-cache decoding Opened 1 month ago | cuichenx | closed - completed | bug | 1 | 0 | 17 days ago |
#2002 [QAD] Request for exact training configurations used in arXiv:2601.20088 Opened 2 months ago | Sekri0 | open | No labels | 5 | 0 | 18 days ago |
#1577 generate_context_logits returns truncated logits when KV prefix cache is enabled, causing silently wrong loglikelihood evaluation Opened 4 months ago | HouChentripleg | open | bug | 1 | 0 | 19 days ago |
#1974 Nvidia Modelopt structured 2:4 weight sparsity showing no speed improvement compared to dense model Opened 2 months ago | Pavan6136 | open | bug | 5 | 0 | 21 days ago |
#2281 ONNX PTQ: no NVFP4 path for convolutional models Opened 23 days ago | geoffrey-delhomme | open | No labels | 0 | 0 | 23 days ago |
#2131 fold_weight crashes with AttributeError: 'NoneType' object has no attribute 'data' on Megatron models with tied word embeddings Opened 1 month ago | babyplutokurt | closed - completed | bug | 1 | 0 | 23 days ago |
#2230 hf_ptq.py: deprecated --auto_quantize_bits CLI flag silently no-ops instead of enabling AutoQuantize Opened 30 days ago | wyattearp | open | No labels | 1 | 0 | 27 days ago |
#2220 [Feature Request] DeepSeek-V4-Flash-0731 NVFP4 checkpoint Opened 1 month ago | nicole-lihui | open | feature request | 2 | 0 | 1 month ago |
#2204 LUT-B Support Opened 1 month ago | brian-dellabetta | open | question | 1 | 0 | 1 month ago |
#2160 [Bug] init_quantized_weights / --low_memory_mode exports numerically broken NVFP4 (quantizes meta tensors before weights load) Opened 1 month ago | spped2000 | open | No labels | 0 | 0 | 1 month ago |
#2158 Can `auto_quant` calculate the score for `kv-cache` separately? Opened 1 month ago | zcfh | open | question | 0 | 0 | 1 month ago |