Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-08-24 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-23 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-22 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-21 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-20 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-19 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-18 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-17 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-16 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-15 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-14 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-13 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-11 | 0 | 0 | 0 | 0 | 0 |
Opened: 0
Closed: 0
Comments: 0
Events: 0
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#413 qint4 (TinyGemmWeightQBitsTensor) is much slower than qint2 for bf16 CUDA training workloads Opened 5 months ago | pyros-projects | closed - not_planned | Stale | 3 | 0 | 4 months ago |
#327 issues with non-contiguous Tensor Opened 2 years ago | bghira | closed - completed | No labels | 3 | 1 | 6 months ago |
#411 GPU Memory Leak During Layer-by-Layer Inference in QuantizedModelForCausalLM (or quantize & freeze) (Layers Not Released After Deletion) Opened 7 months ago | quaternior | closed - completed | No labels | 1 | 0 | 7 months ago |
#404 'TinyGemmWeightQBitsTensor' object has no attribute '_scale' Opened 10 months ago | wikeeyang | closed - completed | Stale | 5 | 0 | 7 months ago |
#405 is this project still maintained? Opened 10 months ago | bghira | closed - completed | No labels | 6 | 0 | 9 months ago |
#406 wrapped tensor does not expose data pointer Opened 10 months ago | bghira | closed - completed | invalid | 1 | 0 | 9 months ago |
#400 Is model parallelism supported? Opened 1 year ago | DominikHil | closed - not_planned | Stale | 2 | 0 | 11 months ago |
#397 module incompatible Opened 1 year ago | Yuyok | closed - not_planned | Stale | 2 | 0 | 11 months ago |
#398 Error loading quantized qwen2.5 vl model Opened 1 year ago | QQQ27 | closed - completed | No labels | 0 | 0 | 1 year ago |
#396 The quantized model does not support FSDP. Opened 1 year ago | Piggy-ch | closed - not_planned | Stale | 2 | 0 | 1 year ago |
#393 Support Multi-GPU calibration Opened 1 year ago | Piggy-ch | closed - not_planned | Stale | 2 | 0 | 1 year ago |
#391 The `bench/generation/metrics/perplexity.py` will produce a wrong result when n_ctx > n_batch Opened 1 year ago | yrom | closed - not_planned | Stale | 2 | 0 | 1 year ago |
#339 Error when quantizing flux (already fixed in master) Opened 2 years ago | samedii | closed - completed | No labels | 2 | 0 | 1 year ago |
#388 Some tensors share memory, this will lead to duplicate memory on disk and potential differences when loading them agai Opened 1 year ago | laborer996 | closed - not_planned | Stale | 3 | 0 | 1 year ago |
#386 RuntimeError: Error building extension 'quanto_cuda': Opened 1 year ago | 932179209 | closed - not_planned | Stale | 3 | 0 | 1 year ago |
#381 can "tests" be used instead of "test" for test folder? Opened 1 year ago | faaany | closed - completed | Stale | 4 | 0 | 1 year ago |
#385 param.dtype of the model is torch.float32 after quantization Opened 1 year ago | adityarajsahu | closed - not_planned | Stale | 5 | 0 | 1 year ago |
#379 bench/kernels/benchmark.py fails due to usage of pruned API disable_extensions Opened 1 year ago | dvrogozh | closed - not_planned | Stale | 2 | 0 | 1 year ago |
#378 When using `Calibration`, only the first batch returns `QTensor` Opened 1 year ago | FredBill1 | closed - not_planned | Stale | 2 | 0 | 1 year ago |
#361 ValueError: invalid literal for int() with base 10: '90a' Opened 2 years ago | MatthewCroughan | closed - not_planned | Stale | 7 | 0 | 1 year ago |
#311 quantize(model, weights=qint4, activations=qint8) produce weights with dtype = torch.uint8 Opened 2 years ago | lifelongeeek | closed - completed | No labels | 3 | 0 | 1 year ago |
#375 When will we be getting full Quanto support for WAN 2.1? Opened 1 year ago | ukaprch | closed - not_planned | Stale | 2 | 0 | 1 year ago |
#376 Cannot use pipe.enable_model_cpu_offload() to quantized diffusers transformers Opened 1 year ago | james-imi | closed - not_planned | Stale | 2 | 0 | 1 year ago |
#374 Can run an int8 quantized model on CUDA? Opened 1 year ago | shencuifeng | closed - not_planned | Stale | 4 | 0 | 1 year ago |
#254 Packages created on the CI are missing cpp and cuda extension files Opened 2 years ago | dacorvo | closed - not_planned | Stale | 12 | 0 | 1 year ago |