Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-15 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-14 | 0 | 0 | 0 | 0 | 90 |
| 2026-09-13 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-11 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-10 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-09 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-08 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-07 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-06 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-05 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-04 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-03 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-02 | 0 | 0 | 0 | 0 | 0 |
Opened: 0
Closed: 0
Comments: 0
Events: 0
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#571 Curious about Exllama+TP Opened 2 years ago | grimulkan | open - reopened | No labels | 15 | 0 | 2 months ago |
#817 Multiple out-of-bounds accesses in QMatrix quant-metadata handling via a crafted GPTQ/EXL2 model Opened 2 months ago | professor-moody | open | No labels | 0 | 0 | 2 months ago |
#814 E8 lattice VQ for KV cache: 2-3 bit with asymmetric K/V — cross-model results Opened 5 months ago | jagmarques | closed - completed | No labels | 2 | 0 | 4 months ago |
#813 [BUG] System freeze during inference with RTX 5060 Ti via Thunderbolt 5 eGPU on Linux (kernel freeze, SSH drop, GPU fan full speed) Opened 6 months ago | Vassili79 | open | bug | 1 | 0 | 6 months ago |
#811 [REQUEST] Please add support for Qwen3VL Opened 11 months ago | sujitvasanth | open | No labels | 1 | 0 | 11 months ago |
#808 [BUG] 2080ti outputs gibberish Opened 11 months ago | frenzybiscuit | open | bug | 1 | 0 | 11 months ago |
#804 [REQUEST]Please support Torch 2.8 Opened 1 year ago | pivtienduc | open | No labels | 2 | 0 | 11 months ago |
#809 [QUESTION] Inquiry about DynamicGenerator: EOS not returned until max_new_tokens reached, despite stop_conditions Opened 11 months ago | keds-rnd | open - reopened | bug | 3 | 0 | 11 months ago |
#810 [REQUEST] Make a version that is on Cuda V13 Opened 11 months ago | thebaconstudio | open | No labels | 0 | 0 | 11 months ago |
#806 [QUESTION] I cannot achieve token/s speed as in your tests even with better GPU Opened 1 year ago | piercarlo62 | open | bug | 0 | 0 | 1 year ago |
#803 [BUG] [Cause Identified & Workarounds Proposed] Infinite Generation for Llama3 & Other Models with Multiple EOS IDs Opened 1 year ago | abgulati | open | bug | 1 | 1 | 1 year ago |
#800 [BUG] Dry sampling over long contexts Opened 1 year ago | Ph0rk0z | closed - completed | bug | 1 | 0 | 1 year ago |
#802 [BUG] Unable to run examples/chat.py Opened 1 year ago | homeworkace | open | bug | 0 | 0 | 1 year ago |
#795 [BUG] ExllamaV2 version >0.2.8 broken for mistral 7b(v0.2) models on Nvidia 2060 Opened 1 year ago | IceFog72 | open | bug | 7 | 0 | 1 year ago |
#784 [BUG] Runtime error when trying to load Qwen3 32B Opened 1 year ago | umar-mq | open | bug | 10 | 0 | 1 year ago |
#798 [REQUEST] Support for Hunyuan-A13B-Instruct Opened 1 year ago | RodriMora | open | No labels | 0 | 1 | 1 year ago |
#797 [BUG] ExLlamaV2Generator import broken across WHL & source repo — Windows 11 + RTX 5090 + CUDA 12.8 build inconsistencies Opened 1 year ago | mindworksmanagement | open | bug | 0 | 0 | 1 year ago |
#777 [BUG]gemma 3 27b exl2 loops nonsense afterwards 2-3 correct paragraphs Opened 1 year ago | ciprianveg | open | bug | 11 | 0 | 1 year ago |
#749 [REQUEST] Please add support for Gemma3. Opened 2 years ago | emzaedu | open | No labels | 38 | 18 | 1 year ago |
#793 [BUG] Exllamav2 quickly devolves into endless repetition in versions newer than 2.8.0 Opened 1 year ago | ZhenyaPav | open | bug | 1 | 0 | 1 year ago |
#789 [BUG] Remove Sentencepiece Opened 1 year ago | kingbri1 | closed - completed | bug | 2 | 0 | 1 year ago |
#724 [REQUEST] Support new SOTA vision model: Qwen 2.5 VL (3B, 7B, 72B) Opened 2 years ago | ThomasBaruzier | closed - completed | No labels | 2 | 2 | 1 year ago |
#792 [QUESTION] Estimate measurements and quantization peak VRAM use in advance Opened 1 year ago | ThomasBaruzier | open | No labels | 0 | 0 | 1 year ago |
#790 [BUG] Silent crash in safetensors call when compiling shards Opened 1 year ago | r0mar0ma | open | bug | 0 | 0 | 1 year ago |
#780 [BUG] Blue Screen MEMORY_MANAGEMENT Error when trying to quantize Gemma3. Opened 1 year ago | Nrgte | open | bug | 2 | 0 | 1 year ago |