Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-07 | 5 | 0 | 11 | 15 | 9 |
| 2026-09-06 | 4 | 6 | 4 | 9 | 16 |
| 2026-09-05 | 3 | 6 | 2 | 4 | 5 |
| 2026-09-04 | 7 | 2 | 5 | 10 | 5 |
| 2026-09-03 | 1 | 0 | 0 | 0 | 2 |
| 2026-09-02 | 1 | 1 | 3 | 4 | 4 |
| 2026-09-01 | 1 | 0 | 5 | 8 | 56 |
| 2026-08-31 | 4 | 2 | 0 | 0 | 0 |
| 2026-08-30 | 2 | 3 | 0 | 0 | 0 |
| 2026-08-29 | 3 | 0 | 0 | 0 | 0 |
| 2026-08-28 | 1 | 27 | 0 | 0 | 0 |
| 2026-08-27 | 3 | 0 | 0 | 0 | 0 |
| 2026-08-26 | 5 | 1 | 0 | 0 | 0 |
| 2026-08-25 | 0 | 0 | 0 | 0 | 0 |
Opened: 22
Closed: 15
Comments: 30
Events: 50
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#1384 [Performance]: GLM-5.2 int4 gs64 on Azure D128ads v7, 4x local NVMe RAID0 Opened 12 hours ago | simone-gasparini | open | No labels | 1 | 0 | 12 hours ago |
#1376 [Feature]: Opened 22 hours ago | ChonSong | open | No labels | 3 | 0 | 12 hours ago |
#1375 [Bug]: GLM 5.3-Flash puts Coli Tune OOM in 1.10.2 Opened 1 day ago | claws61821 | open | No labels | 5 | 0 | 14 hours ago |
#1380 [Performance]: GLM-5.2 int4 gs64 on Azure E64pds v6 Arm64, 4x local NVMe RAID0 Opened 14 hours ago | simone-gasparini | open | No labels | 1 | 0 | 14 hours ago |
#1379 [Performance]: GLM-5.2 int4 gs64 on Azure E64ads v7, 4x local NVMe RAID0 Opened 15 hours ago | simone-gasparini | open | No labels | 1 | 0 | 15 hours ago |
#1378 [Feature]: Add Architecture for OpenAI-Compatible Tool Translation Proxy via Sidecar/Dual-Engine Routing Opened 15 hours ago | ChonSong | open | No labels | 0 | 0 | 15 hours ago |
#1370 Perplexity vs Ollama's Q4_K_M on the same tokens: our int4 experts lose 2.5–3.4 %, the int8 trunk buys nothing — two questions about the format Opened 1 day ago | kreuzzelg | open | No labels | 1 | 0 | 18 hours ago |
#689 [Bug]: deep speculative verify batches are not token-exact vs CPU on CUDA (near-tie flips + lower draft acceptance) Opened 1 month ago | terrizoaguimor | open | bug cuda quality | 8 | 0 | 20 hours ago |
#1339 qwen36 CUDA tier: qt_issue overruns G.is_x with two or more GPUs (stride 8*D into a 32*D buffer) Opened 3 days ago | crichalchemist | closed - completed | No labels | 3 | 0 | 1 day ago |
#1341 qwen36 CUDA tier: int8 containers free the weights the tier still points at (regression from #1334) Opened 3 days ago | crichalchemist | closed - completed | No labels | 3 | 0 | 1 day ago |
#1340 qwen36 CUDA tier: qt_shutdown never signals cv_take, so pthread_join can hang if a group is open Opened 3 days ago | crichalchemist | closed - completed | No labels | 2 | 0 | 1 day ago |
#1331 [Bug]: CUDA VRAM expert tier silently no-ops on int8 checkpoints -- reserves budget, promotes zero experts, no error Opened 4 days ago | jmpmachado | open | No labels | 3 | 0 | 1 day ago |
#1359 [Bug]: release packer does not follow the imports of scripts it packages for subprocess launch Opened 2 days ago | monotophic | closed - completed | No labels | 2 | 0 | 1 day ago |
#1368 [Bug]: 1.10.1 Windows Coli Web GLM 5.3-Flash embed_tokens tensor Opened 1 day ago | claws61821 | closed - completed | No labels | 4 | 0 | 1 day ago |
#1242 [Feature]: Add support for Qwen3.8-Flash-Next Opened 12 days ago | acedogblast | closed - completed | No labels | 7 | 21 | 1 day ago |
#1365 [Bug]: coli doctor reports 2 missing core tensors on a healthy GLM-5.3-Flash: the required list is the GLM-5.2 layout Opened 2 days ago | JustVugg | closed - completed | No labels | 1 | 0 | 2 days ago |
#1040 qwen36 GPU detection: 'none compute-capable' on NVIDIA (GTX 1660 Ti) → CPU fallback Opened 23 days ago | sehHeiden | closed - not_planned | vulkan model-support | 13 | 0 | 2 days ago |
#1352 Discord invite link expired Opened 2 days ago | jtinbergen | closed - completed | No labels | 2 | 0 | 2 days ago |
#683 I trained a small model for this engine from scratch, anyone can use it Opened 1 month ago | view321 | open | discussion | 4 | 2 | 2 days ago |
#1351 [Bug]: single-GPU auto expert tier stops at ~56% of budget: prefix_est divides by the MTP int8 width since #766 (masks #687 on dev) Opened 2 days ago | terrizoaguimor | open | No labels | 2 | 0 | 2 days ago |
#687 [Bug]: CUDA_EXPERT_GB=auto fills all VRAM on single-GPU; lazy dense uploads then fail 60x and fall back to CPU silently Opened 1 month ago | terrizoaguimor | open | bug cuda | 9 | 0 | 2 days ago |
#708 [Experiment] expert-transition-history placement policy vs gate-momentum — controlled A/B for hypothesis #1 Opened 1 month ago | matthewworner | open | enhancement help wanted performance | 10 | 0 | 2 days ago |
#797 Intel Optane Storage on PMEM 100 or PMEM 200 Opened 1 month ago | sigkill | open | performance hardware-owner-needed discussion | 3 | 0 | 2 days ago |
#441 [Performance]: PILOT prefetch runs at NVMe queue-depth 1, and DIRECT=1 re-reads WILLNEED-prefetched pages Opened 2 months ago | cdhdt | open | performance | 12 | 0 | 2 days ago |
#972 [Feature]: Add support for Mistral Large 3 Opened 27 days ago | CorentinWicht | open | feature model-support | 3 | 0 | 2 days ago |