Repository Issue Activity (beta)

justvugg/colibri

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
62
New in 7 Days
26
Closed in 7 Days
17
Average Open Age
21 days
Stale 30+ Days
2
Stale 90+ Days
0
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-075011159
2026-09-06464916
2026-09-0536245
2026-09-04725105
2026-09-0310002
2026-09-0211344
2026-09-01105856
2026-08-3142000
2026-08-3023000
2026-08-2930000
2026-08-28127000
2026-08-2730000
2026-08-2651000
2026-08-2500000
This Week

Opened: 22

Closed: 15

Comments: 30

Events: 50

Top Labels
bug (70)
performance (53)
benchmark (24)
cuda (22)
discussion (22)
feature (20)
model-support (14)
enhancement (13)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#1384 [Performance]: GLM-5.2 int4 gs64 on Azure D128ads v7, 4x local NVMe RAID0

Opened 12 hours ago
simone-gasparini
open
No labels
1012 hours ago

#1376 [Feature]:

Opened 22 hours ago
ChonSong
open
No labels
3012 hours ago

#1375 [Bug]: GLM 5.3-Flash puts Coli Tune OOM in 1.10.2

Opened 1 day ago
claws61821
open
No labels
5014 hours ago

#1380 [Performance]: GLM-5.2 int4 gs64 on Azure E64pds v6 Arm64, 4x local NVMe RAID0

Opened 14 hours ago
simone-gasparini
open
No labels
1014 hours ago

#1379 [Performance]: GLM-5.2 int4 gs64 on Azure E64ads v7, 4x local NVMe RAID0

Opened 15 hours ago
simone-gasparini
open
No labels
1015 hours ago

#1378 [Feature]: Add Architecture for OpenAI-Compatible Tool Translation Proxy via Sidecar/Dual-Engine Routing

Opened 15 hours ago
ChonSong
open
No labels
0015 hours ago

#1370 Perplexity vs Ollama's Q4_K_M on the same tokens: our int4 experts lose 2.5–3.4 %, the int8 trunk buys nothing — two questions about the format

Opened 1 day ago
kreuzzelg
open
No labels
1018 hours ago

#689 [Bug]: deep speculative verify batches are not token-exact vs CPU on CUDA (near-tie flips + lower draft acceptance)

Opened 1 month ago
terrizoaguimor
open
bug
cuda
quality
8020 hours ago

#1339 qwen36 CUDA tier: qt_issue overruns G.is_x with two or more GPUs (stride 8*D into a 32*D buffer)

Opened 3 days ago
crichalchemist
closed - completed
No labels
301 day ago

#1341 qwen36 CUDA tier: int8 containers free the weights the tier still points at (regression from #1334)

Opened 3 days ago
crichalchemist
closed - completed
No labels
301 day ago

#1340 qwen36 CUDA tier: qt_shutdown never signals cv_take, so pthread_join can hang if a group is open

Opened 3 days ago
crichalchemist
closed - completed
No labels
201 day ago

#1331 [Bug]: CUDA VRAM expert tier silently no-ops on int8 checkpoints -- reserves budget, promotes zero experts, no error

Opened 4 days ago
jmpmachado
open
No labels
301 day ago

#1359 [Bug]: release packer does not follow the imports of scripts it packages for subprocess launch

Opened 2 days ago
monotophic
closed - completed
No labels
201 day ago

#1368 [Bug]: 1.10.1 Windows Coli Web GLM 5.3-Flash embed_tokens tensor

Opened 1 day ago
claws61821
closed - completed
No labels
401 day ago

#1242 [Feature]: Add support for Qwen3.8-Flash-Next

Opened 12 days ago
acedogblast
closed - completed
No labels
7211 day ago

#1365 [Bug]: coli doctor reports 2 missing core tensors on a healthy GLM-5.3-Flash: the required list is the GLM-5.2 layout

Opened 2 days ago
JustVugg
closed - completed
No labels
102 days ago

#1040 qwen36 GPU detection: 'none compute-capable' on NVIDIA (GTX 1660 Ti) → CPU fallback

Opened 23 days ago
sehHeiden
closed - not_planned
vulkan
model-support
1302 days ago

#1352 Discord invite link expired

Opened 2 days ago
jtinbergen
closed - completed
No labels
202 days ago

#683 I trained a small model for this engine from scratch, anyone can use it

Opened 1 month ago
view321
open
discussion
422 days ago

#1351 [Bug]: single-GPU auto expert tier stops at ~56% of budget: prefix_est divides by the MTP int8 width since #766 (masks #687 on dev)

Opened 2 days ago
terrizoaguimor
open
No labels
202 days ago

#687 [Bug]: CUDA_EXPERT_GB=auto fills all VRAM on single-GPU; lazy dense uploads then fail 60x and fall back to CPU silently

Opened 1 month ago
terrizoaguimor
open
bug
cuda
902 days ago

#708 [Experiment] expert-transition-history placement policy vs gate-momentum — controlled A/B for hypothesis #1

Opened 1 month ago
matthewworner
open
enhancement
help wanted
performance
1002 days ago

#797 Intel Optane Storage on PMEM 100 or PMEM 200

Opened 1 month ago
sigkill
open
performance
hardware-owner-needed
discussion
302 days ago

#441 [Performance]: PILOT prefetch runs at NVMe queue-depth 1, and DIRECT=1 re-reads WILLNEED-prefetched pages

Opened 2 months ago
cdhdt
open
performance
1202 days ago

#972 [Feature]: Add support for Mistral Large 3

Opened 27 days ago
CorentinWicht
open
feature
model-support
302 days ago

Rows per page:

1–25 of 422