Repository Issue Activity (beta)

thetom/turboquant_plus

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
36
New in 7 Days
1
Closed in 7 Days
0
Average Open Age
143 days
Stale 30+ Days
35
Stale 90+ Days
33
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-0700000
2026-09-0600000
2026-09-0500000
2026-09-0400001
2026-09-0310000
2026-09-0200000
2026-09-01000035
2026-08-3100000
2026-08-3000000
2026-08-2900000
2026-08-2800000
2026-08-2700000
2026-08-2600000
2026-08-2500000
This Week

Opened: 1

Closed: 0

Comments: 0

Events: 0

Top Labels
P1 (13)
type:algorithm (12)
P0 (11)
type:benchmark (6)
type:bug (6)
P2 (3)
P3 (3)
type:port (3)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#99 CUDA vec FA kernel: turbo4 V-cache branch consumes 4 of 8 elements per iteration — half of every head's V output is zero on batch-1 decode (patch attached)

Opened 4 days ago
morpheos-llc
open
No labels
104 days ago

#97 uv pip install turboquant also required

Opened 1 month ago
jwtiyar
open
No labels
001 month ago

#92 Can it be used together with the MTP draft model?

Opened 4 months ago
serchcc
open
No labels
714 months ago

#89 Consider proper attribution for EDEN quantization

Opened 4 months ago
amitport
closed - completed
No labels
1804 months ago

#87 Use scale factor for improvements

Opened 5 months ago
Mitzenmacher
closed - completed
No labels
314 months ago

#27 Upstream: TurboQuant discussion + contribution requirements for llama.cpp

Opened 6 months ago
TheTom
open
type:port
P2
1124 months ago

#88 attn-rotation for native llama quants

Opened 5 months ago
erazortt
closed - completed
No labels
104 months ago

#86 [ROCm] Scale operation fails with "invalid device function" during Gemma 4 loading

Opened 5 months ago
http403
open
No labels
005 months ago

#85 Feature request: prebuilt binary for Windows CPU only

Opened 5 months ago
jadams777
open
No labels
005 months ago

#84 Diagnostic: [RTX 3090]

Opened 5 months ago
chertykov
open
No labels
005 months ago

#77 TurboFlash attention kernel: 13 GB excess Metal compute buffer allocation (commit 0d6b38aad)

Opened 5 months ago
shivam2014
closed - completed
No labels
105 months ago

#82 Question about QJL abliation study

Opened 5 months ago
ryusaeba
open
No labels
005 months ago

#80 time and performance overhead of quantization and dequantization

Opened 5 months ago
Szh1107
open
No labels
005 months ago

#79 Reproducibility of QJL resurrection issue

Opened 5 months ago
zhangsipeng
open
No labels
005 months ago

#74 it run

Opened 5 months ago
ayttop
open
No labels
205 months ago

#72 End-to-end inference test fails with --cache-type turbo4 on CPU

Opened 5 months ago
MELANCHOLY828
open
No labels
005 months ago

#70 Error when building from source

Opened 5 months ago
gabriel93blt
open
No labels
015 months ago

#69 polar coordinate transformation operations

Opened 5 months ago
Keshawn082
open
No labels
005 months ago

#68 llama-quantize crashes with ios_base::failbit when using --tensor-type-file or --tensor-type (regression from #20503)

Opened 5 months ago
dogukanatakul
open
No labels
005 months ago

#42 Enable sparse V dequant on pre-M5 hardware after verification

Opened 5 months ago
TheTom
open
No labels
115 months ago

#60 [Bug] Numerical instability (NaN/Infinite Loop) with Qwen 2.5/3.5 models using turbo4 on RTX 4070Ti

Opened 5 months ago
david01333
open
No labels
015 months ago

#59 Greetings from TurboQuant-MLX & Some Insights to Share! (System Prompt vs Boundary Layers)

Opened 5 months ago
helgklaizar
open
No labels
045 months ago

#57 turbo - cant compile in Windows

Opened 5 months ago
Ezzz-dev
open
No labels
105 months ago

#58 Unsupported gpu architecture 'compute_120a'

Opened 5 months ago
Ezzz-dev
open
No labels
205 months ago

#48 No Difference in tokens/sec - Ministral3 8B Q5_K_M

Opened 5 months ago
MrMuhannadObeidat
open
No labels
315 months ago

Rows per page:

1–25 of 69