Repository Issue Activity (beta)

ikawrakow/ik_llama.cpp

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
60
New in 7 Days
10
Closed in 7 Days
3
Average Open Age
60 days
Stale 30+ Days
41
Stale 90+ Days
25
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-08-2300221
2026-08-2210111
2026-08-2100002
2026-08-2010001
2026-08-1911012
2026-08-1800001
2026-08-1721335
2026-08-1651573
2026-08-1521231
2026-08-1422003
2026-08-131711213
2026-08-1230111
2026-08-1101121
2026-08-1031111
This Week

Opened: 5

Closed: 2

Comments: 6

Events: 7

Top Labels
enhancement (101)
wontfix (24)
bug (10)
help wanted (7)
Refactoring (2)
mainline bug (2)
Usability (1)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#2344 DeepSeek-V4-Flash: all-NaN logits abort survives the f32 DSA fix (#2311), and llama-server's own handler for it can never run

Opened 3 days ago
daimonionnn
open
No labels
409 hours ago

#2346 Bug: LLGuidance fails to compile against ik_llama.cpp due to API divergence from upstream

Opened 2 days ago
freggolopogit
open
No labels
102 days ago

#2336 Bug: Server crashes with Gemma4 when reaching the end of context window

Opened 7 days ago
aadametz
open
No labels
003 days ago

#2335 Bug: [Bug] Qwen3.8 MTP cold-start: first long reasoning stalls, warm-up fixes it

Opened 7 days ago
Aoike123
open
No labels
304 days ago

#2338 CUDA illegal memory access (ggml_cuda_cpy_dest_ptrs_copy) on hybrid-recurrent Qwen3.8-27B with --ctx-checkpoints + -sm graph across 4 GPUs

Opened 7 days ago
gopinath87607
closed - completed
No labels
104 days ago

#2343 Bug: CUDA crash "invalid argument" in `ggml_cuda_flash_attn_ext_mma_f16_case` when running MLA models (576/512, 320/256) on Turing GPUs

Opened 5 days ago
hgeistcomtest
open
No labels
005 days ago

#2203 Feature Request: Kimi K3 model support

Opened 25 days ago
VanceVagell
open
enhancement
505 days ago

#2320 Bug: --cache-ram size limit is not enforced?

Opened 9 days ago
Skelectric
open
No labels
306 days ago

#2155 Slow MXFP4 token generation speed

Opened 1 month ago
cora4
closed - completed
No labels
406 days ago

#2340 Feature Request: llama-server with unsloth studio

Opened 6 days ago
cora4
closed - not_planned
enhancement
006 days ago

#2333 Bug: Unhandled type iq1_m (29) in ggml-cuda/mmq.cuh:112

Opened 7 days ago
giftick
open
No labels
007 days ago

#2331 Feature Request: add dots3-note

Opened 7 days ago
gopinath87607
open
enhancement
007 days ago

#2325 Feature Request: Improve AMD/Vulkan support for IQ4_KS/IQ4_KT quants

Opened 8 days ago
guy915
open
enhancement
108 days ago

#2329 Bug: Qwen 3.8 Thought loops

Opened 8 days ago
frenzybiscuit
closed - completed
No labels
208 days ago

#2318 Bug: -sm graph + MTP speculative decoding fails with Qwen3.8 27B

Opened 9 days ago
joonanykanen
closed - completed
No labels
209 days ago

#2317 DSV4: pos-11 argmax is GPU-architecture dependent on Blackwell (sm_120 picks the pre-#2311 branch; same binary on sm_86 does not)

Opened 9 days ago
rumas77
closed - completed
No labels
309 days ago

#2301 DSV4: forward pass logits diverge from mainline on identical input, degrading DSML tool-call reliability (deterministic repro)

Opened 11 days ago
rumas77
closed - completed
No labels
1309 days ago

#437 Feature Request: support intel amx for further accelerate

Opened 1 year ago
zhaoyukoon
open
enhancement
103310 days ago

#1769 Mimo V2.5 Pro GGUF Fails to Load

Opened 4 months ago
yimbin
open
No labels
17010 days ago

#2288 DeepSeek-V4: most speculative stages crash on long (~12K) prompts — cublasSgemm invalid parameter at ggml-cuda.cu:1882 (ngram-simple unaffected)

Opened 13 days ago
rumas77
closed - completed
No labels
9010 days ago

#2300 DSV4: CPU-only inference asserts nth*work_size <= wsize (ggml.c:24480) during prefill

Opened 11 days ago
rumas77
closed - completed
No labels
1010 days ago

#2183 Bug: Phi-3 GGUFs without phi3.attention.sliding_window fail to load (the compat fallback is unreachable)

Opened 29 days ago
joaomdsg
closed - completed
No labels
0011 days ago

#2185 Bug: mellum + -sm graph aborts at graph build: GGML_ASSERT(nhave > 1) in ggml_reduce

Opened 29 days ago
joaomdsg
closed - completed
No labels
1111 days ago

#2282 Feature Request: adding support for Ling-3.0-flash

Opened 15 days ago
gopinath87607
closed - completed
enhancement
1311 days ago

#2284 Restore slot file does not work

Opened 14 days ago
mohsenbgi
closed - completed
No labels
2011 days ago

Rows per page:

1–25 of 610