Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-07 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-06 | 0 | 5 | 1 | 6 | 1 |
| 2026-09-05 | 0 | 0 | 0 | 1 | 270 |
| 2026-09-04 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-03 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-02 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-01 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-31 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-30 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-29 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-28 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-27 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-26 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-25 | 0 | 0 | 0 | 0 | 0 |
Opened: 1
Closed: 5
Comments: 1
Events: 7
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#2231 Request for Gemma 3 4B & Gemma 4 Integration Support with llama-cpp-python Opened 4 months ago | LeConsulat2 | open | No labels | 2 | 0 | 1 day ago |
#1813 [Feature request] High-level API support for DRY and XTC samplers Opened 2 years ago | ddh0 | closed - completed | No labels | 3 | 6 | 1 day ago |
#1495 _LlamaModel.metadata() does not return `tokenizer.ggml.tokens` Opened 2 years ago | ddh0 | closed - not_planned | No labels | 7 | 0 | 1 day ago |
#1360 [REQUEST] Accept raw token IDs in `stop` parameter Opened 2 years ago | ddh0 | closed - completed | enhancement | 11 | 4 | 1 day ago |
#1335 KV cache quantization fails with GGML_ASSERT Opened 2 years ago | ddh0 | closed - completed | bug upstream | 1 | 0 | 1 day ago |
#878 Possible bug with greedy sampling code: Llama._sample() Opened 3 years ago | ddh0 | closed - not_planned | bug | 1 | 0 | 1 day ago |
#2362 Long whitespace tokens detokenize to empty bytes, breaking tokenize/detokenize round-trip Opened 3 days ago | gbjones | open | No labels | 0 | 0 | 3 days ago |
#1366 Pip Install faling in colab Opened 2 years ago | Orneyfish | closed - completed | No labels | 9 | 3 | 8 days ago |
#2361 [Bug] DFlash 2 speculative decoding crashes and fails in Python due to missing ctx_other and non-autoregressive graph execution Opened 18 days ago | ilperev | open | No labels | 1 | 0 | 18 days ago |
#2342 Llama.close() doesn't clean up chat_handler causing null pointer on second load of a vision model Opened 2 months ago | WarmneoN | closed - completed | No labels | 1 | 0 | 21 days ago |
#2013 Can't install with GPU support with Cuda toolkit 12.9 and Cuda 12.9 Opened 1 year ago | hunainahmedj | open | No labels | 22 | 8 | 22 days ago |
#2352 uv add llama-cpp-python wheels fails for versions above 0.3.30 Opened 1 month ago | Karthik777 | open | No labels | 0 | 0 | 1 month ago |
#1169 Docker llama-cpp libcuda.so.1: cannot open shared object file: No such file or directory Opened 3 years ago | Apotrox | open | No labels | 15 | 0 | 2 months ago |
#2341 Support sm100 and sm120 in CUDA pre‑built wheels Opened 2 months ago | XuehaoSun | open | No labels | 0 | 1 | 2 months ago |
#2029 Access Violation issue facing for exe created using pyinstaller Opened 1 year ago | maniron214 | open | No labels | 5 | 0 | 2 months ago |
#2211 Llama.embed() calls LlamaBatch.add_sequence with old 3-arg signature; missing logits_array Opened 4 months ago | emptyngton | open | No labels | 2 | 0 | 2 months ago |
#1290 OpenAI compatible API not able to connect to llama.cpp web server Opened 2 years ago | hpxiong | closed - completed | bug | 10 | 0 | 2 months ago |
#247 No CUDA toolset found. Opened 3 years ago | bfgball | closed - not_planned | build windows | 17 | 0 | 2 months ago |
#2323 I am passionate for Edge AI and would like to try out Qwen 3.5 2B model using llama-cpp-python. But as of now, support for Qwen 3.5 is pending. I saw there's a PR (https://github.com/abetlen/llama-cpp-python/pull/2132) adding support. Any idea when will this be merged? Opened 2 months ago | imrarethatimaware-alt | closed - completed | No labels | 0 | 1 | 2 months ago |
#635 Documentation of server command line parameters. Opened 3 years ago | arthurwolf | open | documentation | 8 | 0 | 3 months ago |
#2175 How to use "Gemma-4-E4B-it-heretic-GGUF" ? Opened 5 months ago | TimmyHeart | open | No labels | 3 | 1 | 3 months ago |
#1095 Output the final settings used to load the llm Opened 3 years ago | vriesdemichael | closed - not_planned | enhancement | 0 | 1 | 3 months ago |
#1938 Specifying additional_files for model files in directory adds additional copy of directory to download URL Opened 2 years ago | zhudotexe | closed - completed | No labels | 0 | 2 | 3 months ago |
#2103 Pre-built wheels for Python 3.14 and 3.14 free-threaded Opened 9 months ago | clemlesne | closed - completed | No labels | 4 | 2 | 3 months ago |
#2224 Feature Request: Gemma Multimodal Support Opened 4 months ago | rycerzes | closed - completed | No labels | 0 | 0 | 3 months ago |