Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-08-16 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-15 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-14 | 0 | 0 | 0 | 0 | 14 |
| 2026-08-13 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-11 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-10 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-09 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-08 | 1 | 0 | 0 | 0 | 0 |
| 2026-08-07 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-06 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-05 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-04 | 3 | 0 | 0 | 0 | 0 |
| 2026-08-03 | 0 | 3 | 0 | 0 | 0 |
Opened: 0
Closed: 0
Comments: 0
Events: 0
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#148 Request: Bonsai conversion of ThinkingCap-Qwen3.6-27B (halves thinking tokens, composes with ternary) Opened 9 days ago | rcfa | open | No labels | 0 | 2 | 9 days ago |
#147 [27B multimodal] Repeated 14K prompts get no prefix-cache reuse Opened 13 days ago | ZeroEgoist | open | No labels | 2 | 0 | 12 days ago |
#137 Please consider making a mainline compatible dspark draft model or publish some advice on that Opened 19 days ago | gf-mse | open | No labels | 2 | 1 | 12 days ago |
#146 [27B/OpenAI API] Clarify reasoning defaults and per-request opt-out for agentic tool calls Opened 13 days ago | ZeroEgoist | open | No labels | 0 | 0 | 13 days ago |
#145 [CUDA/mainline] Q4 KV causes a 3x-67x prompt-processing slowdown on Ternary-Bonsai-27B Opened 13 days ago | ZeroEgoist | open | No labels | 0 | 0 | 13 days ago |
#81 GLM 5.2 version ? Opened 1 month ago | X-Ryl669 | closed - completed | No labels | 2 | 0 | 13 days ago |
#142 Just comment: bonsai 1.7B on haswell i5 iMac Opened 15 days ago | Aladincz | open | No labels | 2 | 1 | 13 days ago |
#92 Ternary-Bonsai-27B on ROCm / RDNA3: works unpatched (RX 7800 XT, gfx1101) + benchmarks Opened 1 month ago | base-zz | closed - completed | No labels | 5 | 2 | 13 days ago |
#102 Together serverless Ternary Bonsai 27B endpoint metadata points at an AWQ-4bit artifact, catalog says ternary g128 Opened 1 month ago | Astezelex | closed - completed | No labels | 1 | 0 | 13 days ago |
#139 Combination of BONSAI_KV4 and BONSAI_SPECULATIVE on 27B Ternary reports issue with KV cache tpye Opened 17 days ago | gdevenyi | open | No labels | 2 | 1 | 16 days ago |
#114 One-click Windows launcher for Bonsai-demo (community tool, not a PR) Opened 29 days ago | Basearchio | open | No labels | 1 | 1 | 17 days ago |
#131 It is stuck in a loop when used with openclaw model Ternary-Bonsai-27B-Q2_0 Opened 24 days ago | SaitamaTechno | open | No labels | 1 | 0 | 18 days ago |
#130 Bonsai Alternatives Opened 24 days ago | LionelColaso | open | No labels | 1 | 0 | 21 days ago |
#101 Any plans to release the ternary vLLM kernels used in the whitepaper evals? Opened 1 month ago | Astezelex | closed - completed | No labels | 1 | 0 | 25 days ago |
#125 --no-mmproj-offload should be specified for vision Opened 26 days ago | gdevenyi | closed - completed | No labels | 4 | 0 | 25 days ago |
#105 DSpark drafter: measured 32k enablement ceiling + net-negative speedup at depth (mechanism ≠ VRAM) Opened 1 month ago | cnndabbler | open | No labels | 4 | 0 | 25 days ago |
#121 Chance of Qwen 3.8 and/or KIMI K3? Opened 27 days ago | Saknutella | closed - completed | No labels | 1 | 0 | 25 days ago |
#124 Context should not be specified with `-c 0` otherwise fitting fails. Opened 26 days ago | gdevenyi | closed - completed | No labels | 3 | 0 | 25 days ago |
#86 Ornith-1.0 9B version? Opened 1 month ago | hrstoyanov | closed - completed | No labels | 2 | 1 | 27 days ago |
#115 bug(deps): open-webui 0.10.2 routers/audio requires Python *strictly* <3.13 (or traceback) Opened 28 days ago | tildebyte | closed - completed | bug | 2 | 0 | 27 days ago |
#93 Pre-built llama.cpp binaries: Metal fails to compile on macOS 26 Tahoe (M5) Opened 1 month ago | lightoshadow | closed - completed | No labels | 8 | 0 | 27 days ago |
#112 Speculative decoding (draft-dspark) silently falls back to CPU on DGX Spark / unified-memory CUDA devices — 2.35x slower than fixed Opened 29 days ago | thadreber-web | closed - completed | No labels | 1 | 1 | 28 days ago |
#85 is iphone's NPU utilized in Bonsai inference? Opened 1 month ago | ColdCodeCool | closed - completed | No labels | 1 | 0 | 28 days ago |
#113 Reverts to GPU inference when running it on AMD CPU with integrated graphics (only affects llama server) Opened 29 days ago | squirrellyidiot | closed - completed | No labels | 0 | 0 | 28 days ago |
#95 Question: Exact PrismML-MLX commit used for Bonsai-27B-mlx-1bit release Opened 1 month ago | twinfufu | closed - completed | No labels | 3 | 0 | 28 days ago |