Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-08-07 | 6 | 0 | 31 | 32 | 11 |
| 2026-08-06 | 4 | 0 | 8 | 8 | 13 |
| 2026-08-05 | 8 | 1 | 15 | 27 | 15 |
| 2026-08-04 | 10 | 10 | 2 | 10 | 14 |
| 2026-08-03 | 7 | 0 | 13 | 31 | 16 |
| 2026-08-02 | 6 | 0 | 7 | 19 | 141 |
| 2026-08-01 | 7 | 0 | 0 | 0 | 0 |
| 2026-07-31 | 2 | 1 | 0 | 0 | 0 |
| 2026-07-30 | 0 | 0 | 0 | 0 | 0 |
| 2026-07-29 | 4 | 0 | 0 | 0 | 0 |
| 2026-07-28 | 2 | 0 | 0 | 0 | 0 |
| 2026-07-27 | 1 | 0 | 0 | 0 | 0 |
| 2026-07-26 | 5 | 0 | 0 | 0 | 0 |
| 2026-07-25 | 2 | 0 | 0 | 0 | 0 |
Opened: 48
Closed: 11
Comments: 76
Events: 127
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#736 OpenVino support Opened 9 hours ago | DealsBeam | open | No labels | 0 | 0 | 9 hours ago |
#735 ds4-server failed to run with 4xA100/80G (prefill failed) Opened 10 hours ago | snailwei | open | No labels | 1 | 0 | 9 hours ago |
#734 CUDA --ssd-streaming: prefill dies with illegal memory access at exactly 128 prompt tokens Opened 13 hours ago | pmasala | open | No labels | 0 | 0 | 11 hours ago |
#732 CUDA --ssd-streaming: MMQ prefill tier OOMs on a 12 GB card (main b030961) Opened 15 hours ago | pmasala | open | No labels | 0 | 0 | 11 hours ago |
#733 DSpark net-negative on M3 Ultra + MXFP4 despite 82.8% accept rate — independent confirmation of #695 replay/verify overhead Opened 14 hours ago | lifodetails | open | No labels | 0 | 0 | 14 hours ago |
#629 CUDA SSD streaming: host-RAM tier for routed experts (working branch, 1.65-3.1x decode) Opened 9 days ago | pmasala | open | No labels | 6 | 0 | 16 hours ago |
#521 gfx1201 (RDNA4, R9700) build fails: __builtin_amdgcn_wmma_f32_16x16x16_f16_w32 not supported Opened 1 month ago | cm999club | open | No labels | 6 | 0 | 18 hours ago |
#731 DSpark: per-accepted-token replay cancels ~100% of the speculative saving (M3 Ultra measurements, net_saved = -3.98 s) Opened 19 hours ago | justinz0920-png | open | No labels | 6 | 0 | 18 hours ago |
#639 --ssd-streaming-cache-experts is silently inert on CUDA — expert cache never populated, hit rate 0 Opened 7 days ago | nexus-cw | open | No labels | 12 | 0 | 18 hours ago |
#726 M3 Ultra 512 Testing results with DSv4F 0731 Opened 1 day ago | trueimage | open | No labels | 5 | 0 | 18 hours ago |
#658 DSpark speculative decode breaks --temp 0 greedy identity (batch-verify vs single-token KV numerics) Opened 5 days ago | nexus-cw | closed - completed | No labels | 4 | 0 | 1 day ago |
#724 DSpark proposals systematically wrong under --ssd-streaming (DeepSeek-V4-Flash-0731, Metal): enabling patch observed, miss_first=100% Opened 1 day ago | johnpippett | open | No labels | 1 | 0 | 1 day ago |
#695 DSpark scheduler break-even model appears not to account for replay cost (Metal, net -1327ms despite 70% accept rate) Opened 3 days ago | datanerdie | open | No labels | 6 | 0 | 1 day ago |
#705 ds4f-q2-q4 cannot be served resident on 128 GB DGX Spark (GB10) — global OOM during startup span preparation Opened 2 days ago | adamlawi | open | No labels | 2 | 0 | 1 day ago |
#721 CUDA loader: Q4-class GGUF load transient is ~1.8× mapped tensor size on unified memory (GB10) — freezes 128 GB DGX Spark Opened 1 day ago | snhwang | open | No labels | 0 | 1 | 1 day ago |
#719 Show and tell: DSBrain (macOS menu bar for ds4-server) Opened 2 days ago | derkan | open | No labels | 0 | 1 | 2 days ago |
#496 GLM 5.2 error: `ds4: required tensor is missing: token_embd.weight` Opened 1 month ago | links486 | open | No labels | 1 | 0 | 2 days ago |
#693 gguf-tools splicer cannot read the MXFP4 GGUF this repo publishes (unsupported GGML tensor type 39) Opened 3 days ago | datanerdie | open | No labels | 2 | 0 | 2 days ago |
#713 Laguna server batch prefill fails at >=7k tokens: routed MoE intermediate quantize launch invalid argument Opened 2 days ago | nexus-cw | open | No labels | 0 | 0 | 2 days ago |
#685 ds4-server reasoning_content deltas fragmented mid-token at arbitrary byte offsets (OpenAI-compatible streaming) Opened 3 days ago | cylentsec | open | No labels | 1 | 0 | 2 days ago |
#660 New 0731 iq2 thinks forever! Opened 5 days ago | arkham000 | open | No labels | 3 | 0 | 2 days ago |
#444 Disk KV eviction preferentially removes short-prefix cache files causing disk cache miss & full pre-fill when Disk KV Store is full Opened 2 months ago | srinathh | open | No labels | 6 | 2 | 2 days ago |
#698 Extend live-prefix rewind (#668) beyond GLM: DeepSeek V4 Flash on Metal still cold-prefills exact-prefix retries Opened 3 days ago | Snail3D | open | No labels | 1 | 0 | 2 days ago |
#691 KV cache reuse breaks for OpenAI tools clients that don't replay reasoning_content Opened 3 days ago | hennas-waifson | open | No labels | 1 | 0 | 2 days ago |
#585 CUDA OOM regression on DGX Spark (GB10) Opened 18 days ago | xXRaidohXx | open | No labels | 4 | 1 | 2 days ago |