Repository Issue Activity (beta)

antirez/ds4

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
167
New in 7 Days
50
Closed in 7 Days
12
Average Open Age
34 days
Stale 30+ Days
73
Stale 90+ Days
0
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-08-0760313211
2026-08-06408813
2026-08-0581152715
2026-08-04101021014
2026-08-0370133116
2026-08-0260719141
2026-08-0170000
2026-07-3121000
2026-07-3000000
2026-07-2940000
2026-07-2820000
2026-07-2710000
2026-07-2650000
2026-07-2520000
This Week

Opened: 48

Closed: 11

Comments: 76

Events: 127

Top Labels
cuda (6)
bug (4)
ideas-to-explore (3)
priority (3)
speed (3)
correctness (2)
kv-cache (2)
metal (2)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#736 OpenVino support

Opened 9 hours ago
DealsBeam
open
No labels
009 hours ago

#735 ds4-server failed to run with 4xA100/80G (prefill failed)

Opened 10 hours ago
snailwei
open
No labels
109 hours ago

#734 CUDA --ssd-streaming: prefill dies with illegal memory access at exactly 128 prompt tokens

Opened 13 hours ago
pmasala
open
No labels
0011 hours ago

#732 CUDA --ssd-streaming: MMQ prefill tier OOMs on a 12 GB card (main b030961)

Opened 15 hours ago
pmasala
open
No labels
0011 hours ago

#733 DSpark net-negative on M3 Ultra + MXFP4 despite 82.8% accept rate — independent confirmation of #695 replay/verify overhead

Opened 14 hours ago
lifodetails
open
No labels
0014 hours ago

#629 CUDA SSD streaming: host-RAM tier for routed experts (working branch, 1.65-3.1x decode)

Opened 9 days ago
pmasala
open
No labels
6016 hours ago

#521 gfx1201 (RDNA4, R9700) build fails: __builtin_amdgcn_wmma_f32_16x16x16_f16_w32 not supported

Opened 1 month ago
cm999club
open
No labels
6018 hours ago

#731 DSpark: per-accepted-token replay cancels ~100% of the speculative saving (M3 Ultra measurements, net_saved = -3.98 s)

Opened 19 hours ago
justinz0920-png
open
No labels
6018 hours ago

#639 --ssd-streaming-cache-experts is silently inert on CUDA — expert cache never populated, hit rate 0

Opened 7 days ago
nexus-cw
open
No labels
12018 hours ago

#726 M3 Ultra 512 Testing results with DSv4F 0731

Opened 1 day ago
trueimage
open
No labels
5018 hours ago

#658 DSpark speculative decode breaks --temp 0 greedy identity (batch-verify vs single-token KV numerics)

Opened 5 days ago
nexus-cw
closed - completed
No labels
401 day ago

#724 DSpark proposals systematically wrong under --ssd-streaming (DeepSeek-V4-Flash-0731, Metal): enabling patch observed, miss_first=100%

Opened 1 day ago
johnpippett
open
No labels
101 day ago

#695 DSpark scheduler break-even model appears not to account for replay cost (Metal, net -1327ms despite 70% accept rate)

Opened 3 days ago
datanerdie
open
No labels
601 day ago

#705 ds4f-q2-q4 cannot be served resident on 128 GB DGX Spark (GB10) — global OOM during startup span preparation

Opened 2 days ago
adamlawi
open
No labels
201 day ago

#721 CUDA loader: Q4-class GGUF load transient is ~1.8× mapped tensor size on unified memory (GB10) — freezes 128 GB DGX Spark

Opened 1 day ago
snhwang
open
No labels
011 day ago

#719 Show and tell: DSBrain (macOS menu bar for ds4-server)

Opened 2 days ago
derkan
open
No labels
012 days ago

#496 GLM 5.2 error: `ds4: required tensor is missing: token_embd.weight`

Opened 1 month ago
links486
open
No labels
102 days ago

#693 gguf-tools splicer cannot read the MXFP4 GGUF this repo publishes (unsupported GGML tensor type 39)

Opened 3 days ago
datanerdie
open
No labels
202 days ago

#713 Laguna server batch prefill fails at >=7k tokens: routed MoE intermediate quantize launch invalid argument

Opened 2 days ago
nexus-cw
open
No labels
002 days ago

#685 ds4-server reasoning_content deltas fragmented mid-token at arbitrary byte offsets (OpenAI-compatible streaming)

Opened 3 days ago
cylentsec
open
No labels
102 days ago

#660 New 0731 iq2 thinks forever!

Opened 5 days ago
arkham000
open
No labels
302 days ago

#444 Disk KV eviction preferentially removes short-prefix cache files causing disk cache miss & full pre-fill when Disk KV Store is full

Opened 2 months ago
srinathh
open
No labels
622 days ago

#698 Extend live-prefix rewind (#668) beyond GLM: DeepSeek V4 Flash on Metal still cold-prefills exact-prefix retries

Opened 3 days ago
Snail3D
open
No labels
102 days ago

#691 KV cache reuse breaks for OpenAI tools clients that don't replay reasoning_content

Opened 3 days ago
hennas-waifson
open
No labels
102 days ago

#585 CUDA OOM regression on DGX Spark (GB10)

Opened 18 days ago
xXRaidohXx
open
No labels
412 days ago

Rows per page:

1–25 of 290