Repository Issue Activity (beta)

deepseek-ai/deepspec

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
26
New in 7 Days
2
Closed in 7 Days
1
Average Open Age
37 days
Stale 30+ Days
18
Stale 90+ Days
0
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-08-1500000
2026-08-14000026
2026-08-1300000
2026-08-1200000
2026-08-1100000
2026-08-1021000
2026-08-0900000
2026-08-0800000
2026-08-0700000
2026-08-0600000
2026-08-0500000
2026-08-0400000
2026-08-0300000
2026-08-0200000
This Week

Opened: 2

Closed: 1

Comments: 0

Events: 0

Top Labels

No label distribution is available yet.

Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#81 Target feature selection for DSpark

Opened 5 days ago
sleepyshep
open
No labels
005 days ago

#80 Target feature selection for DSpark on hybrid GDN models

Opened 5 days ago
sleepyshep
closed - completed
No labels
005 days ago

#21 Support online training without precomputing the target cache

Opened 2 months ago
Ofir408
open
No labels
1010 days ago

#77 [Off-topic] Question for Damai Dai / Chengqi Deng: has profile-based MoE routing been explored?

Opened 14 days ago
washingtoneimae-dot
open
No labels
0014 days ago

#76 [Question] Why DFlash is worse than EAGLE3 for Gemma4-12B in paper's experiment

Opened 17 days ago
DeclK
open
No labels
0017 days ago

#73 推理效率异常

Opened 24 days ago
team109
closed - completed
No labels
6022 days ago

#71 DSpark Qwen3 draft modeling silently ignores `partial_rotary_factor`, producing checkpoints whose config mismatches the trained weights (breaks vLLM serving AL)

Opened 24 days ago
arthurgao2003
open
No labels
1023 days ago

#70 feat: support Multi-Token Prediction (MTP) training for LLM and speech modalities

Opened 25 days ago
pengyanai
open
No labels
0125 days ago

#49 一个关于生成范式的远期思考(非Bug报告)

Opened 1 month ago
madheu
open
No labels
11027 days ago

#57 Release speculators used for ablations

Opened 1 month ago
sheelfshah
open
No labels
121 month ago

#61 请问deepseek官方是否有针对新版本模型写作文笔能力修正?

Opened 1 month ago
9DHans
open
No labels
201 month ago

#52 [RFC] White-Box Architectural Reference for DSpark: Spherical Normalization, 3-Tier Semantic Cache & Unified Multimodal Token

Opened 1 month ago
Xuan-yi-yan
open
No labels
201 month ago

#64 [Proposal] PathOracle: Hidden-State Prediction Layer Skipping for Prefill Acceleration, Complementary to DSpark

Opened 1 month ago
xhy-h
open
No labels
001 month ago

#60 Feat: Compatible with Gamme4 MoE structure

Opened 1 month ago
waiting-xia
open
No labels
031 month ago

#53 Any plan for supporting DS v4 Flash/Pro DSpark Training?

Opened 1 month ago
singzhou
closed - completed
No labels
101 month ago

#56 Question about prefill mtp forward

Opened 1 month ago
MARD1NO
closed - completed
No labels
001 month ago

#51 reading my chats not funny

Opened 1 month ago
ezydubs
open
No labels
001 month ago

#43 Where is hardware-aware prefix scheduler implemented?

Opened 1 month ago
lshAlgorithm
closed - completed
No labels
401 month ago

#46 RFC: Support Domino speculative decoding in DeepSpec?

Opened 1 month ago
jianuo-huang
open
No labels
101 month ago

#5 我看到sglang和vllm都已提供该特性,是否对minimax等其他模型也具有同样推理提升?

Opened 2 months ago
HardenGale
open
No labels
301 month ago

#44 [Draft Checkpoint] eagle3_qwen3_4b_ttt7 hidden_size mismatch (config: 2560, weights: 5120)

Opened 1 month ago
joooooy1
closed - completed
No labels
101 month ago

#22 DSpark trained models for qwen3.5-27B

Opened 2 months ago
Ofir408
open
No labels
1101 month ago

#36 launch_sglang_server.sh ignores CUDA_VISIBLE_DEVICES and always launches GPUs 0-7

Opened 2 months ago
Oxygen56
open
No labels
002 months ago

#35 Concurrent training data generation can corrupt resume order and buffer unbounded results

Opened 2 months ago
morluto
open
No labels
002 months ago

#34 DSpark train/loss metric is rank-weighted instead of globally weighted

Opened 2 months ago
morluto
open
No labels
002 months ago

Rows per page:

1–25 of 34