Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-08-15 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-14 | 0 | 0 | 0 | 0 | 26 |
| 2026-08-13 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-11 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-10 | 2 | 1 | 0 | 0 | 0 |
| 2026-08-09 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-08 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-07 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-06 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-05 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-04 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-03 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-02 | 0 | 0 | 0 | 0 | 0 |
Opened: 2
Closed: 1
Comments: 0
Events: 0
No label distribution is available yet.
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#81 Target feature selection for DSpark Opened 5 days ago | sleepyshep | open | No labels | 0 | 0 | 5 days ago |
#80 Target feature selection for DSpark on hybrid GDN models Opened 5 days ago | sleepyshep | closed - completed | No labels | 0 | 0 | 5 days ago |
#21 Support online training without precomputing the target cache Opened 2 months ago | Ofir408 | open | No labels | 1 | 0 | 10 days ago |
#77 [Off-topic] Question for Damai Dai / Chengqi Deng: has profile-based MoE routing been explored? Opened 14 days ago | washingtoneimae-dot | open | No labels | 0 | 0 | 14 days ago |
#76 [Question] Why DFlash is worse than EAGLE3 for Gemma4-12B in paper's experiment Opened 17 days ago | DeclK | open | No labels | 0 | 0 | 17 days ago |
#73 推理效率异常 Opened 24 days ago | team109 | closed - completed | No labels | 6 | 0 | 22 days ago |
#71 DSpark Qwen3 draft modeling silently ignores `partial_rotary_factor`, producing checkpoints whose config mismatches the trained weights (breaks vLLM serving AL) Opened 24 days ago | arthurgao2003 | open | No labels | 1 | 0 | 23 days ago |
#70 feat: support Multi-Token Prediction (MTP) training for LLM and speech modalities Opened 25 days ago | pengyanai | open | No labels | 0 | 1 | 25 days ago |
#49 一个关于生成范式的远期思考(非Bug报告) Opened 1 month ago | madheu | open | No labels | 11 | 0 | 27 days ago |
#57 Release speculators used for ablations Opened 1 month ago | sheelfshah | open | No labels | 1 | 2 | 1 month ago |
#61 请问deepseek官方是否有针对新版本模型写作文笔能力修正? Opened 1 month ago | 9DHans | open | No labels | 2 | 0 | 1 month ago |
#52 [RFC] White-Box Architectural Reference for DSpark: Spherical Normalization, 3-Tier Semantic Cache & Unified Multimodal Token Opened 1 month ago | Xuan-yi-yan | open | No labels | 2 | 0 | 1 month ago |
#64 [Proposal] PathOracle: Hidden-State Prediction Layer Skipping for Prefill Acceleration, Complementary to DSpark Opened 1 month ago | xhy-h | open | No labels | 0 | 0 | 1 month ago |
#60 Feat: Compatible with Gamme4 MoE structure Opened 1 month ago | waiting-xia | open | No labels | 0 | 3 | 1 month ago |
#53 Any plan for supporting DS v4 Flash/Pro DSpark Training? Opened 1 month ago | singzhou | closed - completed | No labels | 1 | 0 | 1 month ago |
#56 Question about prefill mtp forward Opened 1 month ago | MARD1NO | closed - completed | No labels | 0 | 0 | 1 month ago |
#51 reading my chats not funny Opened 1 month ago | ezydubs | open | No labels | 0 | 0 | 1 month ago |
#43 Where is hardware-aware prefix scheduler implemented? Opened 1 month ago | lshAlgorithm | closed - completed | No labels | 4 | 0 | 1 month ago |
#46 RFC: Support Domino speculative decoding in DeepSpec? Opened 1 month ago | jianuo-huang | open | No labels | 1 | 0 | 1 month ago |
#5 我看到sglang和vllm都已提供该特性,是否对minimax等其他模型也具有同样推理提升? Opened 2 months ago | HardenGale | open | No labels | 3 | 0 | 1 month ago |
#44 [Draft Checkpoint] eagle3_qwen3_4b_ttt7 hidden_size mismatch (config: 2560, weights: 5120) Opened 1 month ago | joooooy1 | closed - completed | No labels | 1 | 0 | 1 month ago |
#22 DSpark trained models for qwen3.5-27B Opened 2 months ago | Ofir408 | open | No labels | 1 | 10 | 1 month ago |
#36 launch_sglang_server.sh ignores CUDA_VISIBLE_DEVICES and always launches GPUs 0-7 Opened 2 months ago | Oxygen56 | open | No labels | 0 | 0 | 2 months ago |
#35 Concurrent training data generation can corrupt resume order and buffer unbounded results Opened 2 months ago | morluto | open | No labels | 0 | 0 | 2 months ago |
#34 DSpark train/loss metric is rank-weighted instead of globally weighted Opened 2 months ago | morluto | open | No labels | 0 | 0 | 2 months ago |