Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-09 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-08 | 1 | 0 | 2 | 3 | 2 |
| 2026-09-07 | 0 | 0 | 0 | 1 | 222 |
| 2026-09-06 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-05 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-04 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-03 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-02 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-01 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-31 | 2 | 0 | 0 | 0 | 0 |
| 2026-08-30 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-29 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-28 | 1 | 1 | 0 | 0 | 0 |
| 2026-08-27 | 0 | 0 | 0 | 0 | 0 |
Opened: 1
Closed: 0
Comments: 2
Events: 4
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#2370 [Bug] save_model fires unconditionally at the final step regardless of --save-interval; final optimizer-state write can kill the job Opened 1 day ago | kugua2025aigc-code | open | No labels | 1 | 0 | 1 day ago |
#1136 [BUG] 稳定训一半OOM,已开cp Opened 9 months ago | glennccc | open | No labels | 7 | 0 | 2 days ago |
#1487 Training hangs indefinitely after rollout phase completes on 8 B200 GPUs with CP=4, TP=2 Opened 8 months ago | euclidgame | open | No labels | 6 | 0 | 4 days ago |
#200 UX/Unintended bug: sample.reward is None on aborted sample Opened 1 year ago | casper-hansen | open | No labels | 5 | 0 | 5 days ago |
#2338 [Question] realign 使用本轮 response 长度判断 Opened 10 days ago | stivory0 | open | question | 1 | 0 | 5 days ago |
#397 if we use GRPO and args.kl_coef is non-zero, is the KL computation incorrect? Opened 1 year ago | adol001 | open | No labels | 2 | 0 | 5 days ago |
#1462 cannot import name 'GPU_MEMORY_TYPE_CUDA_GRAPH' from 'sglang.srt.constants' Opened 8 months ago | CHRIS123540 | open | No labels | 4 | 0 | 5 days ago |
#543 grade_answer_verl fails when rm_type="boxed_math" strips the \boxed{} wrapper upstream Opened 11 months ago | jsheng112 | open | No labels | 3 | 0 | 5 days ago |
#1369 【BUG】wrong implementation of zero-centered mean when normalizing advantages Opened 8 months ago | yaokl-nju | open | No labels | 1 | 0 | 5 days ago |
#2214 Your project is on StackMap — a curated map of the AI stack Opened 2 months ago | hoghweed | open | No labels | 3 | 0 | 6 days ago |
#2339 [Discussion] Support background evaluation while fully-async training rollout is running Opened 10 days ago | zhuyijie88 | open | question | 1 | 0 | 7 days ago |
#2344 [Bug] Fully-async multi-agent rollout crashes on nested groups Opened 8 days ago | looput | open | bug | 0 | 0 | 8 days ago |
#2176 [Bug] 过采样关停的时候judge不会被杀掉 Opened 2 months ago | Fu-Dayuan | closed - completed | bug | 3 | 0 | 13 days ago |
#2336 [Bug] Qwen3.5-VL grounding RL silently wrong: bundled SGLang image drops H/W rows of [3,T] M-RoPE (upstream sgl-project/sglang#35345, fix #35744) Opened 13 days ago | yszhli | open | No labels | 0 | 0 | 13 days ago |
#2324 如何才能快速训练GLM 5.2 Opened 16 days ago | zyh190507 | open | question | 1 | 0 | 14 days ago |
#2332 [Bug] W&B initialization fails with wandb 0.28.2 because wandb.util.generate_id was removed Opened 14 days ago | bcol23 | open | bug | 1 | 0 | 14 days ago |
#2329 [Question] Multiple critic updates per actor update (SAO-style faster value update)? Opened 15 days ago | ycoliver | open | question | 0 | 0 | 15 days ago |
#2285 [Question/Performance] Full message history is re-tokenized on every coding-agent turn Opened 22 days ago | zhuyijie88 | open | question | 3 | 0 | 15 days ago |
#2314 [Question] 什么时候支持ling3.0呢? Opened 17 days ago | jmgytc | open | question | 0 | 0 | 17 days ago |
#1111 Add cache to process inputs Opened 9 months ago | nanjiangwill | open | No labels | 2 | 0 | 17 days ago |
#2253 [Bug] Training/log-prob forward materializes full-vocab fp32 logits over the entire packed sequence → OOM for large-vocab, long multi-turn RL Opened 1 month ago | yc32768 | open | bug | 1 | 0 | 17 days ago |
#2188 [Bug] Colocate weight update fails with torch_memory_saver/offload because PyTorch CUDA IPC _share_cuda_ raises cudaErrorInvalidValue Opened 2 months ago | Winnie-Lian | open | bug | 2 | 0 | 17 days ago |
#2201 GLM-5.2 SparseMLA TileLang backward returns NaN gradients from finite inputs Opened 2 months ago | zhoutong-hai | open | No labels | 2 | 0 | 17 days ago |
#2209 Delta weight sync (NCCL) produces NaN weights on Qwen3.5-122B MoE → "probability tensor contains inf/nan" crash Opened 2 months ago | leofan-i | open | No labels | 1 | 0 | 17 days ago |
#2299 [Bug] Agent adapters rewrite malformed tool-call arguments before the client sees them Opened 20 days ago | EazyReal | open | No labels | 0 | 0 | 20 days ago |