Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-07 | 3 | 6 | 2 | 15 | 5 |
| 2026-09-06 | 1 | 1 | 3 | 6 | 2 |
| 2026-09-05 | 0 | 4 | 10 | 14 | 7 |
| 2026-09-04 | 2 | 2 | 4 | 9 | 4 |
| 2026-09-03 | 0 | 2 | 5 | 12 | 490 |
| 2026-09-02 | 0 | 2 | 0 | 0 | 0 |
| 2026-09-01 | 3 | 48 | 0 | 0 | 0 |
| 2026-08-31 | 3 | 0 | 0 | 0 | 0 |
| 2026-08-30 | 2 | 1 | 0 | 0 | 0 |
| 2026-08-29 | 0 | 1 | 0 | 0 | 0 |
| 2026-08-28 | 1 | 3 | 0 | 0 | 0 |
| 2026-08-27 | 2 | 1 | 0 | 0 | 0 |
| 2026-08-26 | 1 | 3 | 0 | 0 | 0 |
| 2026-08-25 | 1 | 0 | 0 | 0 | 0 |
Opened: 9
Closed: 65
Comments: 24
Events: 56
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#10058 [RFC] Reward model (value-head RM) support in infer/deploy via vLLM Opened 9 hours ago | rrrxxx0510 | open | No labels | 0 | 1 | 9 hours ago |
#10057 About qwen3.5 mtp sft, multi-modality data Opened 10 hours ago | guoyingying432 | open | question | 0 | 0 | 10 hours ago |
#9973 enable_channel_loss crashes with IndexError when combined with sequence parallelism (sequence_parallel_size > 1) Opened 15 days ago | lxc1851588 | closed - completed | enhancement | 1 | 0 | 11 hours ago |
#10050 `resume_from_checkpoint` loses the epoch seed: the rebuilt sampler replays the epoch-0 order, so already-trained samples are re-trained Opened 1 day ago | Si1w | closed - completed | bug | 0 | 0 | 12 hours ago |
#9991 GRPO训练,用colocate模式在rollout阶段会出现模型参数或者优化器参数没有被offload导致oom的问题 Opened 13 days ago | 8125345 | closed - completed | question | 7 | 0 | 16 hours ago |
#10053 support for deepseek v4 flash vision exp Opened 17 hours ago | dsplog | open | enhancement | 0 | 0 | 17 hours ago |
#8256 [Bug] predict_with_generate=True causes 1D dimension collapse of eval_prediction.predictions in custom metrics (DDP) Opened 6 months ago | coldchair | closed - completed | bug | 3 | 0 | 20 hours ago |
#9892 支持Ling-3.0-flash和Ling-3.0-tiny和Muse-Glimmer-30B Opened 28 days ago | chuxiliyixiaosa | closed - completed | enhancement | 11 | 0 | 21 hours ago |
#10041 Qwen3-Omni使用add_special_tokens参数sft报错 Opened 3 days ago | zhao-huan | closed - completed | bug | 2 | 0 | 23 hours ago |
#9508 Example for training on SWE (agentic software-engineering) tasks? Opened 3 months ago | dipta007 | open | question stale | 3 | 0 | 24 hours ago |
#9518 Logging Per-Sample Evaluation Loss During Inference in Swift PT for Qwen3-VL Opened 3 months ago | somesh2002 | open | question stale | 1 | 0 | 24 hours ago |
#9240 Qwen3.5 9B async模式GRPO指定freeze vit时训练挂掉 Opened 4 months ago | zhenfenxiao | open | bug stale | 3 | 0 | 2 days ago |
#9371 fp8_param_gather does not reduce VRAM consumation comparing with fp16 in megatron Opened 4 months ago | edgeinfinity1 | closed - not_planned | bug stale | 3 | 0 | 2 days ago |
#9507 Multimodal SFT silently hangs forever on certain clips: load_audio (librosa→audioread) and decord video decode have no timeout (DataLoader deadlock) Opened 3 months ago | LukeLIN-web | open | stale | 1 | 0 | 2 days ago |
#10013 [NPU][Megatron GRPO] 开启 PP 后 P2P 通信在 HcclGroupEnd 报错并导致进程崩溃 Opened 8 days ago | mayunaise | open | bug | 1 | 0 | 3 days ago |
#10040 使用npu进行qwen35 27B megatron sft时,contex parallel设置是否能正确生效呢 Opened 3 days ago | tusiqi1 | closed - completed | question | 1 | 0 | 3 days ago |
#10024 Bug: unexpected behavior with special characters in input Opened 7 days ago | Groudonbrulant | open | No labels | 2 | 0 | 3 days ago |
#10022 求教:--loss_scale ignore_empty_think是否会影响SFT的think/no-think的混合训练? Opened 7 days ago | Ray-sjtu | open | question | 1 | 0 | 3 days ago |
#10014 Image augmentation is a bit tricky, tied to models, and hard to do for training data only Opened 8 days ago | sliedes | open | question | 2 | 0 | 3 days ago |
#10006 Data containing multiple images is not supported for the Qwen3.5 series. Opened 9 days ago | Skyer19 | open | bug | 1 | 0 | 3 days ago |
#10016 Data loaders somehow pull in CUDA Opened 7 days ago | sliedes | closed - completed | bug | 0 | 0 | 3 days ago |
#9399 dist_muon训练时卡住 Opened 4 months ago | JuntaoLiu01 | closed - not_planned | question stale | 5 | 0 | 3 days ago |
#9454 grpo多轮训练中,没有设置dynamic_sample,但是报错Padding free mode is not supported for dynamic sample Opened 3 months ago | fyw1999 | closed - not_planned | bug stale | 3 | 0 | 3 days ago |
#9435 【NPU】swift grpo qwen3-32b,开启sleep_level和vllm_tensor_parallel_size后报错 Opened 3 months ago | leuitong | closed - not_planned | bug stale | 2 | 0 | 4 days ago |
#9436 Consider using pre-built Flash Attention kernels via `kernels` Opened 3 months ago | sayakpaul | closed - not_planned | stale | 2 | 3 | 4 days ago |