Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-08-24 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-23 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-22 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-21 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-20 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-19 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-18 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-17 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-16 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-15 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-14 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-13 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-11 | 0 | 0 | 0 | 0 | 0 |
Opened: 0
Closed: 0
Comments: 0
Events: 0
No label distribution is available yet.
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#727 Eval datasets for reasoning Opened 2 months ago | connerlambden | closed - completed | No labels | 1 | 0 | 2 months ago |
#685 Interest in dataset based on ACM SGU? Opened 1 year ago | radoslav11 | closed - completed | No labels | 3 | 0 | 2 months ago |
#723 New aproach to developing AGI Opened 3 months ago | Kwoargus | open | No labels | 0 | 0 | 3 months ago |
#343 in SFT script, distributed training got stuck if set `packing=false` Opened 2 years ago | ChenDRAG | open | No labels | 9 | 1 | 5 months ago |
#692 RuntimeError: The size of tensor a (0) must match the size of tensor b (5120) at non-singleton dimension 1 Opened 1 year ago | george1459 | open | No labels | 1 | 0 | 5 months ago |
#719 Dependency conflicts Opened 6 months ago | DisGuiser9 | open | No labels | 0 | 0 | 6 months ago |
#341 Bug in SFT script Opened 2 years ago | ChenDRAG | open | No labels | 15 | 0 | 7 months ago |
#239 Why does the loss start at 0 when I train GRPO, and then possibly increase? Opened 2 years ago | hellen9527 | closed - completed | No labels | 41 | 0 | 7 months ago |
#715 Is there a Chinese version of the dataset in this project? Opened 8 months ago | fxb392 | open | No labels | 0 | 0 | 8 months ago |
#714 grad_norm nan Opened 9 months ago | weiyezhimeng | closed - completed | No labels | 2 | 0 | 9 months ago |
#713 grad_norm nan Opened 9 months ago | weiyezhimeng | closed - completed | No labels | 0 | 0 | 9 months ago |
#712 grad_norm nan Opened 9 months ago | weiyezhimeng | closed - completed | No labels | 0 | 0 | 9 months ago |
#118 Facing OOM during DRPO stage Opened 2 years ago | Some-random | closed - completed | No labels | 2 | 0 | 9 months ago |
#710 SFT Trains the Model on the Entire Sequence Opened 10 months ago | RohollahHS | open | No labels | 0 | 1 | 10 months ago |
#707 Forward reward always 0 Opened 10 months ago | ZhaoyangLi-1 | open | No labels | 0 | 0 | 10 months ago |
#705 Checkpoint selection when using cot sft Opened 11 months ago | lzhptr | open | No labels | 0 | 0 | 11 months ago |
#704 why kl = nan when grpo train? Opened 1 year ago | uilstong | open | No labels | 0 | 0 | 1 year ago |
#702 Request for License Information for Mixture-of-Thoughts Dataset Opened 1 year ago | mmyymmyy321 | open | No labels | 0 | 0 | 1 year ago |
#694 GSPO update request Opened 1 year ago | LuozhuZhang | closed - completed | No labels | 0 | 2 | 1 year ago |
#403 Instead of rising steadily, the reward fluctuates wildly Opened 1 year ago | Hasuer | open | No labels | 11 | 0 | 1 year ago |
#699 Does the SFT Framework Support Fine-Tuning for DeepSeek-R1-Distill-Qwen-7B Opened 1 year ago | ZJUCQR | open | No labels | 0 | 0 | 1 year ago |
#698 Question about evaluating AIME24 Accuracy Opened 1 year ago | ZJUCQR | open | No labels | 0 | 0 | 1 year ago |
#695 vllm prepends two BOS for LLama Opened 1 year ago | wenquanlu | open | No labels | 0 | 0 | 1 year ago |
#684 Latest version of flash_attn no longer works with torch==2.6.0 Opened 1 year ago | george1459 | open | No labels | 2 | 15 | 1 year ago |
#444 How to increase the context window from 4k to 32k on qwen models ? Opened 1 year ago | Jeremmmyyyyy | closed - completed | No labels | 6 | 0 | 1 year ago |