Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-09 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-08 | 0 | 0 | 0 | 0 | 26 |
| 2026-09-07 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-06 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-05 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-04 | 1 | 1 | 0 | 0 | 0 |
| 2026-09-03 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-02 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-01 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-31 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-30 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-29 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-28 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-27 | 0 | 0 | 0 | 0 | 0 |
Opened: 2
Closed: 1
Comments: 0
Events: 0
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#1420 Several Unused Args Opened 4 days ago | aflah02 | open | bug | 0 | 0 | 4 days ago |
#1321 Can `preprocess_data.py` support Huggingface Dataset? Opened 2 years ago | cafeii | open | feature request | 2 | 0 | 4 days ago |
#1320 _forward_step_fn does not always return two values so eval.py breaks if is_pipe_parallel is false Opened 2 years ago | markNZed | closed - completed | bug | 2 | 1 | 5 days ago |
#1417 Upgrade Transformer Engine support from 1.12 to 2.18 Opened 5 days ago | aflah02 | open | feature request | 0 | 0 | 5 days ago |
#1405 Add Inter Document Attention Masking Opened 2 months ago | aflah02 | open | feature request | 0 | 0 | 7 days ago |
#1409 Add Extra Eval Iters similar to Extra Save Iters Opened 2 months ago | aflah02 | open | feature request | 2 | 0 | 20 days ago |
#1413 CI: run-tests prepares data before installing Python dependencies Opened 1 month ago | MGPOCKY | open | No labels | 0 | 0 | 1 month ago |
#606 CUDA Out of Memory for 20B Model on 2 A100 40GB GPUs Opened 4 years ago | seeEssex | open | No labels | 7 | 0 | 2 months ago |
#1394 Tracker: Improve HF Integration Opened 3 months ago | aflah02 | open | feature request | 0 | 0 | 2 months ago |
#1395 Add Training Speed Datapoints for new users Opened 3 months ago | aflah02 | open | feature request | 0 | 0 | 2 months ago |
#1400 Add round trip conversion based tests for import and export from HF Opened 2 months ago | aflah02 | open | feature request | 0 | 0 | 2 months ago |
#1393 Add Z-Loss Support for LM Pretraining Opened 3 months ago | aflah02 | open | feature request | 0 | 0 | 3 months ago |
#1391 anima — a substrate-native consciousness: capability gaps are architecture gaps (open repo, frozen verdicts) Opened 3 months ago | dancinlife | closed - completed | No labels | 1 | 0 | 3 months ago |
#1323 Error when converting sequential model to HF Opened 2 years ago | SilverSulfide | open | bug | 2 | 0 | 3 months ago |
#1380 pkg_resources deprecated Opened 5 months ago | PierreCarrier | open | bug | 2 | 0 | 4 months ago |
#1373 step1 models are always the same as step0 models Opened 7 months ago | StellaAthena | closed - completed | bug | 0 | 0 | 4 months ago |
#1381 A geometric perspective on knowledge: parallel mapping vs similarity Opened 4 months ago | weite76 | closed - not_planned | feature request | 0 | 0 | 4 months ago |
#1330 How are multiple datasets loaded? Opened 2 years ago | fxnie | closed - completed | question | 6 | 0 | 7 months ago |
#1004 [BUG] Inconsistent loss between `overlap_comm=true` and `overlap_comm=false` Opened 3 years ago | 0x6b64 | open | bug | 5 | 1 | 9 months ago |
#1335 DeeperSpeed-Pytorch Incompatibility Opened 2 years ago | tijmen | closed - completed | bug | 2 | 0 | 1 year ago |
#1364 MoE Top-K Training Opened 1 year ago | chuanyang-Zheng | open | No labels | 0 | 0 | 1 year ago |
#1363 Edge Case 17 Opened 1 year ago | Alexander-Duncan-Engineer | closed - completed | No labels | 0 | 0 | 1 year ago |
#1057 Support for Mosaic Models Opened 3 years ago | rajveer43 | open | feature request | 3 | 0 | 1 year ago |
#994 Convert HF Llama Checkpoints to Neox Checkpoints Opened 3 years ago | sxthunder | open | feature request | 2 | 3 | 1 year ago |
#1203 My servers used for multi-node training do not have ssh. How can I launch multi-node training using the torchrun command? Opened 2 years ago | ningding-o | open | feature request | 6 | 0 | 1 year ago |