Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-07 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-06 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-05 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-04 | 1 | 3 | 4 | 14 | 107 |
| 2026-09-03 | 2 | 0 | 0 | 0 | 0 |
| 2026-09-02 | 1 | 1 | 0 | 0 | 0 |
| 2026-09-01 | 0 | 1 | 0 | 0 | 0 |
| 2026-08-31 | 1 | 0 | 0 | 0 | 0 |
| 2026-08-30 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-29 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-28 | 0 | 4 | 0 | 0 | 0 |
| 2026-08-27 | 1 | 0 | 0 | 0 | 0 |
| 2026-08-26 | 3 | 0 | 0 | 0 | 0 |
| 2026-08-25 | 0 | 0 | 0 | 0 | 0 |
Opened: 4
Closed: 5
Comments: 4
Events: 14
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#3883 Opt out of upcoming SAC × saved_tensors_hooks change (pytorch/pytorch#190581) Opened 1 month ago | ezyang | closed - completed | wip under review scheduled_release | 3 | 1 | 3 days ago |
#3973 Qwen3.5 monkeypatch `get_cu_seqlens` crashes under sample_packing: `.view(-1)` on a non-contiguous tensor Opened 4 days ago | TomMoeras | closed - completed | No labels | 2 | 0 | 3 days ago |
#3967 Expose FP8 Scaling Recipe (Rowwise) Instead of Hardcoded Tensorwise Opened 7 days ago | ved1beta | closed - completed | enhancement | 0 | 0 | 3 days ago |
#3981 Add support for new flash-attn (and kernel) interface in ring-flash-attention Opened 3 days ago | NanoCode012 | open | enhancement | 0 | 0 | 3 days ago |
#3972 Qwen3.5 monkeypatch reads `layer_type` but transformers 5.14.1 defines `block_type` (AttributeError in first forward pass) Opened 4 days ago | TomMoeras | closed - completed | No labels | 2 | 0 | 3 days ago |
#3971 Async GRPO fails with `num_tiles` when data producer and prefetch are enabled Opened 5 days ago | dudeperf3ct | open | bug good first issue | 1 | 1 | 3 days ago |
#2650 Deprecated `flash_attn_fuse_qkv` toggle for the latest transformers (4.51.3 or lower) Opened 1 year ago | DaizeDong | closed - completed | enhancement | 1 | 0 | 6 days ago |
#3955 [New model] Add Qwen4 Opened 12 days ago | NanoCode012 | closed - completed | enhancement wip | 1 | 0 | 6 days ago |
#3952 nemotron_h sample-packing monkeypatch fails on both transformers 5.14.1 and 5.15.0 (the versions in the recent official images) Opened 12 days ago | john-broadway | closed - completed | No labels | 4 | 0 | 10 days ago |
#3951 nemotron_h: 4-bit quantization packs `out_proj`; the mamba fused kernel reads `out_proj.weight` directly and crashes at step 0 (falcon_h1 already has this skip) Opened 12 days ago | john-broadway | closed - completed | No labels | 4 | 0 | 10 days ago |
#3959 `vllm-serve` ignores `revision_of_model` and loads the default Hugging Face main branch Opened 11 days ago | dudeperf3ct | closed - completed | bug | 0 | 0 | 11 days ago |
#3940 TokensPerSecondCallback captures resume_from_checkpoint before auto-resume detection resolves it, silently zeroing token metrics Opened 17 days ago | AmirF194 | open | No labels | 2 | 0 | 13 days ago |
#3203 OOM for causal lm evaluation and missing logging Opened 11 months ago | maximrepidgey | open | bug waiting for reporter | 3 | 0 | 18 days ago |
#3836 check_batch_size_fields rejects configs that rely on the documented micro_batch_size/gradient_accumulation_steps defaults Opened 2 months ago | ErenAta16 | closed - completed | No labels | 1 | 0 | 21 days ago |
#3754 `chat_template` Strategy Never Trains the Gemma Turn Terminator Opened 2 months ago | thad0ctor | closed - completed | bug | 4 | 0 | 29 days ago |
#3890 model support system follow-ups Opened 1 month ago | thad0ctor | open | enhancement | 1 | 0 | 29 days ago |
#3626 [Feature] Automatic LoRA rank recommendation based on dataset size Opened 5 months ago | rehan243 | open | good first issue | 4 | 0 | 1 month ago |
#3848 MultipackBatchSampler.generate_batches() stale-reference bug corrupts total_token_slots and can crash with IndexError Opened 2 months ago | ErenAta16 | closed - completed | No labels | 0 | 0 | 1 month ago |
#3801 PEFT bump pending Opened 2 months ago | thad0ctor | closed - completed | enhancement | 9 | 0 | 1 month ago |
#3908 MLflow/WandB/Comet/Trackio env vars not set on Ray Train worker (use_ray: true) Opened 1 month ago | fialhocoelho | open | No labels | 0 | 0 | 1 month ago |
#3608 Ring Attention w/ document packing produces different results Opened 5 months ago | Ueaj-Kerman | open | bug | 9 | 0 | 2 months ago |
#3856 kks Opened 2 months ago | Pythonaja | closed - not_planned | No labels | 0 | 0 | 2 months ago |
#3857 asj Opened 2 months ago | Pythonaja | closed - not_planned | enhancement | 0 | 0 | 2 months ago |
#3787 Expert-granularity CPU offload for quantized MoE: 30B-A3B QLoRA in ~7 GB (single resident expert layer, extends the layer_offloading idea) Opened 2 months ago | pjordanandrsn | open | waiting for reporter | 11 | 0 | 2 months ago |
#2396 EXTREMELY SLOW (unusable) towards end of tokenization of dataset with long multi turn conversations Opened 2 years ago | Nero10578 | open | bug | 15 | 2 | 2 months ago |