Repository Issue Activity (beta)

axolotl-ai-cloud/axolotl

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
107
New in 7 Days
5
Closed in 7 Days
5
Average Open Age
416 days
Stale 30+ Days
100
Stale 90+ Days
86
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-0700000
2026-09-0600000
2026-09-0500000
2026-09-0413414107
2026-09-0320000
2026-09-0211000
2026-09-0101000
2026-08-3110000
2026-08-3000000
2026-08-2900000
2026-08-2804000
2026-08-2710000
2026-08-2630000
2026-08-2500000
This Week

Opened: 4

Closed: 5

Comments: 4

Events: 14

Top Labels
bug (362)
enhancement (183)
waiting for reporter (49)
good first issue (16)
wip (15)
help wanted (11)
waiting on upstream (10)
under review (9)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#3883 Opt out of upcoming SAC × saved_tensors_hooks change (pytorch/pytorch#190581)

Opened 1 month ago
ezyang
closed - completed
wip
under review
scheduled_release
313 days ago

#3973 Qwen3.5 monkeypatch `get_cu_seqlens` crashes under sample_packing: `.view(-1)` on a non-contiguous tensor

Opened 4 days ago
TomMoeras
closed - completed
No labels
203 days ago

#3967 Expose FP8 Scaling Recipe (Rowwise) Instead of Hardcoded Tensorwise

Opened 7 days ago
ved1beta
closed - completed
enhancement
003 days ago

#3981 Add support for new flash-attn (and kernel) interface in ring-flash-attention

Opened 3 days ago
NanoCode012
open
enhancement
003 days ago

#3972 Qwen3.5 monkeypatch reads `layer_type` but transformers 5.14.1 defines `block_type` (AttributeError in first forward pass)

Opened 4 days ago
TomMoeras
closed - completed
No labels
203 days ago

#3971 Async GRPO fails with `num_tiles` when data producer and prefetch are enabled

Opened 5 days ago
dudeperf3ct
open
bug
good first issue
113 days ago

#2650 Deprecated `flash_attn_fuse_qkv` toggle for the latest transformers (4.51.3 or lower)

Opened 1 year ago
DaizeDong
closed - completed
enhancement
106 days ago

#3955 [New model] Add Qwen4

Opened 12 days ago
NanoCode012
closed - completed
enhancement
wip
106 days ago

#3952 nemotron_h sample-packing monkeypatch fails on both transformers 5.14.1 and 5.15.0 (the versions in the recent official images)

Opened 12 days ago
john-broadway
closed - completed
No labels
4010 days ago

#3951 nemotron_h: 4-bit quantization packs `out_proj`; the mamba fused kernel reads `out_proj.weight` directly and crashes at step 0 (falcon_h1 already has this skip)

Opened 12 days ago
john-broadway
closed - completed
No labels
4010 days ago

#3959 `vllm-serve` ignores `revision_of_model` and loads the default Hugging Face main branch

Opened 11 days ago
dudeperf3ct
closed - completed
bug
0011 days ago

#3940 TokensPerSecondCallback captures resume_from_checkpoint before auto-resume detection resolves it, silently zeroing token metrics

Opened 17 days ago
AmirF194
open
No labels
2013 days ago

#3203 OOM for causal lm evaluation and missing logging

Opened 11 months ago
maximrepidgey
open
bug
waiting for reporter
3018 days ago

#3836 check_batch_size_fields rejects configs that rely on the documented micro_batch_size/gradient_accumulation_steps defaults

Opened 2 months ago
ErenAta16
closed - completed
No labels
1021 days ago

#3754 `chat_template` Strategy Never Trains the Gemma Turn Terminator

Opened 2 months ago
thad0ctor
closed - completed
bug
4029 days ago

#3890 model support system follow-ups

Opened 1 month ago
thad0ctor
open
enhancement
1029 days ago

#3626 [Feature] Automatic LoRA rank recommendation based on dataset size

Opened 5 months ago
rehan243
open
good first issue
401 month ago

#3848 MultipackBatchSampler.generate_batches() stale-reference bug corrupts total_token_slots and can crash with IndexError

Opened 2 months ago
ErenAta16
closed - completed
No labels
001 month ago

#3801 PEFT bump pending

Opened 2 months ago
thad0ctor
closed - completed
enhancement
901 month ago

#3908 MLflow/WandB/Comet/Trackio env vars not set on Ray Train worker (use_ray: true)

Opened 1 month ago
fialhocoelho
open
No labels
001 month ago

#3608 Ring Attention w/ document packing produces different results

Opened 5 months ago
Ueaj-Kerman
open
bug
902 months ago

#3856 kks

Opened 2 months ago
Pythonaja
closed - not_planned
No labels
002 months ago

#3857 asj

Opened 2 months ago
Pythonaja
closed - not_planned
enhancement
002 months ago

#3787 Expert-granularity CPU offload for quantized MoE: 30B-A3B QLoRA in ~7 GB (single resident expert layer, extends the layer_offloading idea)

Opened 2 months ago
pjordanandrsn
open
waiting for reporter
1102 months ago

#2396 EXTREMELY SLOW (unusable) towards end of tokenization of dataset with long multi turn conversations

Opened 2 years ago
Nero10578
open
bug
1522 months ago

Rows per page:

1–25 of 600