Repository Issue Activity (beta)

huggingface/open-r1

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
290
New in 7 Days
0
Closed in 7 Days
0
Average Open Age
380 days
Stale 30+ Days
290
Stale 90+ Days
290
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-08-2400000
2026-08-2300000
2026-08-2200000
2026-08-2100000
2026-08-2000000
2026-08-1900000
2026-08-1800000
2026-08-1700000
2026-08-1600000
2026-08-1500000
2026-08-1400000
2026-08-1300000
2026-08-1200000
2026-08-1100000
This Week

Opened: 0

Closed: 0

Comments: 0

Events: 0

Top Labels

No label distribution is available yet.

Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#727 Eval datasets for reasoning

Opened 2 months ago
connerlambden
closed - completed
No labels
102 months ago

#685 Interest in dataset based on ACM SGU?

Opened 1 year ago
radoslav11
closed - completed
No labels
302 months ago

#723 New aproach to developing AGI

Opened 3 months ago
Kwoargus
open
No labels
003 months ago

#343 in SFT script, distributed training got stuck if set `packing=false`

Opened 2 years ago
ChenDRAG
open
No labels
915 months ago

#692 RuntimeError: The size of tensor a (0) must match the size of tensor b (5120) at non-singleton dimension 1

Opened 1 year ago
george1459
open
No labels
105 months ago

#719 Dependency conflicts

Opened 6 months ago
DisGuiser9
open
No labels
006 months ago

#341 Bug in SFT script

Opened 2 years ago
ChenDRAG
open
No labels
1507 months ago

#239 Why does the loss start at 0 when I train GRPO, and then possibly increase?

Opened 2 years ago
hellen9527
closed - completed
No labels
4107 months ago

#715 Is there a Chinese version of the dataset in this project?

Opened 8 months ago
fxb392
open
No labels
008 months ago

#714 grad_norm nan

Opened 9 months ago
weiyezhimeng
closed - completed
No labels
209 months ago

#713 grad_norm nan

Opened 9 months ago
weiyezhimeng
closed - completed
No labels
009 months ago

#712 grad_norm nan

Opened 9 months ago
weiyezhimeng
closed - completed
No labels
009 months ago

#118 Facing OOM during DRPO stage

Opened 2 years ago
Some-random
closed - completed
No labels
209 months ago

#710 SFT Trains the Model on the Entire Sequence

Opened 10 months ago
RohollahHS
open
No labels
0110 months ago

#707 Forward reward always 0

Opened 10 months ago
ZhaoyangLi-1
open
No labels
0010 months ago

#705 Checkpoint selection when using cot sft

Opened 11 months ago
lzhptr
open
No labels
0011 months ago

#704 why kl = nan when grpo train?

Opened 1 year ago
uilstong
open
No labels
001 year ago

#702 Request for License Information for Mixture-of-Thoughts Dataset

Opened 1 year ago
mmyymmyy321
open
No labels
001 year ago

#694 GSPO update request

Opened 1 year ago
LuozhuZhang
closed - completed
No labels
021 year ago

#403 Instead of rising steadily, the reward fluctuates wildly

Opened 1 year ago
Hasuer
open
No labels
1101 year ago

#699 Does the SFT Framework Support Fine-Tuning for DeepSeek-R1-Distill-Qwen-7B

Opened 1 year ago
ZJUCQR
open
No labels
001 year ago

#698 Question about evaluating AIME24 Accuracy

Opened 1 year ago
ZJUCQR
open
No labels
001 year ago

#695 vllm prepends two BOS for LLama

Opened 1 year ago
wenquanlu
open
No labels
001 year ago

#684 Latest version of flash_attn no longer works with torch==2.6.0

Opened 1 year ago
george1459
open
No labels
2151 year ago

#444 How to increase the context window from 4k to 32k on qwen models ?

Opened 1 year ago
Jeremmmyyyyy
closed - completed
No labels
601 year ago

Rows per page:

1–25 of 397