Repository Issue Activity (beta)

ml-gsai/llada

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
88
New in 7 Days
0
Closed in 7 Days
0
Average Open Age
449 days
Stale 30+ Days
88
Stale 90+ Days
83
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-2000000
2026-09-1900000
2026-09-1800000
2026-09-17000088
2026-09-1600000
2026-09-1500000
2026-09-1400000
2026-09-1300000
2026-09-1200000
2026-09-1100000
2026-09-1000000
2026-09-0900000
2026-09-0800000
2026-09-0700000
This Week

Opened: 0

Closed: 0

Comments: 0

Events: 0

Top Labels

No label distribution is available yet.

Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#137 `confidence_eos_eot_inf` does not correctly suppress EOS confidence

Opened 1 month ago
johnrunjiali-svg
open
No labels
001 month ago

#136 [Bug] A single padding token can trigger Math SDPA fallback and a quadratic GPU-memory spike [`flash_attn_varlen_func` not supported]

Opened 1 month ago
zpointS
open
No labels
001 month ago

#135 Verify evals on Papers with Code

Opened 2 months ago
NielsRogge
open
No labels
002 months ago

#133 Segfault loading LLaDA-8B-Instruct under transformers >= 5 (new core_model_loading path) — works with <5 pin

Opened 2 months ago
ghost
open
No labels
002 months ago

#132 iLLaDA-8B-Instruct: LLaDA's generate.py with mask_id=5 alone produces degenerate token-repetition output by iLLaDA

Opened 3 months ago
lanstonchu
closed - completed
No labels
403 months ago

#122 How to evaluate the peft model

Opened 9 months ago
zaiquanyang
open
No labels
103 months ago

#121 Unconditional Generation

Opened 9 months ago
Yuxuan314
open
No labels
204 months ago

#131 LLaDA inference logits differ across NVIDIA GPUs (Blackwell vs 4090) with same weights/code

Opened 5 months ago
Kevin-lkw
open
No labels
005 months ago

#114 Feature Request: Enable output_attentions=True for LLaDA Models

Opened 11 months ago
mli746
open
No labels
406 months ago

#130 AttributeError: 'LLaDAModelLM' object has no attribute 'all_tied_weights_keys'

Opened 6 months ago
josesho
closed - completed
No labels
106 months ago

#129 Typo in Algorithm 5: Masking ratio inverted?

Opened 6 months ago
lazitech
open
No labels
006 months ago

#128 📋 Documentation Enhancement Suggestion

Opened 6 months ago
croviatrust
open
No labels
006 months ago

#125 Usage of CFG for Conditional Generation Benchmarks

Opened 7 months ago
JeiminJeon
open
No labels
007 months ago

#123 [Docs] Add PRISM test-time scaling recipe / reference for LLaDA-8B-Instruct in README

Opened 8 months ago
viiika
open
No labels
008 months ago

#119 cpu memory OOM when using lm-eval to evaluate LLADA

Opened 10 months ago
duterscmy
open
No labels
508 months ago

#120 RoPE and attention mask mismatch

Opened 9 months ago
probablyabot
open
No labels
409 months ago

#103 Problems when using lm_eval and calculating ppl

Opened 1 year ago
ZijianYY
open
No labels
239 months ago

#36 Performance Optimization: Significantly Speed Up LLaDA Generation in generate.py

Opened 2 years ago
mochibuilds
closed - completed
No labels
399 months ago

#62 Is float64 necessary?

Opened 1 year ago
maomaocun
open
No labels
309 months ago

#44 Inference Speed

Opened 2 years ago
jacklishufan
open
No labels
31010 months ago

#115 [Help Needed] mmengine.Config.fromfile() resolves wrong path for local configs

Opened 11 months ago
SIKAI-C
open
No labels
4011 months ago

#112 Evaluation is too slow

Opened 1 year ago
Algolzw
closed - completed
No labels
2011 months ago

#106 Unexpectedly high CPU memory usage when running GSM8K with LLaDA

Opened 1 year ago
XiangZhang-zx
open
No labels
2011 months ago

#107 Update the WeChat QR code

Opened 1 year ago
jiangzizi
open
No labels
3011 months ago

#113 Failed to finetune llada moe with math competition dataset

Opened 1 year ago
MengAiDev
closed - completed
No labels
001 year ago

Rows per page:

1–25 of 115