Repository Issue Activity (beta)

evolvinglmms-lab/lmms-eval

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
24
New in 7 Days
4
Closed in 7 Days
4
Average Open Age
171 days
Stale 30+ Days
16
Stale 90+ Days
13
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-2000221
2026-09-1910251
2026-09-1801011
2026-09-1703034
2026-09-1620000
2026-09-15000025
2026-09-1400000
2026-09-1310000
2026-09-1200000
2026-09-1100000
2026-09-1000000
2026-09-0900000
2026-09-0801000
2026-09-0700000
This Week

Opened: 3

Closed: 4

Comments: 4

Events: 11

Top Labels
stale (181)
enhancement (39)
bug (28)
help wanted (12)
discussion (6)
question (2)
documentation (1)
duplicate (1)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#1541 VideoMME score is substantially higher than the reported LLaVA-OneVision-7B result with 32 frames

Opened 2 days ago
jiali456
open
bug
4013 hours ago

#1284 MMMU test dataset can be evaluated locally now

Opened 6 months ago
ywwynm
closed - completed
enhancement
203 days ago

#1505 mathvision: _fix_sqrt corrupts indexed roots (\sqrt[3]{8} → \sqrt{[}3]{8}), silently failing the latex2sympy equivalence path

Opened 16 days ago
feiiiiii5
open
No labels
203 days ago

#1519 PyAV packet fallback retains the entire decoded video before frame sampling

Opened 7 days ago
pentaoa
closed - completed
No labels
003 days ago

#1534 vllm_generate crashes on video tasks for zai-org/GLM-4.1V-9B-Thinking

Opened 4 days ago
jbyczkow
closed - completed
bug
103 days ago

#1530 vllm_generate ignores trust_remote_code, breaks all custom-code video models

Opened 4 days ago
jbyczkow
closed - completed
bug
103 days ago

#1509 mathvision: find_math_answer deletes the \text command name but keeps its braces — \boxed{\text{Sunday}} is scored wrong against gold Sunday

Opened 16 days ago
feiiiiii5
open
No labels
104 days ago

#1511 mathvision: find_math_answer keeps the FIRST \\boxed answer when a model retracts it and boxes a corrected one

Opened 16 days ago
feiiiiii5
open
No labels
104 days ago

#1507 mathvision: _fix_fracs is gated on "sqrt" — \frac12-style shorthand never normalizes, silently failing is_equal for sqrt-free answers

Opened 16 days ago
feiiiiii5
open
No labels
104 days ago

#1300 qwen3_5是文本图像多模态模型,但目前只能测评文本

Opened 5 months ago
dnut4398
open
bug
318 days ago

#1432 Decord causes segfaults on teardown

Opened 1 month ago
microslaw
closed - completed
bug
3012 days ago

#1215 [Installation] lmms-eval --tasks list returns empty list after pip install

Opened 7 months ago
YirongWho
closed - completed
bug
3020 days ago

#1240 [Bug] NCCL warning on error path: destroy_process_group() not called

Opened 7 months ago
akawincent
closed - completed
No labels
1021 days ago

#925 MVBench dataset average result

Opened 10 months ago
ylllllll
closed - completed
No labels
3121 days ago

#1481 OpenAI-compat: --model openai with base_url /v1; leave azure_openai off

Opened 24 days ago
cursor[bot]
closed - not_planned
No labels
1023 days ago

#1219 'Qwen3_VL' object has no attribute 'fps'

Opened 7 months ago
dfm021101
closed - completed
bug
2023 days ago

#1254 📋 Documentation Enhancement Suggestion

Opened 6 months ago
croviatrust
closed - not_planned
No labels
0023 days ago

#1483 Verify evals on Papers with Code

Opened 23 days ago
NielsRogge
open
No labels
0023 days ago

#1404 [Bug] chat `vllm` model drops video timestamps (to_openai_messages), degrading Qwen3-VL/Qwen3.5 on motion benchmarks

Opened 2 months ago
njb-nvidia
closed - completed
No labels
0024 days ago

#1476 transform data before feeding into the model

Opened 25 days ago
nedatghd
closed - completed
No labels
0025 days ago

#1393 `qwen2_5_vl` video metadata fix in #1269 does not affect `second_per_grid_ts`; effective workaround is to pass fps, but fps > 2 can break temporal position ids

Opened 2 months ago
jianglhui123
open
bug
1027 days ago

#1436 [Bug] MMMU open-ended scoring flattens candidate lists before eval_open

Opened 1 month ago
LioEinaudi
closed - completed
No labels
0027 days ago

#1319 OpenAI chat adapter drops text-task ctx, causing lm-eval-harness task mismatch

Opened 5 months ago
babyplutokurt
closed - completed
bug
0027 days ago

#1434 [Bug] BLINK answer extraction misparses verbose responses ("The correct answer is (B)" → "T"), penalizing small/verbose models

Opened 1 month ago
akawincent
closed - completed
No labels
1030 days ago

#1397 [Bug] SGLang batch reuses the first request's generation kwargs

Opened 2 months ago
CoderChen01
closed - completed
No labels
002 months ago

Rows per page:

1–25 of 497