Repository Issue Activity (beta)

bigcode-project/bigcode-evaluation-harness

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
61
New in 7 Days
1
Closed in 7 Days
0
Average Open Age
639 days
Stale 30+ Days
60
Stale 90+ Days
60
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-08-2400000
2026-08-2300001
2026-08-2210000
2026-08-2100000
2026-08-2000000
2026-08-1900000
2026-08-1800000
2026-08-1700000
2026-08-1600000
2026-08-1500000
2026-08-1400000
2026-08-1300000
2026-08-1200000
2026-08-1100000
This Week

Opened: 1

Closed: 0

Comments: 0

Events: 0

Top Labels
good first issue (10)
enhancement (7)
help wanted (4)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#325 Optional EvalPort export for task datasets and pass@k results

Opened 2 days ago
adhabnr-ux
open
No labels
002 days ago

#322 Community task: Helium open benchmarks

Opened 2 months ago
connerlambden
closed - completed
No labels
102 months ago

#320 📝 Integration Proposal: CAJAL — Scientific Paper Agent for BigCode

Opened 4 months ago
Agnuxo1
open
No labels
004 months ago

#318 Question: is there room for long horizon "spec tension" tests around complex code tasks?

Opened 7 months ago
onestardao
closed - completed
No labels
105 months ago

#304 Unable to execute the MultiPL-E task for C++

Opened 1 year ago
JohnneyQin
open
No labels
109 months ago

#315 Wrong Hugging Face link for Spider dataset in bigcode-evaluation-harness/docs/README.md line 434

Opened 1 year ago
354246695
open
No labels
101 year ago

#311 Improve pass@1 Score on Humaneval

Opened 1 year ago
showlibia
closed - completed
No labels
011 year ago

#313 Support configurability of FIM tokens on SantaCoder

Opened 1 year ago
Jay-Roberts
open
No labels
001 year ago

#224 Multiple-E Go test file name suffix does not contain _test.go

Opened 2 years ago
sagtanih
closed - completed
No labels
011 year ago

#240 Some questions about APPS

Opened 2 years ago
virt9
closed - completed
No labels
201 year ago

#308 testing Humaneval of qwen-2.5-7B-coder-instruct

Opened 1 year ago
zxiangx
open
No labels
021 year ago

#307 how to add new model?

Opened 1 year ago
pengzhangzhi
open
No labels
001 year ago

#306 When executing languages such as JS and GO in Multiple-E, the generated results suddenly end

Opened 1 year ago
fxnie
open
No labels
001 year ago

#266 What is `fine-tuning` in task submission?

Opened 2 years ago
zhimin-z
open
No labels
011 year ago

#303 Code-Llama-7B-Python 4 Bit Error on HumanEval

Opened 1 year ago
wilyub
open
No labels
001 year ago

#300 is there a benchmark page on the benchmark results evaluated using bigcode-evaluation-harness

Opened 2 years ago
yxchng
open
No labels
102 years ago

#131 'HumanEval' object has no attribute 'dataset'

Opened 3 years ago
dongguanting
closed - completed
No labels
742 years ago

#192 Potentially extra slow inference when using LoRA adapter

Opened 3 years ago
sadaisystems
open
No labels
202 years ago

#271 Evaluating a Model with a Local Dataset in an Offline Environment

Opened 2 years ago
ankush13r
open
No labels
902 years ago

#289 HumanEval-X Go evaluation raises error

Opened 2 years ago
nielstron
closed - completed
No labels
202 years ago

#297 Select device: GPU for model eval

Opened 2 years ago
kn0wn-cyber
open
No labels
002 years ago

#290 HumanEval-X generation appears to not time out called subproceses

Opened 2 years ago
nielstron
open
No labels
002 years ago

#288 Adding an assistant_model Argument for Speculative Decoding

Opened 2 years ago
ilyasoulk
open
No labels
002 years ago

#283 Docker image for multiple evalulation broken

Opened 2 years ago
Extirpater
open
No labels
102 years ago

#287 Could you share a completed file of generations_mbppplus.json

Opened 2 years ago
FearandDreams1123
open
No labels
002 years ago

Rows per page:

1–25 of 164