Repository Issue Activity (beta)

vllm-project/llm-compressor

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
52
New in 7 Days
11
Closed in 7 Days
8
Average Open Age
46 days
Stale 30+ Days
25
Stale 90+ Days
5
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-08-2410383
2026-08-2301113
2026-08-2211121
2026-08-2122265
2026-08-20205178
2026-08-1930268
2026-08-1804393
2026-08-1720001
2026-08-1600000
2026-08-15102103
2026-08-1400263
2026-08-1300000
2026-08-1210253
2026-08-1101001
This Week

Opened: 9

Closed: 8

Comments: 17

Events: 49

Top Labels
bug (338)
enhancement (171)
good first issue (93)
stale (59)
question (46)
vllm (31)
awq (28)
keep-open (23)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#3065 [Bug]: autowrap_forward fails with NameError for models defined in __main__ under cProfile/runpy

Opened 5 days ago
Isitthakkar11
open
bug
201 day ago

#3083 [Bug]: requires_compute_capability crashes at collection time on Apple Silicon

Opened 2 days ago
Isitthakkar11
open
bug
002 days ago

#3057 [Bug] sequential pipeline cannot trace models whose forward is decorated without functools.wraps (Qwen3.8 / Qwen3_5GatedDeltaNet)

Opened 5 days ago
malaiwah
closed - completed
bug
tracing
222 days ago

#3037 Qwen3.5-MoE misses ARCH_TO_2D_MAPPINGS, forcing a 2D→3D→2D expert round-trip that doubles expert memory at load

Opened 7 days ago
wenis
open
No labels
602 days ago

#2296 [AWQ] add option to take smooth layer quantization into accout

Opened 7 months ago
HDCharles
closed - not_planned
enhancement
good first issue
awq
1003 days ago

#3054 [Bug] AutoRoundModifier: calibration inputs retained on GPU exhaust VRAM before optimization

Opened 5 days ago
xesdiny
open
autoround
113 days ago

#2939 GPTQ silently corrupts fused-expert MoE (Qwen3.6-A3B) on 0.12.0 — fixed on main by #2847/#2848, please cut a release

Opened 1 month ago
narfian
closed - completed
No labels
513 days ago

#3069 [Bug]: Cannot dispatch MOE correctly if using naming `w1,w3` or `block_sparse_moe`

Opened 4 days ago
xin3he
open
bug
103 days ago

#3071 [Bug]: Qwen3-VL-2B (dense) AWQ/GPTQ calibration fails, vision tower not FX-traceable (rot_pos_emb CUDA driver error), even with visual.* in ignore

Opened 3 days ago
gxaurab
open
bug
003 days ago

#2984 [Bug]: Loading models with DDP leads to race condition on HF `cached_files`

Opened 25 days ago
kylesayrs
open
bug
good first issue
604 days ago

#2936 How to achieve high accuracy on NVFP4?

Opened 1 month ago
IEI-mjx
open
enhancement
1404 days ago

#3064 [Bug]: tools/collect_env.py crashes on Apple Silicon (torch.mps has no get_device_name)

Opened 5 days ago
Isitthakkar11
open
bug
104 days ago

#2949 [Feature] Zero-copy DDP weight sharing for memory-constrained large MoE quantization

Opened 1 month ago
xesdiny
open
No labels
1505 days ago

#3059 [Bug]: Noisy gpt_oss test

Opened 5 days ago
kylesayrs
open
bug
good first issue
105 days ago

#2741 [Feature] Native intra-calibration resume in `SequentialPipeline`

Opened 3 months ago
pasta-paul
open
stale
105 days ago

#3040 [Bug] Fix ignore list for linear_attn with Qwen3.8

Opened 7 days ago
dsikka
open
bug
good first issue
good follow-up issue
415 days ago

#2919 Q3 Roadmap

Opened 1 month ago
dsikka
open
RFC
keep-open
ROADMAP
015 days ago

#2522 [Bug]: Evaluate AWQ and GPTQ for Gemma

Opened 5 months ago
Jeevi10
closed - completed
bug
stale
3106 days ago

#2690 [RFC] HIGGS Integration into llm-compressor

Opened 4 months ago
krishnateja95
open
enhancement
RFC
stale
506 days ago

#3032 [Feature] Multimodal calibration dataset support in oneshot

Opened 9 days ago
sudo-0x2a
open
enhancement
226 days ago

#3023 [Bug] AutoRoundModifier: input_capture_hook accumulates GPU memory during optimization when gradient_accumulate_steps > 1

Opened 12 days ago
xesdiny
closed - completed
No labels
116 days ago

#2693 [Bug]: `compressed-tensors` W4A8 INT (int4 weights + int8 activations) fails to find any compatible kernel on H100

Opened 4 months ago
RedHeartSecretMan
closed - completed
bug
506 days ago

#2975 [ModelFreePTQ] Memory issues with multi-gpu

Opened 28 days ago
kylesayrs
closed - completed
enhancement
good first issue
model_free_ptq
106 days ago

#1939 HOW To: Quantization: Qwen/Qwen3-VL-30B-A3B-Instruct AWQ

Opened 10 months ago
JartX
closed - completed
qwen
awq
6807 days ago

#2979 [Kimi-K3] Integrate KimiSparseMoeBlock with LinearExperts2D for REAP Support

Opened 27 days ago
kylesayrs
open
enhancement
good first issue
moe
2110 days ago

Rows per page:

1–25 of 735