Repository Issue Activity (beta)

internlm/lmdeploy

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
346
New in 7 Days
3
Closed in 7 Days
0
Average Open Age
485 days
Stale 30+ Days
334
Stale 90+ Days
325
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-0900001
2026-09-081011346
2026-09-0700000
2026-09-0600000
2026-09-0500000
2026-09-0400000
2026-09-0320000
2026-09-0200000
2026-09-0100000
2026-08-3101000
2026-08-3000000
2026-08-2910000
2026-08-2800000
2026-08-2700000
This Week

Opened: 3

Closed: 0

Comments: 1

Events: 1

Top Labels
awaiting response (226)
Stale (174)
backlog (19)
planned feature (4)
mllm (3)
wontfix (2)
documentation (1)
duplicate (1)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#4944 [Bug] TurboMind SIGSEGV crash-loop triggered by session_len truncation path (deterministic libc offset across 3 crashes)

Opened 3 days ago
zambalee
open
No labels
103 days ago

#4933 [Bug] input_embeddings are omitted from prefix-cache identity, causing wrong KV reuse

Opened 7 days ago
wildoranges
open
No labels
007 days ago

#4930 [Bug] Mooncake external KV keys omit KV format and weights lineage, with no default tenant isolation

Opened 7 days ago
wildoranges
open
No labels
207 days ago

#4824 [Feature] DeepSeek-V4 MoE is Blackwell-only via DeepGEMM — is a non-Blackwell path wanted?

Opened 1 month ago
Hakureirm
open
No labels
9010 days ago

#4530 [Feature] support DFlash: Block Diffusion for Flash Speculative Decoding

Opened 5 months ago
hicofeng
open
planned feature
3010 days ago

#4905 [Feature] 咱们最新版本的lmdeploy可以加载这个 Qwen3.8-Flash吗?

Opened 15 days ago
simonjhy
open
No labels
1010 days ago

#4889 [Feature] 什么时候turbomind能够开始支持MTP或者DFlash之类的加速功能?

Opened 21 days ago
simonjhy
open
No labels
1010 days ago

#4917 [Bug] Prefix-cache KV is not invalidated by /update_weights: hot weight update silently reuses stale KV (wrong outputs, cached_tokens still reported)

Opened 12 days ago
wildoranges
closed - completed
No labels
1011 days ago

#4899 [Bug] PyTorch engine: Qwen3.5 hybrid-GDN + AWQ checkpoint crashes on TP>1 - 'AwqLinear' object has no attribute 'block_size'

Opened 17 days ago
u5084757454-collab
open
No labels
0417 days ago

#4781 [Bug] legacy /generate's non-streaming disconnect branch discards its 400 response, returns 200 null instead

Opened 2 months ago
AmirF194
closed - completed
No labels
1018 days ago

#4883 [Bug] pytorch backend silently hangs under concurrent decoding with qwen3_5_mtp speculative decoding

Opened 22 days ago
matrix72c
open
No labels
1021 days ago

#4879 [Feature] 当前qwen-next等SSM类模型prefix cache缓存命中率低

Opened 23 days ago
Tsundoku958
open
No labels
0022 days ago

#4863 [Bug] turbomind can not support fp8 in SM75

Opened 27 days ago
bltcn
closed - completed
No labels
2022 days ago

#4870 [Feature] lmdeploy 框架的 API 层:默认不支持思维深度设置为 xhigh 吗?

Opened 25 days ago
simonjhy
closed - completed
No labels
0025 days ago

#4864 [Feature] 什么时候能够开始支持 https://huggingface.co/Qwen/Qwen3.8-27B?

Opened 27 days ago
simonjhy
closed - completed
No labels
1027 days ago

#4865 [Feature] sm75 can support awq/gptq/fp8/w8a8 in turbomind

Opened 27 days ago
bltcn
open
No labels
0027 days ago

#4849 [Bug] int4 KV cache: the quantization range is stretched to include 0 when the packed head width is not a power of two

Opened 29 days ago
truong-v
closed - completed
No labels
0028 days ago

#4804 [Bug] Unauthenticated pickle deserialization via ZMQ in disaggregated serving → RCE

Opened 1 month ago
AAtomical
closed - completed
No labels
1029 days ago

#4832 [Feature] 现在最新的0.15.0可以支持最新的deepseek-v4-flash-0731吗?

Opened 1 month ago
simonjhy
open
No labels
2030 days ago

#4329 [Bug] 基于https://github.com/InternLM/lmdeploy/pull/4320 构建Docker Image,启动GLM-4.7-Flash依然报错

Opened 7 months ago
simonjhy
closed - completed
No labels
901 month ago

#4400 [Bug] 在lmdeploy中加载glm-4.7-flash参数如何配置?

Opened 6 months ago
simonjhy
closed - completed
No labels
601 month ago

#3781 [Bug] fill_kv_cache 实现有bug

Opened 1 year ago
zxy1123
closed - completed
No labels
101 month ago

#4715 [Bug] 在使用qwen3.5-122B的时候,模型是多模态模型,但是在使用lmdeploy收到请求的时候,报错了

Opened 2 months ago
simonjhy
closed - not_planned
awaiting response
Stale
301 month ago

#4743 [Bug] INT4 KV cache (--quant-policy 4) breaks qwen3coder tool call parser for Qwen3-Coder-30B

Opened 2 months ago
zambalee
closed - not_planned
awaiting response
Stale
301 month ago

#4776 [Bug] legacy /v1/completions disconnect-abort calls nonexistent AsyncEngine.stop_session

Opened 2 months ago
AmirF194
closed - completed
No labels
101 month ago

Rows per page:

1–25 of 1,039