Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-08-24 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-23 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-22 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-21 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-20 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-19 | 1 | 0 | 1 | 1 | 1 |
| 2026-08-18 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-17 | 0 | 1 | 2 | 5 | 2 |
| 2026-08-16 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-15 | 0 | 0 | 0 | 0 | 1 |
| 2026-08-14 | 1 | 1 | 0 | 0 | 0 |
| 2026-08-13 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-12 | 1 | 0 | 1 | 2 | 3 |
| 2026-08-11 | 0 | 0 | 0 | 0 | 0 |
Opened: 1
Closed: 0
Comments: 1
Events: 1
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#2378 Qwen3.6-35B-A3B (Qwen3_5MoeForConditionalGeneration): every completion request fails with 'moe experts forward / PendingIsqLayer is in an invalid transitional state' Opened 5 days ago | rainfall-rajesh | open | No labels | 1 | 0 | 5 days ago |
#1714 Qwen 3 VL support for GGUF models? Opened 9 months ago | thevishnupradeep | closed - completed | new feature | 4 | 2 | 7 days ago |
#2262 Cargo Crate out of date Opened 2 months ago | dhleong | open | bug | 2 | 4 | 7 days ago |
#2125 GGUF loading fails for Qwen3.5 models with `Unknown GGUF architecture qwen35` Opened 4 months ago | rlaskaehd | open | new feature | 6 | 3 | 7 days ago |
#2367 Don't force images and other media to come first in context for VLMs Opened 9 days ago | copumpkin | open | new feature | 0 | 0 | 9 days ago |
#2361 Support qwen35 architecture Opened 19 days ago | dreyn74 | closed - completed | new feature | 2 | 0 | 9 days ago |
#2363 Chat template output differs from transformers on tool-calling prompts Opened 12 days ago | GregoryBolshakov | open | No labels | 0 | 0 | 12 days ago |
#1951 Support realtime API for audio streaming Opened 6 months ago | npuichigo | open | new feature | 1 | 0 | 12 days ago |
#156 Model Wishlist Opened 2 years ago | EricLBuehler | open | models | 187 | 3 | 12 days ago |
#1114 cuda+cpu mode panic at unwrap Opened 2 years ago | grinapo | closed - completed | bug | 8 | 0 | 18 days ago |
#1948 Docker: Missing mistralrs binary and libc dependencies in cpu-0.7 image Opened 6 months ago | hidehiroanto | open | bug | 1 | 0 | 21 days ago |
#2340 Unsupported group_size: 6 on CUDA Opened 1 month ago | svenstaro | open | bug | 0 | 0 | 24 days ago |
#2360 Unknown GGUF architecture `gemma4` Opened 24 days ago | ralphIsidore | closed - completed | No labels | 1 | 0 | 24 days ago |
#2024 paged_attention: O(N²) memory allocator thrashing and continuous batching failure Opened 5 months ago | glaziermag | closed - completed | No labels | 1 | 1 | 24 days ago |
#2343 Concurrent text-serving throughput does not scale above serial despite PagedAttention engaged (v0.9.0, H100, Llama-3.2-1B bf16) Opened 1 month ago | dharmater | open | No labels | 0 | 0 | 1 month ago |
#2323 Batched multimodal prefill crashes with rank mismatch; mixed-length image batches silently corrupt output (Gemma 4) Opened 2 months ago | achandi | open | No labels | 1 | 0 | 1 month ago |
#2336 Web UI: sending an image returns HTTP 500 `invalid image source` (v1 API works) Opened 1 month ago | subin9 | open | No labels | 0 | 0 | 1 month ago |
#2335 Llama-3.2-11B-Vision (mllama): errors with PagedAttention by default, and image inference crashes on CUDA with `--paged-attn off` Opened 1 month ago | subin9 | open | No labels | 0 | 0 | 1 month ago |
#2329 sampler: greedy paths report log of raw logit instead of log of probability Opened 2 months ago | dx2102 | open | No labels | 0 | 0 | 1 month ago |
#2328 Re-export pinned candle versions (or at least candle-core::Device) Opened 2 months ago | thiezn | open | new feature | 0 | 0 | 2 months ago |
#2322 MemoryUsage::query runs sysinfo System::new_all() in the attention hot path (CPU + integrated CUDA) Opened 2 months ago | achandi | open | No labels | 0 | 0 | 2 months ago |
#2321 PagedAttentionScheduler::schedule livelocks with concurrent multimodal requests Opened 2 months ago | achandi | open | No labels | 0 | 0 | 2 months ago |
#2128 Add architecture support for `paddleocr_vl` (PaddleOCR-VL-1.5) Opened 4 months ago | nicolas-geysse | open | No labels | 2 | 0 | 2 months ago |
#2160 GGUF Qwen3 MoE on Metal loads successfully but fails on first inference via CUDA-only indexed_moe_forward path Opened 3 months ago | j-zuilkowski | closed - completed | No labels | 1 | 0 | 2 months ago |
#2272 All looks good on GB10 with 0.8.21 but doesn't work due to CUDA mismatch (?) Opened 2 months ago | misureaudio | closed - completed | bug | 4 | 0 | 2 months ago |