-
-
Notifications
You must be signed in to change notification settings - Fork 20.3k
All issues
Issue creation is restricted in this repository
- #42770 · WoosukKwon opened
on May 15, 2026 21 - #44280 · BugenZhao opened
on Jun 2, 2026 21 - #50001 · ywang96 opened
on Jul 27, 2026 9
Issues
is:issue state:open
is:issue state:open
Search results
[Bug]: Prefix caching is ineffective on Mamba-2/GDN hybrid (Qwen3_5MoeForConditionalGeneration)
bugSomething isn't workingSomething isn't workingStatus: Open.#51250 In vllm-project/vllm;- Status: Open.#51240 In vllm-project/vllm;
[Feature]: per request kv cache write logic
feature requestNew feature or requestNew feature or requestStatus: Open.#51234 In vllm-project/vllm;[Rocm] Fix AITER_MLA issues for the kimik3-dspark model with Agentic workload
rocmRelated to AMD ROCmRelated to AMD ROCmStatus: Open.#51232 In vllm-project/vllm;[Bug]: seed is ignored on CPU when another request in the batch has no seed
bugSomething isn't workingSomething isn't workingStatus: Open.#51226 In vllm-project/vllm;[CI/Build]: source_file_dependencies still point at pre-libtorch_stable csrc paths, so kernel jobs no longer trigger
rocmRelated to AMD ROCmRelated to AMD ROCmStatus: Open.#51225 In vllm-project/vllm;[Feature]: DeepSeek V4: option to preserve <think>/</think> markers in chat completion content
feature requestNew feature or requestNew feature or requestStatus: Open.#51223 In vllm-project/vllm;- Status: Open.#51220 In vllm-project/vllm;
- Status: Open.#51214 In vllm-project/vllm;
- Status: Open.#51212 In vllm-project/vllm;
- Status: Open.#51211 In vllm-project/vllm;
[Bug]: Inkling-Small-NVFP4 wont start on 2 H200 GPUs.
bugSomething isn't workingSomething isn't workingStatus: Open.#51205 In vllm-project/vllm;