Bug Vllm 051 Tensor Parallel 2 Hang Issue 6370 Vllm Project

[Bug]: vLLM 0.5.1 tensor parallel 2 hang · Issue #6370 · vllm-project ...
[Bug]: vLLM 0.5.1 tensor parallel 2 hang · Issue #6370 · vllm-project ...
vLLM Production Deployment 2026: Multi-GPU Tensor Parallel + FP8 Docker ...
vLLM Production Deployment 2026: Multi-GPU Tensor Parallel + FP8 Docker ...
Bug - vllm not working for L4 GPUs and tensor_parallel_size > 1 · Issue ...
Bug - vllm not working for L4 GPUs and tensor_parallel_size > 1 · Issue ...
Possible sampling parameter bug in VLLM Server · Issue #2754 · vllm ...
Possible sampling parameter bug in VLLM Server · Issue #2754 · vllm ...
Scaling LLM Inference: Data, Pipeline & Tensor Parallelism in vLLM ...
Scaling LLM Inference: Data, Pipeline & Tensor Parallelism in vLLM ...
[Bug]: Error when using tensor_parallel in v0.6.1 · Issue #8397 · vllm ...
[Bug]: Error when using tensor_parallel in v0.6.1 · Issue #8397 · vllm ...
VLLM tensor-parallel and RegexLogitsProcessor · Issue #524 · dottxt-ai ...
VLLM tensor-parallel and RegexLogitsProcessor · Issue #524 · dottxt-ai ...
[Bug]: VLLM 0.8.2 OOM error (No error in 0.7.3 version) · Issue #15664 ...
[Bug]: VLLM 0.8.2 OOM error (No error in 0.7.3 version) · Issue #15664 ...
[Bug]: vllm hangs after model download / load · Issue #7303 · vllm ...
[Bug]: vllm hangs after model download / load · Issue #7303 · vllm ...
[Bug]: tensor model parallel group is not initialized · Issue #3639 ...
[Bug]: tensor model parallel group is not initialized · Issue #3639 ...
Tensor Parallel on A10G - llamav2 · Issue #634 · vllm-project/vllm · GitHub
Tensor Parallel on A10G - llamav2 · Issue #634 · vllm-project/vllm · GitHub
[Bug]: Cannot use OLMoE with tensor parallel higher than 1 · Issue ...
[Bug]: Cannot use OLMoE with tensor parallel higher than 1 · Issue ...
[Bug]: v0.4.1 VLLM_USE_MODELSCOPE not working · Issue #4362 · vllm ...
[Bug]: v0.4.1 VLLM_USE_MODELSCOPE not working · Issue #4362 · vllm ...
[Bug]: Error After Model Load in vllm 0.7.0 (No Issue in vllm 0.6.6 ...
[Bug]: Error After Model Load in vllm 0.7.0 (No Issue in vllm 0.6.6 ...
[Bug]: vllm 0.8.3 serve error · Issue #15457 · vllm-project/vllm · GitHub
[Bug]: vllm 0.8.3 serve error · Issue #15457 · vllm-project/vllm · GitHub
[Bug]: vllm stopped at vLLM is using nccl==2.21.5 · Issue #16772 · vllm ...
[Bug]: vllm stopped at vLLM is using nccl==2.21.5 · Issue #16772 · vllm ...
[Bug]: 单gpu没有任何反应(设置tensor_parallel_size=1模型加载失败) · Issue #7136 · vllm ...
[Bug]: 单gpu没有任何反应(设置tensor_parallel_size=1模型加载失败) · Issue #7136 · vllm ...
[Bug]: VLLM_USE_V1=1 failed with deepseek-v3 · Issue #12956 · vllm ...
[Bug]: VLLM_USE_V1=1 failed with deepseek-v3 · Issue #12956 · vllm ...
[Bug]: requests with response_format cause vllm to hang with pipeline ...
[Bug]: requests with response_format cause vllm to hang with pipeline ...
[Bug]: Unable to serve Llama3 using vLLM Docker container · Issue #4725 ...
[Bug]: Unable to serve Llama3 using vLLM Docker container · Issue #4725 ...
[Bug]: vllm failed to run two instance with one gpu · Issue #10533 ...
[Bug]: vllm failed to run two instance with one gpu · Issue #10533 ...
[Bug]: Error while importing vllm since v0.6.6 · Issue #11683 · vllm ...
[Bug]: Error while importing vllm since v0.6.6 · Issue #11683 · vllm ...
bug: vllm-router latest version incompatible with vllm 0.10.0 · Issue ...
bug: vllm-router latest version incompatible with vllm 0.10.0 · Issue ...
[Bug]: vLLM 0.4.2 8xH100 init failed · Issue #5785 · vllm-project/vllm ...
[Bug]: vLLM 0.4.2 8xH100 init failed · Issue #5785 · vllm-project/vllm ...
[Bug]: vllm is crashed on v0.5.3.post1 · Issue #7161 · vllm-project ...
[Bug]: vllm is crashed on v0.5.3.post1 · Issue #7161 · vllm-project ...
[Bug]: Single-Node data parallel (--data-parallel-size=4) leads to vLLM ...
[Bug]: Single-Node data parallel (--data-parallel-size=4) leads to vLLM ...
[Bug]: vllm serve Exception in ASGI application · Issue #10215 · vllm ...
[Bug]: vllm serve Exception in ASGI application · Issue #10215 · vllm ...
[Bug]: Issue Running Qwen2.5-VL-7B-Instruct with vLLM Due to ...
[Bug]: Issue Running Qwen2.5-VL-7B-Instruct with vLLM Due to ...
[Bug]: vLLM crashes on tokenized embedding input · Issue #11375 · vllm ...
[Bug]: vLLM crashes on tokenized embedding input · Issue #11375 · vllm ...
[Bug]: Unable to embed any text using the vLLM CPU server · Issue #9379 ...
[Bug]: Unable to embed any text using the vLLM CPU server · Issue #9379 ...
[Bug]: vllm(v0.5.3)openai接口中,流式返回取消了usage信息?请问如何设置 · Issue #7179 · vllm ...
[Bug]: vllm(v0.5.3)openai接口中,流式返回取消了usage信息?请问如何设置 · Issue #7179 · vllm ...
[Bug]: vLLM CPU mode broken Unable to get JIT kernel for brgemm · Issue ...
[Bug]: vLLM CPU mode broken Unable to get JIT kernel for brgemm · Issue ...
vLLM running on a Ray Cluster Hanging on Initializing · Issue #2826 ...
vLLM running on a Ray Cluster Hanging on Initializing · Issue #2826 ...
[Bug]: VLLM Build Using Docker Error Deploy · Issue #15376 · vllm ...
[Bug]: VLLM Build Using Docker Error Deploy · Issue #15376 · vllm ...
Expert Parallelism and Mixed Parallelism Strategies in vLLM | Jarvis ...
Expert Parallelism and Mixed Parallelism Strategies in vLLM | Jarvis ...
Expert Parallelism and Mixed Parallelism Strategies in vLLM | Jarvis ...
Expert Parallelism and Mixed Parallelism Strategies in vLLM | Jarvis ...

Loading image details...

Source
Dimensions