Published event
Hardware
ApiUpdate
1 source(s)
v0.18.0
Summary
v0.18.0 vllm-project / vllm Public Uh oh! There was an error while loading.
Why it matters
This ApiUpdate is relevant to the technology intelligence record because it involves OpenAI, Mistral AI, DeepSeek, NVIDIA. The source article should remain the factual reference for follow-up coverage.
Key facts
- vllm-project / vllm Public Uh oh!
- There was an error while loading.
- Notifications You must be signed in to change notification settings Fork 22.7k Star 92.7k v0.18.0 khluu released this 20 Mar 21:31 · 7117 commits to main since this release v0.18.0 bcf2be9 vLLM v0.18.0 Known issues Degraded accuracy when serving Qwen3.5 with FP8 KV cache on B200 ( #37618 ) If you previously ran into CUBLAS_STATUS_INVALID_VALUE and had to use a workaround in v0.17.0 , you can reinstall torch 2.10.0 .
- PyTorch published an updated wheel that addresses this bug.
- Highlights This release features 445 commits from 213 contributors (61 new)!
- gRPC Serving Support : vLLM now supports gRPC serving via the new --grpc flag ( #36169 ), enabling high-performance RPC-based serving alongside the existing HTTP/REST interface.
Entities in this story
Related events