Published event
Hardware
PolicyChange
1 source(s)
v0.27.0
Summary
v0.27.0 vllm-project / vllm Public Uh oh! There was an error while loading.
Why it matters
This PolicyChange is relevant to the technology intelligence record because it involves DeepSeek, NVIDIA, Meta, Cohere. The source article should remain the factual reference for follow-up coverage.
Key facts
- vllm-project / vllm Public Uh oh!
- There was an error while loading.
- Notifications You must be signed in to change notification settings Fork 22.7k Star 92.7k v0.27.0 khluu released this 10 Aug 21:18 · 2556 commits to main since this release v0.27.0 4bdc8a7 vLLM v0.27.0 Release Notes Highlights This release features 561 commits from 242 contributors (64 new)!
- Kimi K3 support with a full stack landing in one release: core model files and kernels ( #50089 , #50000 ), Python ( #50093 ) and Rust ( #50104 ) frontends, AttnRes kernels ( #50090 ), DeepGEMM support ( #50458 ), compressed-tensors quantized checkpoints ( #50500 ), DSpark AR fusion ( #50242 ), and an option to shard the shared expert instead of replicating it ( #50656 ).
- More new models : Qwen3.5 text-only dense and MoE models ( #50210 ) with EVS video token pruning ( #48912 ), K-EXAONE-2.0-750B-A37B ( #50524 ), VaultGemma via the Transformers modeling backend ( #49803 ), and jina-embeddings-v5-text-nano ( #50688 ).
- PyTorch 2.13.0 upgrade along with torchvision 0.28.0 and Triton 3.7.1 ( #48155 ) — this is a breaking environment change; XPU ( #48677 ) and CPU ( #50412 ) followed to torch 2.13 as well.
Entities in this story
Related events