Published event
Hardware OpenSourceRelease 1 source(s)

v0.9.1

Updated September 26, 2026 · 2:46 PM · source date June 10, 2025

Summary

v0.9.1 vllm-project / vllm Public Uh oh! There was an error while loading.

Why it matters

This OpenSourceRelease is relevant to the technology intelligence record because it involves DeepSeek, GitHub, Amazon Web Services, AMD. The source article should remain the factual reference for follow-up coverage.

Key facts
  • vllm-project / vllm Public Uh oh!
  • There was an error while loading.
  • Notifications You must be signed in to change notification settings Fork 22.7k Star 92.7k v0.9.1 github-actions released this 10 Jun 18:30 · 14957 commits to main since this release v0.9.1 b6553be This commit was created on GitHub.com and signed with GitHub’s verified signature .
  • GPG key ID: B5690EEEBB952194 Verified Learn about vigilant mode .
  • ( #18284 ), Add Multi-Modal model support for Neuron ( #18921 ), Support quantization on neuron ( #18283 ) Platform: Make torch distributed process group extendable ( #18763 ) Engine features Add Lora Support to Beam Search ( #18346 ) Add rerank support to run_batch endpoint ( #16278 ) CLI: add run batch ( #18804 ) Server: custom logging ( #18403 ), allowed_token_ids in ChatCompletionRequest ( #19143 ) LLM API: make use_tqdm accept a callable for custom progress bars ( #19357 ) perf: [KERNEL] Sampler.
  • by @aws-satyajith in #18284 [LoRA] Add LoRA support for InternVL by @jeejeelee in #18842 [Doc] Remove redundant spaces from compatibility_matrix.md by @windsonsea in #18891 [doc] add CLI doc by @reidliu41 in #18871 [Bugfix] Fix misleading information in the documentation by @jeejeelee in #18845 [Misc] Replace TODO in serving transcription by @NickLucche in #18895 [Bugfix] Ensure tensors are contiguous during serialisation by @lgeiger in #18860 [BugFix] Update pydantic to fix error on python 3.10 by @ProExpertProg in #18852 Fix an error in dummy weight loading for quantization models by @Chenyaaang in #18855 [Misc][Tools][Benchmark] Add benchmark_serving supports for llama.cpp.
Entities in this story
Related events