Published event
ArtificialIntelligence
ProductUpdate
1 source(s)
v0.32.6
Summary
v0.32.6 ollama / ollama Public Notifications You must be signed in to change notification settings Fork 18k Star 182k v0.32.6 github-actions released this 04 Aug 18:49 · 198 commits to main since this release v0.32.6 c82ebbd This commit was created on GitHub.com and signed with GitHub’s verified signature . GPG key ID: B5690EEEBB952194 Verified Learn about vigilant mode .
Why it matters
This ProductUpdate is relevant to the technology intelligence record because it involves Apple, OpenAI, GitHub, llama. The source article should remain the factual reference for follow-up coverage.
Key facts
- ollama / ollama Public Notifications You must be signed in to change notification settings Fork 18k Star 182k v0.32.6 github-actions released this 04 Aug 18:49 · 198 commits to main since this release v0.32.6 c82ebbd This commit was created on GitHub.com and signed with GitHub’s verified signature .
- GPG key ID: B5690EEEBB952194 Verified Learn about vigilant mode .
- What's Changed Qwen3.5 is faster on Apple GPUs: the MLX engine now uses the model's MTP head for speculative decoding automatically /v1/chat/completions streaming now matches OpenAI's wire format: role only on the first chunk, finish_reason on its own chunk, and usage in a separate chunk with stream_options.include_usage .
- Truncated OpenAI responses now report finish_reason: "length" instead of "tool_calls" .
- ollama run kimi-k3 now offers kimi-k3:cloud for cloud-only models that publish no default tag, instead of failing.
- TUI fixes: pipe-delimited prose no longer renders as a table, Enter accepts the highlighted @ file completion, and /prompt scrolling is no longer laggy.
Entities in this story
Related events