Published event
ArtificialIntelligence ProductUpdate 1 source(s)

v0.32.6

Updated September 26, 2026 · 2:47 PM · source date August 4, 2026

Summary

v0.32.6 ollama / ollama Public Notifications You must be signed in to change notification settings Fork 18k Star 182k v0.32.6 github-actions released this 04 Aug 18:49 · 198 commits to main since this release v0.32.6 c82ebbd This commit was created on GitHub.com and signed with GitHub’s verified signature . GPG key ID: B5690EEEBB952194 Verified Learn about vigilant mode .

Why it matters

This ProductUpdate is relevant to the technology intelligence record because it involves Apple, OpenAI, GitHub, llama. The source article should remain the factual reference for follow-up coverage.

Key facts
  • ollama / ollama Public Notifications You must be signed in to change notification settings Fork 18k Star 182k v0.32.6 github-actions released this 04 Aug 18:49 · 198 commits to main since this release v0.32.6 c82ebbd This commit was created on GitHub.com and signed with GitHub’s verified signature .
  • GPG key ID: B5690EEEBB952194 Verified Learn about vigilant mode .
  • What's Changed Qwen3.5 is faster on Apple GPUs: the MLX engine now uses the model's MTP head for speculative decoding automatically /v1/chat/completions streaming now matches OpenAI's wire format: role only on the first chunk, finish_reason on its own chunk, and usage in a separate chunk with stream_options.include_usage .
  • Truncated OpenAI responses now report finish_reason: "length" instead of "tool_calls" .
  • ollama run kimi-k3 now offers kimi-k3:cloud for cloud-only models that publish no default tag, instead of failing.
  • TUI fixes: pipe-delimited prose no longer renders as a table, Enter accepts the highlighted @ file completion, and /prompt scrolling is no longer laggy.
Entities in this story
Related events