Release tracker
Verified local LLM release changes
Scan exact versions, dates, change types, verification state, and primary evidence in one place.
Verified means checked against the cited source. Vendor performance claims are attributed, not independently reproduced.
- DateReleasedReleasev0.33.2macOS app and Claude Desktop proxy fixesRelease trackerEntities
- Ollama
ChangesVerificationVerifiedCheckedBreakingUnknownSummaryOllama restores following the system appearance, hands the macOS app off to an already-running instance, and prevents model-catalog updates from interrupting in-flight Claude Desktop proxy requests.
Evidence - DateReleasedReleasev0.33.0Claude Desktop gateway, caching, and packaging fixesRelease trackerEntities
- Ollama
ChangesVerificationVerifiedCheckedBreakingUnknownSummaryOllama's v0.33.0 release lets Claude Desktop use Ollama as a third-party gateway provider, fixes prompt caching for recurrent-layer models, and fixes Linux and Windows packaging.
Evidence- v0.33.0Ollama
- DateReleasedReleasev0.32.15Ollama reports v0.32.15 caching model metadata and normalizing Qwen 3.8 system messagesRelease trackerEntities
- Ollama
- Qwen 3.8
ChangesVerificationVerifiedCheckedBreakingUnknownSummaryOllama reports that v0.32.15 caches resolved model metadata, fixes chat and generate wedging after a parser error, and normalizes Qwen 3.8 system messages.
Evidence- v0.32.15Ollama
- DateReleasedReleasev0.32.14WebP transcoding and Qwen system-message handlingRelease trackerEntities
- Ollama
ChangesVerificationVerifiedCheckedBreakingUnknownSummaryThe official notes list WebP transcoding for llama-server and Qwen renderer tolerance for non-leading system messages.
Evidence- v0.32.14Ollama
- DateReleasedReleasev0.32.12Qwen3.8-27B supportRelease trackerEntities
- Qwen 3.8 27B
- Ollama
ChangesVerificationVerifiedCheckedBreakingUnknownSummaryThe release adds Qwen3.8-27B support, with Ollama stating particular optimizations for Apple Silicon.
Evidence- v0.32.12Ollama
- DateReleasedReleasev0.32.10NVFP4 prefill and repeat-penalty defaultsRelease trackerEntities
- Ollama
ChangesVerificationVerifiedCheckedBreakingUnknownSummaryOllama reports about 7–8% faster prefill on the Qwen3.6 and Muse Glimmer NVFP4 MLX models with a global scale, changes the default to 1.0 instead of 1.1 for models without `repeat_penalty`, and fixes a shared-digest blob-verification case.
Evidence- v0.32.10Ollama
- DateReleasedReleasev0.32.9Nemotron 3 architecture supportRelease trackerEntities
- NVIDIA Nemotron 3.5 Lightning
- Ollama
ChangesVerificationVerifiedCheckedBreakingUnknownSummaryOllama's release notes add the Nemotron 3 architecture and describe NVIDIA Nemotron 3.5 Lightning as an open 30B MoE model with 3B active parameters.
Evidence- v0.32.9Ollama