← Back to Model Beat
Models·Jun 12·all news from June 12, 2026

ollama v0.30.8

## What's Changed * Fixed selecting the wrong provider in some cases * Improved prompt caching by decoupling it from context shift for better KV cache reuse * More stable MLX inference with hardened linear and embedding layers * MLX runner now creates snapshots during prompt processing and speculative decoding for improved reliability * Improved recurrent model support with per-boundary states from the gated-delta kernels **Full Changelog**: https://github.com/ollama/ollama/compare/v0.30.7...v0.30.8

Covered by 1 source

Related stories

ModelsDiffusionGemma: 4x faster text generationJun 10 · 6 sourcesModelsClaude Fable 5 and Claude Mythos 5 - AnthropicJun 9 · 4 sourcesModelsPredicting model behavior before release by simulating deploymentJun 16 · 3 sourcesModelsFrom Chatbots to Collaborators: AI’s Next EraJun 15 · 39 sources