← Back to Model Beat
Models·2d ago·all news from August 19, 2026

ollama v0.32.15

## What's Changed * New desktop onboarding flow on first launch * Caches resolved model metadata between requests, cutting time-to-first-token by roughly half (TTFT dropped from ~995 ms to ~524 ms in benchmarks) * Fixes a bug where chat and generate could wedge after a mid-stream parser error * **Qwen 3.8** system messages are now normalized so non-leading system messages are handled consistently * MLX and llama.cpp dependency updates ## New Contributors * @gaugarg-nv made their first contribution in https://github.com/ollama/ollama/pull/17752 **Full Changelog**: https://github.com/ollama/ollama/compare/v0.32.14...v0.32.15

Covered by 1 source

Related stories

ModelsDeepSeek Unveils Test Model to Rival Anthropic’s Opus 4.8Aug 19 · 14 sourcesModelsIntroducing ChatGPT for Teens: Built for learning, backed by protectionsAug 18 · 7 sourcesModelsChina’s open-weight AI models are prompting US players to reconsider their strategy.Aug 16 · 20 sourcesModelsGPT-5.6 Sol drives OpenAI's revenue surge as it regains ground on AnthropicAug 20 · 2 sources