← Back to Model Beat
Models·1d ago·all news from September 14, 2026

ollama v0.34.1

## What's Changed * MLX safetensors no longer experimental. GGUF model creation now requires using llama.cpp tooling for safetensor conversion and quantization. * Improved MLX memory handling on Apple Silicon * Runaway repeat token detection now requires 100 repeat tokens for reduced false positives (e.g. OCR) * is much faster on large model libraries (3.1 s → 294 ms cold in testing), and model capabilities are now reported consistently. * MLX and llama.cpp updates **Full Changelog**: https://github.com/ollama/ollama/compare/v0.34.0...v0.34.1-rc1

Covered by 1 source

Related stories

ModelsApple releases iOS 27 with Siri AI overhaulSep 14 · 6 sourcesModelsPerplexity trusts GPT-6 Astra with end-to-end systemsSep 12 · 4 sourcesModelsIntroducing Gemini 3.8 Live and 3.8 Live Extended ThinkingSep 15 · 2 sourcesModelsAnthropic Says Yemeni Cell Used Claude in Missile DevelopmentSep 11 · 6 sources