← Back to Model Beat
Models·Jun 5·all news from June 5, 2026

ollama v0.30.6

# New models - [Gemma 4 QAT weights](https://ollama.com/library/gemma4): the Gemma 4 family is now optimized with Quantization-Aware Training (QAT) to dramatically reduce memory requirements and maximize on-device performance. Look for the tags ending in : - - - - - ## What's Changed * now integrates with [Oh My Pi](https://omp.sh), an AI coding agent with IDE integration * MLX embedding layers now use NVFP4 global scale for improved quantization on Apple Silicon **Full Changelog**: https://github.com/ollama/ollama/compare/v0.30.5...v0.30.6

Covered by 1 source

Related stories

ModelsClaude Fable 5 and Claude Mythos 5 - AnthropicJun 9 · 4 sourcesModelsAlibaba's Qwen Team Launches Qwen3.7-Plus, Adding Vision, Deep Reasoning, Tool Invocation, and Autonomous Iteration on the Bailian Platform - MarkTechPostJun 2 · 3 sourcesModelsFluid, natural voice translation with Gemini 3.5 Live TranslateJun 9ModelsIntroducing Gemma 4 12B: a unified, encoder-free multimodal modelJun 9