v0.30.6
Summary
<h1>New models</h1> <ul> <li><a href="https://ollama.com/library/gemma4" rel="nofollow">Gemma 4 QAT weights</a>: the Gemma 4 family is now optimized with Quantization-Aware Training (QAT) to dramatically reduce memory requirements and maximize on-device performance. Look for the tags ending in <code>-qat</code>: <ul> <li><code>gemma4:e2b-it-qat</code></li> <li><code>gemma4:e4b-it-qat</code></li> <li><code>gemma4:12b-it-qat</code></li> <li><code>gemma4:26b-a4b-it-qat</code></li> <li><code>gemma4:31b-it-qat</code></li> </ul> </li> </ul> <h2>What's Changed</h2> <ul> <li><code>ollama launch omp</code> now integrates with <a href="https://omp.sh" rel="nofollow">Oh My Pi</a>, an AI coding agent with IDE integration</li> <li>MLX embedding layers now use NVFP4 global scale for improved quantization on Apple Silicon</li> </ul> <p><strong>Full Changelog</strong>: <a class="commit-link" href="https://github.com/ollama/ollama/compare/v0.30.5...v0.30.6"><tt>v0.30.5...v0.30.6</tt></a></p>
News Radar provides aggregated summaries. Full content and copyright remain with the original publisher.