Skip to content

ollama

v0.30.8 Feature

This release adds 4 notable features for engineering teams evaluating rollout.

Published 1mo Model Serving & MLOps
✓ No known CVEs patched
Read the diff → Tool health → What is this tool? →

✓ No known CVEs patched in this version

Topics

deepseek gemma gemma3 glm go gpt-oss
+8 more
llama llama3 llm llms minimax mistral ollama qwen

Summary

AI summary

Fixed ollama launch provider selection bug.

Changes in this release

Feature Low

MLX runner now creates snapshots during prompt processing and speculative decoding for improved reliability

MLX runner now creates snapshots during prompt processing and speculative decoding for improved reliability

Source: llm_adapter@2026-06-12

Confidence: high

Feature Low

Improves recurrent model support with per-boundary states from the gated-delta kernels

Improves recurrent model support with per-boundary states from the gated-delta kernels

Source: llm_adapter@2026-06-12

Confidence: high

Performance Low

Improves prompt caching by decoupling it from context shift for better KV cache reuse

Improves prompt caching by decoupling it from context shift for better KV cache reuse

Source: llm_adapter@2026-06-12

Confidence: high

Performance Low

Increases stability of MLX inference with hardened linear and embedding layers

Increases stability of MLX inference with hardened linear and embedding layers

Source: llm_adapter@2026-06-12

Confidence: high

Bugfix Medium

Fixes `ollama launch` selecting the wrong provider in some cases

Fixes `ollama launch` selecting the wrong provider in some cases

Source: llm_adapter@2026-06-12

Confidence: high

Full changelog

What's Changed

  • Fixed ollama launch selecting the wrong provider in some cases
  • Improved prompt caching by decoupling it from context shift for better KV cache reuse
  • More stable MLX inference with hardened linear and embedding layers
  • MLX runner now creates snapshots during prompt processing and speculative decoding for improved reliability
  • Improved recurrent model support with per-boundary states from the gated-delta kernels

Full Changelog: https://github.com/ollama/ollama/compare/v0.30.7...v0.30.8

Weekly OSS security release digest.

The CVE patches and breaking changes that affected production tools this week. One email, every Sunday.

No spam, unsubscribe anytime.

Share this release

Track ollama

Get notified when new releases ship.

Sign up free

About ollama

Get up and running with Kimi-K2.5, GLM-5, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.

All releases →

Related context

Related tools

Beta — feedback welcome: [email protected]