This release adds 3 notable features for engineering teams evaluating rollout.
Published 1mo
Model Serving & MLOps
✓ No known CVEs patched
✓ No known CVEs patched in this version
Topics
deepseek
gemma
gemma3
glm
go
gpt-oss
+8 more
llama
llama3
llm
llms
minimax
mistral
ollama
qwen
Summary
AI summaryUpdates llm, launch, and llama across a mixed release.
Full changelog
What's Changed
- launch: add thinking capability detection to opencode by @hoyyeva in https://github.com/ollama/ollama/pull/15434
- launch: auto-install Claude Code by @hoyyeva in https://github.com/ollama/ollama/pull/16802
- launch: auto-install opencode when missing by @hoyyeva in https://github.com/ollama/ollama/pull/16806
- discover: fix inverted iGPU/dGPU Vulkan classification on Windows hybrid graphics by @Sahil170595 in https://github.com/ollama/ollama/pull/16669
- mlxrunner: unify and tune speculative decoding by @jessegross in https://github.com/ollama/ollama/pull/16791
- launch/codex: detect model drift when Codex App UI switches by @BruceMacD in https://github.com/ollama/ollama/pull/16864
- llama: add sm_86 architecture to cuda_v13_windows preset by @anishesg in https://github.com/ollama/ollama/pull/16834
- llm: size mmproj offload by projector memory by @dhiltgen in https://github.com/ollama/ollama/pull/16866
- docs: document max think level by @ParthSareen in https://github.com/ollama/ollama/pull/16877
- llm: preserve generation headroom for shifted prompts by @ParthSareen in https://github.com/ollama/ollama/pull/16856
- llama: default qwen2.5vl window attention metadata by @dhiltgen in https://github.com/ollama/ollama/pull/16868
- llm: use host Vulkan loader on Windows by @dhiltgen in https://github.com/ollama/ollama/pull/16869
- mlx: update and fix CUDA JIT packaging by @dhiltgen in https://github.com/ollama/ollama/pull/16871
- llm: fix ollama ps double-counting mmap'd weights on partial offload by @discobot in https://github.com/ollama/ollama/pull/16709
- docs: redesign docs landing and integrations overview by @hoyyeva in https://github.com/ollama/ollama/pull/16807
- server: align generate with native chat templates by @dhiltgen in https://github.com/ollama/ollama/pull/16878
- jetson: add CC 87 for CUDA v13 by @dhiltgen in https://github.com/ollama/ollama/pull/16628
- llama.cpp version update by @dhiltgen in https://github.com/ollama/ollama/pull/16548
New Contributors
- @Sahil170595 made their first contribution in https://github.com/ollama/ollama/pull/16669
- @anishesg made their first contribution in https://github.com/ollama/ollama/pull/16834
- @discobot made their first contribution in https://github.com/ollama/ollama/pull/16709
Full Changelog: https://github.com/ollama/ollama/compare/v0.30.10...v0.30.11-rc0
Weekly OSS security release digest.
The CVE patches and breaking changes that affected production tools this week. One email, every Sunday.
No spam, unsubscribe anytime.
Share this release
About ollama
Get up and running with Kimi-K2.5, GLM-5, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Beta — feedback welcome: [email protected]