This release includes 1 security fix for security teams reviewing exposed deployments.
Published 20d
Model Serving & MLOps
✓ No known CVEs patched
This release patches 1 known CVE
Topics
deepseek
gemma
gemma3
glm
go
gpt-oss
+8 more
llama
llama3
llm
llms
minimax
mistral
ollama
qwen
Summary
AI summaryFixed loading of models when file paths contain non‑UTF‑8 characters.
Full changelog
What's Changed
- Enabled flash attention on older NVIDIA GPUs (compute capability 6.x)
- iGPU can now offload vision models with padding to fit available memory
- Fixed structured output for thinking models when thinking is disabled
- Hardened GGUF model creation
ollama launchfor Claude Code now disables telemetry by default- Fixed loading models on paths with non-UTF-8 characters
- Updated the MLX and llama.cpp engines
New Contributors
- @kevinpark1217 made their first contribution in https://github.com/ollama/ollama/pull/16949
Full Changelog: https://github.com/ollama/ollama/compare/v0.31.1...v0.31.2
Security Fixes
- Hardened GGUF model creation
Weekly OSS security release digest.
The CVE patches and breaking changes that affected production tools this week. One email, every Sunday.
No spam, unsubscribe anytime.
Share this release
About ollama
Get up and running with Kimi-K2.5, GLM-5, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Beta — feedback welcome: [email protected]