This release adds 3 notable features for engineering teams evaluating rollout.
Published 1mo
LLM Frameworks
✓ No known CVEs patched
✓ No known CVEs patched in this version
Topics
ai
apple-silicon
benchmarks
cli
gguf
gpu
+7 more
huggingface
inference
llm
local-llm
ollama
python
vram
Summary
AI summaryAdded Markdown ranking output and new speed/fit filtering options with improved color thresholds.
Full changelog
Added
- Markdown ranking output with
--markdown/-mfor pasteable GitHub issues, READMEs, Slack, and Discord. - Runtime-first ranking tables now show memory, estimated speed, fit type, and published date by default.
--speed any|usable|fastand the shorter--fit gpualias for full-GPU recommendations.--vram-headroomand--ram-budgetfor safer fit planning when runtimes or background processes need memory.
Changed
- Speed colors now reflect practical generation speed: red under 4 tok/s, yellow from 4-10, green from 10-30, and bright green at 30+ tok/s.
--detailsrestores the download-focused metadata table when needed.
Fixed
- Invalid
--vram-headroomvalues are now rejected even in CPU-only runs.
Weekly OSS security release digest.
The CVE patches and breaking changes that affected production tools this week. One email, every Sunday.
No spam, unsubscribe anytime.
Share this release
Track Find the best local LLM for your hardware, ranked by benchmarks
Get notified when new releases ship.
Sign up freeAbout Find the best local LLM for your hardware, ranked by benchmarks
All releases →Related context
Related tools
Beta — feedback welcome: [email protected]