Skip to content

This release adds 3 notable features for engineering teams evaluating rollout.

Published 1mo LLM Frameworks
✓ No known CVEs patched
Read the diff → Tool health → What is this tool? →

✓ No known CVEs patched in this version

Topics

ai apple-silicon benchmarks cli gguf gpu
+7 more
huggingface inference llm local-llm ollama python vram

Summary

AI summary

Added Markdown ranking output and new speed/fit filtering options with improved color thresholds.

Full changelog

Added

  • Markdown ranking output with --markdown / -m for pasteable GitHub issues, READMEs, Slack, and Discord.
  • Runtime-first ranking tables now show memory, estimated speed, fit type, and published date by default.
  • --speed any|usable|fast and the shorter --fit gpu alias for full-GPU recommendations.
  • --vram-headroom and --ram-budget for safer fit planning when runtimes or background processes need memory.

Changed

  • Speed colors now reflect practical generation speed: red under 4 tok/s, yellow from 4-10, green from 10-30, and bright green at 30+ tok/s.
  • --details restores the download-focused metadata table when needed.

Fixed

  • Invalid --vram-headroom values are now rejected even in CPU-only runs.

Weekly OSS security release digest.

The CVE patches and breaking changes that affected production tools this week. One email, every Sunday.

No spam, unsubscribe anytime.

Share this release

Track Find the best local LLM for your hardware, ranked by benchmarks

Get notified when new releases ship.

Sign up free

About Find the best local LLM for your hardware, ranked by benchmarks

All releases →

Beta — feedback welcome: [email protected]