Matchboxmatchbox
← Back to match

llama-optimus

Automatically finds the best llama.cpp performance flags for your hardware

Desktopfreeglobal

llama-optimus is a free, open-source Python tool that automatically tunes llama.cpp's performance flags using Optuna, finding the settings that give the best tokens-per-second on a user's specific hardware. It's aimed at local LLM users running llama.cpp who don't want to manually trial-and-error dozens of performance flags.

Categories
AI toolsdeveloper toolsperformance optimization

Full match profile

Behind the summary, Matchbox keeps a richer profile of llama-optimus - the signals our matcher actually reads to decide when to surface it. It stays private; claim the listing to see and control it.

  • Problem & pain-point mapping
  • Who we surface it to (audience fit)
  • What it's a strong alternative to
  • Trust & credibility signals

Try Matchbox with your own problem

Describe what is not working - we’ll show you whether llama-optimus (or something else) actually fits.