Matchboxmatchbox
← Back to match

llamactl

Unified management and routing dashboard for llama.cpp, MLX, and vLLM local LLM models

PlatformWebfreeglobal

llamactl is a self-hosted unified management and routing dashboard for local LLM inference across llama.cpp, MLX, and vLLM backends, with a built-in model downloader for Hugging Face models. It supports dynamic multi-model instances with on-demand loading, automatic idle timeout, and LRU eviction, aimed at developers running multiple local LLM backends who want to manage and route them from one place.

Categories
AI infrastructurelocal LLM tooling

Full match profile

Behind the summary, Matchbox keeps a richer profile of llamactl - the signals our matcher actually reads to decide when to surface it. It stays private; claim the listing to see and control it.

  • Problem & pain-point mapping
  • Who we surface it to (audience fit)
  • What it's a strong alternative to
  • Trust & credibility signals

Try Matchbox with your own problem

Describe what is not working - we’ll show you whether llamactl (or something else) actually fits.