← Back to match
llmux
Run and manage vLLM and llama.cpp model servers from one terminal UI and CLI.
Desktopfreeglobal
llmux runs and manages both vLLM and llama.cpp model servers from one terminal UI and CLI, with shared Docker Compose profiles, live tokens-per-second stats and cross-backend port/GPU conflict checks. It's aimed at developers running both inference backends who want unified management instead of separate tooling for each.
Categories
AI ToolsDeveloper Tools

