Prometheus Inference Platform
Self-hosted platform for running LLM inference across your own hardware
PlatformWebfreeglobal

Prometheus Inference Platform is a self-hosted, production-grade LLM inference platform combining a FastAPI gateway with authentication and per-model access scopes, a multi-backend model manager across llama.cpp, MLX, vLLM and SGLang on distributed hosts, an admin dashboard, and full tracing. It's built for teams that want to run LLM inference on their own hardware with production controls, rather than depending entirely on a cloud AI API.
Categories
AI infrastructureself-hosted LLM tools
Something wrong with this listing — dead link, not a real product, wrong info?

