← Back to match
vllm-mlx
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon, with continuous batching.
Desktopfree
vllm-mlx is a free, open-source high-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon, with continuous batching. It's built for developers running local LLM inference on Apple Silicon, filling a gap most generic tools don't address directly.
Categories
local ai infrastructuredeveloper tools
Something wrong with this listing — dead link, not a real product, wrong info?

