← Back to match
Squish
Fastest way to run local LLMs on Apple Silicon, sub-second model loads.
Desktopfreeglobal
Squish is a local LLM runtime for Apple Silicon offering sub-second model loads and better throughput, tail latency and full-response time than Ollama, with an OpenAI/Ollama-compatible interface and no cloud or API keys. It's aimed at Mac users who want the fastest possible local LLM inference on their own hardware.
Categories
AI Tools

