Matchboxmatchbox
← Back to match

Reame

CPU-first LLM inference server built to run useful models on cheap ARM and shared-vCPU hardware

PlatformWebfreeglobal

Reame is a CPU-first LLM inference server built on llama.cpp, designed to run usable models efficiently on cheap hardware like shared vCPUs, free-tier cloud boxes, and 2-core ARM machines rather than requiring a GPU. It exposes an OpenAI-compatible API, aimed at developers who want to self-host an LLM on modest CPU-only hardware.

Categories
AI infrastructurelocal LLM tooling

Full match profile

Behind the summary, Matchbox keeps a richer profile of Reame - the signals our matcher actually reads to decide when to surface it. It stays private; claim the listing to see and control it.

  • Problem & pain-point mapping
  • Who we surface it to (audience fit)
  • What it's a strong alternative to
  • Trust & credibility signals

Try Matchbox with your own problem

Describe what is not working - we’ll show you whether Reame (or something else) actually fits.