← Back to match
ggrun
llama.cpp launcher that places big MoE models across mismatched multi-GPU rigs using exact VRAM math.
Desktopfreeglobal
ggrun is a free, open-source launcher for llama.cpp and ik_llama.cpp that calculates exact VRAM, RAM, and per-GPU bandwidth to place large mixture-of-experts models across a mismatched multi-GPU rig, rather than requiring identical GPUs. It's aimed at local LLM hobbyists trying to run big models on hardware they've assembled over time rather than a uniform GPU setup.
Categories
developer toolsai infrastructure

