Run a model too big for one GPU in LM Studio across two PCs over LAN (distributed inference via llama.cpp RPC): two DLLs + one env var GGML_RPC_SERVERS, no LM Studio changes. Windows/Vulkan/AMD tested. Home-lab hack by Robert + Kod (AI).
-
Updated
Sep 16, 2026 - C++