Skip to content
#

distributed-llm

Here are 8 public repositories matching this topic...

Run a model too big for one GPU in LM Studio across two PCs over LAN (distributed inference via llama.cpp RPC): two DLLs + one env var GGML_RPC_SERVERS, no LM Studio changes. Windows/Vulkan/AMD tested. Home-lab hack by Robert + Kod (AI).

  • Updated Sep 16, 2026
  • C++

Add this topic to your repo

To associate your repository with the distributed-llm topic, visit your repo's landing page and select "manage topics."

Learn more