Run large language models across several consumer machines over ordinary Ethernet. A model that fits in no single GPU you own can fit across all of them.

consumer-hardware distributed-inference edge-ai gpu llama-cpp llm local-llm pipeline-parallelism
4 Open Issues Need Help Last updated: Aug 15, 2026

Open Issues Need Help

View All on GitHub
help wanted good first issue hardware-report

Run large language models across several consumer machines over ordinary Ethernet. A model that fits in no single GPU you own can fit across all of them.

Shell
#consumer-hardware#distributed-inference#edge-ai#gpu#llama-cpp#llm#local-llm#pipeline-parallelism
help wanted good first issue

Run large language models across several consumer machines over ordinary Ethernet. A model that fits in no single GPU you own can fit across all of them.

Shell
#consumer-hardware#distributed-inference#edge-ai#gpu#llama-cpp#llm#local-llm#pipeline-parallelism
help wanted good first issue

Run large language models across several consumer machines over ordinary Ethernet. A model that fits in no single GPU you own can fit across all of them.

Shell
#consumer-hardware#distributed-inference#edge-ai#gpu#llama-cpp#llm#local-llm#pipeline-parallelism

Run large language models across several consumer machines over ordinary Ethernet. A model that fits in no single GPU you own can fit across all of them.

Shell
#consumer-hardware#distributed-inference#edge-ai#gpu#llama-cpp#llm#local-llm#pipeline-parallelism