Distributed AI compute across everyday computers. Contribute to nibia-ai/fabric development by creating an account on GitHub.
I built NIBIA Fabric, an open-source distributed LLM inference system that pools CPU and RAM across macOS, Linux, and Windows machines.
It builds on llama.cpp, with capacity-aware scheduling, adaptive memory reservation, persistent tensor caching, and an OpenAI-compatible API.
I’ve validated the current alpha on three physical machines across several models, including a 30B-class Qwen model, as well as GPT-OSS 20B. Feedback on the architecture and use cases is very welcome.
3 comments
H4kken2 days ago
Sounds pretty cool, do you have the link to the github so we could take a look ?
Also, who do you plan on targeting with this promising project? like idk if the average ai user nowadays would like to bother with self hosting everything, but i might be wrong
i feel like a more enterprise oriented scope could interest more ppl
kickstartaioffi2 days ago
very interesting
abelop2 days ago
Thanks! Happy to answer any questions about how it works or the architecture.
Read the full thread on Hacker News →
Related stories
- The Verge · 0 points · 3 days ago
- Can you forget how you feel about Meta?theverge.comThe Verge · 0 points · 9 days ago
- The Verge · 0 points · 5 days ago
- Can John Ternus find Apple’s next big thing?theverge.comThe Verge · 0 points · 10 days ago
- Hacker News · 73 points · 9 days ago
- The Verge · 0 points · 12 days ago