
Mesh LLM: Run 235B Models Across Your Home Lab with iroh
A new distributed inference system pools GPU resources across multiple machines and exposes them through a single OpenAI-compatible API. No RDMA, no NVLink - just QUIC and your existing hardware.












