Lumabri Enables Napster-like Peer-to-Peer LLM Distribution

TL;DR. Lumabri introduces a pure C engine allowing peer-to-peer distribution and inference of large Mixture-of-Experts models. - The system lets any machine serve or chat with a model, with inference data streamed on demand. - Local mirroring ensures subsequent inferences are served quickly, even if the original server goes offline. - The engine operates identically on CPUs and GPUs, widening access for model contribution and use.

Sources

Back to QLANKR News