Distributed Post-Quantum LLM Inference
No GPU? No problem. Petals shards massive model layers across the Polygone network. Your machine handles a slice. The network handles the rest.
Each relay transfers encrypted hidden states (ML-KEM-1024) to the next. No relay sees the full model output. Computation is blind.