three.ws · Decentralized inference
Same prompt, two providers, real responses. The left card hits three.ws's default inference (Claude). The right card routes the identical request through a decentralized AI gateway whose orchestrators run open-weight models on commodity GPUs. Phase 4 of the roadmap: decoupling inference from any single vendor so agents stay online when one provider doesn't.
Settings · decentralized gateway model
three.ws default
Claude · Anthropic
latency —
in —
out —
awaiting prompt…
three.ws decentralized
Open-weight · decentralized
latency —
in —
out —
awaiting prompt…
ready.
Powered by Livepeer AI Gateway on the decentralized side. Claude reply provided by Anthropic. The public decentralized gateway is rate-limited and occasionally slow — typical p50 latency is 2–10 s for an 8B-param model.