Local AI.
Infinite Scale.
A highly-scalable distributed inference architecture. It features a high-throughput Go gateway orchestrating real-time SSE streams from decoupled PyTorch gRPC workers, paired with Supabase for seamless edge-to-cloud state synchronization.