Voice agents on CPU vs the GPU incumbents: latency, cost, and a deliberate comparison
Two numbers frame this post. On Lokutor, a production voice agent starts speaking in 139 ms at the median (TTS streaming time to first audio, one CPU thread), and a full minute of a running agent cost
lokutor.hashnode.dev4 min read