VCVijay Choudharyinbettechmagnetics.hashnode.dev·3d ago · 22 min readFrom Event to Endpoint: Engineering a Real-Time Sports Data Architecture That Doesn't LagA goal is scored. Eleven seconds later, your app still shows 0 to 0. Meanwhile a user in another tab sees 1 to 0, and a third user sees the goal twice. Nothing crashed. Every service was up. The archi00
VCVijay Choudharyinbettechmagnetics.hashnode.dev·3d ago · 16 min readWhy "Sub-50ms" Claims Are Mostly Meaningless Without Knowing Which Latency They MeasureEvery API landing page has a number on it. "Sub-50ms." "Real-time." "Lightning fast." It sits next to a green dot and a globe animation, and it is supposed to settle the argument before you have writt00
MUMohammed Usmaniinmohammed-usmani.hashnode.dev·Sep 29 · 11 min readFrom 2.44 s to 1.81 s: working through my voice agent's latency list, and the fast model that went silentTL;DR Part 1 ended with a list of things to try. I tried all of them on the same phone agent, now running a recruiter's pre-screening call. At the agent, the median reply went from 2.44 s to 1.81 s 00
RVRutu Varhadiinmy-networkingblog.hashnode.dev·Sep 28 · 5 min readCan Network Latency be used as a 'Distance Sensor'?Imagine you are sitting in Mumbai; you send a data packet to the other side of the world and get a response within a fraction of a second. You know exactly how long the journey took. But can you tell 10
MUMohammed Usmaniinmohammed-usmani.hashnode.dev·Sep 26 · 11 min readLiveKit turn detection on real phone calls: why a 2-second average hid a 3.6-second stallTL;DR I built a voice agent you can call from a real phone (SIP through LiveKit Cloud) and measured every turn instead of trusting averages. On a test call the average reply took about 2 seconds, bu00
NHNasim Hossain Rabbiinblog.nasimhossain.dev·Sep 14 · 10 min readLittle's Law: The Formula Behind Every Good System Design InterviewYou're in a system design interview and you're asked to design a banking application. Customers can: Login View accounts Check transactions Transfer money Download statements The interviewer ad10
APabhishek pareekinabhishekpareek.dev·Aug 20 · 7 min readReading Latency Effectively TL;DR Mean latency describes total latency averaged across requests.p50 describes the middle of the distribution.p90 and p99 describe progressively higher boundaries in the tail.The latency SLO measur10
Mmarcuscheninvoicelatency.hashnode.dev·Aug 10 · 8 min readTwo weeks before launch, every turn was green and the call still diedThe dashboard was a wall of green. Word error rate under 5 percent. Intent classification at 94 percent on our eval set. Response appropriateness, graded by a rubric we trusted, sitting comfortably in00
Mmarcuscheninvoicelatency.hashnode.dev·Aug 10 · 8 min readThe guardrail fired at 1.4 seconds. The caller had heard the sentence at 1.1.Two weeks ago I wrote about putting a guardrail in front of our voice agent, on the input, where a caller had talked the model out of its own refund policy. This is the other half of that job, and it 00
RBRekha Battuinwhy-it-works.hashnode.dev·Aug 6 · 5 min readWhy a Fast CPU Still Spends Most of Its Time WaitingLearning System Design through first principles. I recently started learning System Design, and one of the resources I'm following is ByteByteGo. Their explanation of latency is excellent—they explain00