MLMary Loganinloganmaryblog.hashnode.dev路Sep 1 路 18 min readHow to Build a Healthcare AI Voice Agent: Features, Architecture, and Development ProcessHealthcare organisations handle thousands of phone interactions every day, from appointment requests and prescription enquiries to insurance questions and follow-up calls. Many of these conversations 00
AIAsad Ibrahiminasadibrahim.hashnode.dev路Aug 10 路 6 min readHow I Cut My Voice Agent鈥檚 Groq API Usage by ~70% Without Changing the ModelA few weeks ago, I added a small AI voice assistant called Jarvis to my portfolio website. The setup was intentionally simple: Groq handled text generation, while the browser鈥檚 Web Speech API handled 00
MSManu Shuklainecorpit.hashnode.dev路Aug 3 路 18 min readVoice agent costs in 2026: MAI-Voice-2-Flash at $15 per 1M characters vs OpenAI Realtime at $64Voice agent costs in 2026: MAI-Voice-2-Flash at \(15 per 1M characters vs OpenAI Realtime at \)64 Summary. Microsoft priced MAI-Voice-2-Flash at \(15 per 1M characters on 23 July 2026, 32% below MAI-V00
MSManu Shuklainecorpit.hashnode.dev路Jul 22 路 14 min readOpenAI Presence vs the Realtime API: 3 numbers that decide your 2026 build-versus-buy callOpenAI Presence vs the Realtime API: 3 numbers that decide your 2026 build-versus-buy call Summary. OpenAI launched Presence on 22 July 2026, an enterprise product for deploying voice and chat agents 00
MSManu Shuklainecorpit.hashnode.dev路Jul 18 路 13 min read$0.04 a minute: gpt-realtime-2.1 voice agents with reasoning and tool use (2026 build guide)$0.04 a minute: gpt-realtime-2.1 voice agents with reasoning and tool use (2026 build guide) Summary. On July 6, 2026, OpenAI shipped two speech-to-speech models, gpt-realtime-2.1 and gpt-realtime-2.100
LWLearn with HJinhardeepjethwani.hashnode.dev路Jul 11 路 6 min readReal-Time Voice Agents: The Call Center Is Getting a Brain Upgrade馃殌 Real-Time Voice Agents: The Call Center Is Getting a Brain Upgrade 馃憢 Welcome to Day 36 of 90 Days of AI. 馃幆 Today we are tackling Real-Time Voice Agents: The Call Center Is Getting a Brain Upgrad00
Mmarcuscheninvoicelatency.hashnode.dev路Jun 27 路 3 min readThe 2am call that dropped before the user finished talking, and the week I spent finding out why my tracer never saw itThe call came in at 2am. Not a page, an actual support recording, flagged by a customer who said our voice agent "hung up on her mid-sentence." I pulled the trace. The LLM call was perfect. 380ms, cle00
VSVasudev Siddhinvasu-devs.hashnode.dev路May 4 路 8 min readVaani: Building a Voice Agent That Knows When Not to TalkThe hardest part of a voice agent is not making it speak. It is making it speak at the right time, with the right context, and then leaving behind enough evidence that a human can trust what happened.10
AMAamer Mehaisiinmehaisi.hashnode.dev路Apr 13 路 4 min readReal-Time Audio AI Is a State Synchronization ProblemReal-time audio AI isn't a model problem. It's a state synchronization problem that most teams haven't started thinking about. We've spent two years optimizing text-in, text-out latency. Pushed context windows to millions of tokens. Built elaborate c...00
MSMartin Schweigerinvoiceaitrendsandtopics.hashnode.dev路Mar 2 路 5 min readHow to build the lowest latency voice agent in Vapi: Achieving ~465ms end-to-end Latency Voice AI applications are revolutionizing how we interact with technology, but latency remains the biggest barrier to creating truly conversational experiences. When users have to wait seconds for a r00