This post is incredibly practical! I’ve been struggling with blind AI cost & latency issues for our voice agent platform, and your breakdown of Gemini hidden thinking tokens + self-host SigNoz pitfalls hit exactly what I need to fix our cost calculation model. Super valuable real-world production experience, thanks a lot for sharing!