HNHiro Nakamurainhironakamura-ai.hashnode.dev·Aug 26 · 7 min readMetaRoCE: Why a Million-GPU Cluster Must Tolerate LossI keep a small rule when reading distributed-systems headlines: if the design needs every packet to behave politely, the design has already picked the wrong scale. That rule came back to me when I rea00
HNHiro Nakamurainhironakamura-ai.hashnode.dev·Aug 19 · 7 min readAgentic Git Hosting Needs a Review Queue, Not Faster Pull RequestsTechCrunch reported on August 18 that Cursor is launching a code-hosting platform to rival GitHub. Cursor calls Origin a forge for the agentic era. The practical consequence is uncomfortable for anyon00
HNHiro Nakamurainhironakamura-ai.hashnode.dev·Jul 22 · 7 min readDiffusion LLM Caching Is a Scheduler Problem, Not a Cache ProblemI was looking at a denoising trace last week and kept circling the same awkward question: why was the runtime redoing work for tokens we had already declared stable? The profiler did not have a satisf00
HNHiro Nakamurainhironakamura-ai.hashnode.dev·Jul 15 · 7 min readA 3.9 GB Model Does Not Make a Phone Agent: Runtime Memory Does3.9 GB is the number that matters in PrismML's July 14 Bonsai 27B announcement. It is also the number most likely to send a mobile team in the wrong direction. Treating a model file size as deployment00
HNHiro Nakamurainhironakamura-ai.hashnode.dev·Jun 17 · 7 min readOpenRouter Fusion: the synthesis step does the work, not the panelThe headline from OpenRouter this week is the kind that makes you stop scrolling. A panel of three cheaper models, fused together, landed within one point of Claude Fable 5 on a hard research benchmar00