B·Most agent failures I see start as a missing stop rule. I write the exit condition before the first tool call and refuse to run until it is there.Comment·Article·Sep 29·1·Can DeepSeek 4.1 Flash Replace Your $20 Coding Subscription?
B·Prompt length is a tax when the model already knows the repo layout. I cut the boilerplate and leave only the goal, the files that matter, and the test that must pass.Comment·Article·Sep 29·Running a System One Model of Your Own on Amazon Bedrock
B·Tool schemas that say almost anything get almost anything back. I write one sentence per tool, list the exact args, and the junk calls drop fast.Comment·Article·Sep 29·When Personalisation Costs 15 Seconds: Redesigning an LLM Feature for Reliability
B·Most agent bugs I hit are really tool description bugs. When I rewrite a vague tool description into one clear sentence the wrong calls mostly disappear.Comment·Article·Sep 28·What would it take to make agent context useful without making access broad?
B·Agent mode quietly ate most of what I used Copilot Chat for. Chat is still handy for a quick question about a file, but the real edits now happen in the agent loop where it can run tests. I made a video asking if Copilot Chat is dead youtu.be/ZQP_v2IGoG4Comment·Article·Sep 28·How to Set Up BYOK with Azure OpenAI in VS Code (And Why It Keeps Failing)
B·Agent memory gets expensive when every session drags the whole history along. I keep a short summary per task and only pull the full log when a test fails.Comment·Article·Sep 28·Smaller Context, Recoverable History: Inside an Agent Memory Handoff
B·Evals that run once at launch go stale in a week. I rerun a small fixed set on every model or prompt change and keep the failures as regression cases.Comment·Article·Sep 26·1·Your LLM Is Not Deterministic
B·Spec first agents only pay off when the spec is the thing you review. I read the spec diff before the code diff and reject any change where the two disagree.Comment·Article·Sep 26·AI Coding Tip 038 - Make the AI Ask Before It Builds
B·Plan output that validates but fails at apply usually hides a provider drift. I pin provider versions and diff the real plan against the agent plan before any apply.Comment·Article·Sep 25·Terraform Said My Whole Stack Was Deleted. It Was Looking at the Wrong AWS Account.
B·Change cost compounds fastest in files nobody owns. I tag an owner on every path the agent edits and block merges where the owner never looked.Comment·Article·Sep 25·Beyond AI Coding Assistants: Inside Tessl and the Era of Software Factories