Agent Network Access Control: Egress Is the Real Risk
An agent with a shell and open network egress can exfiltrate anything it can read. How to scope outbound access without breaking the toolchain.
ReadPractical writing for developers building with large language models — how they work, how to pick one, and how to keep the bill predictable.
An agent with a shell and open network egress can exfiltrate anything it can read. How to scope outbound access without breaking the toolchain.
ReadAverages hide every interesting agent failure. Which counters and histograms to define, what to use as the denominator, and how to alert without drowning in noise.
ReadPer-call prompts, session grants, capability scoping and policy engines. How the main agent permission models behave once real work runs through them.
ReadA prompt edit is a production change with no compiler and no stack trace. How to give prompts a stable identity, trace them, roll them out and roll them back.
ReadAn unbounded queue in front of an agent fleet hides overload until every task is stale. Queue design, backpressure signals and shedding work on purpose.
ReadA prompt edit can silently break a workflow that worked yesterday. Here is how to build a regression suite for an agent that actually blocks merges.
ReadPicking up a stopped agent run safely: why it stopped changes how you resume, and why the world may have moved while the agent was not looking.
ReadAsking a model to check its own work sometimes fixes real errors and sometimes invents new ones. What separates the two, and how to build for it.
ReadA shell tool is the most useful and most dangerous thing you can give an agent. How to grant it without handing over the whole machine.
ReadGiving an agent explicit states and legal transitions cuts wandering and makes failures debuggable. What it buys, what it costs, and when it is overkill.
ReadBreaking work into subtasks is the standard fix for agents that flail. Here is how to size the pieces so decomposition helps instead of adding overhead.
ReadAgents fail intermittently, so a single passing run proves nothing. How to build a harness that measures rates rather than checking assertions.
ReadShowing 13–24 of 404 articles