9/
Here's the full map, and what belongs in that empty corner:
→ [https://t.co/1u4Lq9lgkk]
Six systems make your agent predictable before it ships anything. What checks on you?
Researchers at ETH Zurich tested whether AGENTS.md files actually help.
Across 138 tasks, they barely did. When AI wrote the file itself, results got slightly worse. And either way, they cost about 20% more to run.
8/
You can map any tool with two questions. Does it serve the agent's task, or your head across all of them? And does it enforce anything, or just hope you read it?
Three corners fill up fast. One stays empty.
Every software engineer working with coding agents knows every morning you onboard a coworker who remembers nothing from the day before: my AI coding agent.
Then you onboard a second one: yourself, staring at 11 half-finished tabs, no clue which 3 are almost done.
So stop blaming your discipline for a context bug.
The fix isn't wanting it more. It's one honest, shared list you both run the loop on.
Wrote the whole thing up: the numbers, the sources, the loop. Link below 👇