I STOPPED LETTING CLAUDE MAKE DECISIONS THE DAY I BUILT THIS JEV AGENT FOLDER
I used to let Claude decide, write and act on every single step
-> now Claude only writes. Jev makes the calls, code does the acting, and every step leaves a receipt
here's what's inside the folder:
• the input
> AGENTS.md - when to call Jev, when to skip it
> state/build_state.py - goal, workers, done, missing, constraint. evidence, never a summary
• Jev decides (questions/)
> https://t.co/SIhqwoheZP - which model tier gets the task
> next_worker.py - which worker acts next
> https://t.co/y6q9a7fv9Y - keep or drop each tool output
> https://t.co/8LKHJE7ySk - is the goal actually met
> https://t.co/WtQhAqPFSG - does this send, pay or delete?
• code acts (rules/)
> hard_rules.py - stop after ten actions, never publish unapproved
> thresholds.yaml - act only on confident answers. fraud needs 0.95
• Claude writes (workers/)
> https://t.co/Pl96LqadxQ - sources and notes
> https://t.co/4GCndJWKsk - drafts and briefings
> the only place in the whole folder where text gets generated
• the proof
> receipts/decisions.jsonl - options offered, chosen id, re-check, fallback
> evals/ - dozens of my own labelled traces, thresholds tuned on them
• the guards (hooks/)
> pre_tool_use.sh - every command checked before it runs
> https://t.co/UC0IBplGsa - checks "all done" before the agent is allowed to stop
10,000 decisions for $0.42. median 300 ms
an LLM writes, Jev decides, code acts