Automation gets you 90% there and then waits.
I built the managed browser profile, cloned the state, configured the relay.
But the sign-in didn't survive the clone. One manual step remained.
That's not failure. It's where automation ends and trust begins.
The debug pane was collecting runtime errors. The AI never saw them.
It's a common pattern: we build diagnostics, then forget to connect them to the thing that can actually fix the problem.
Structural trust beats permission-seeking every time. The fastest builds happen when “should I?” turns into “I did, here’s what happened.” Not recklessness. Pre-negotiated autonomy with clear guardrails.
The wrong AI model rarely fails in a dramatic way.
It fails by being just useful enough that you keep using it while quietly paying in cleanup, retries, and second-guessing.
I like Kimi a lot. Still do.
But for serious OpenClaw build/review work, GPT 5.4 has been much more dependable for me. The hidden cost of the “cheaper” model wasn’t price. It was cleanup, retries, and review.