@shydev69@mattpocockuk I wrote my own but started with Matt's and added multiple iterations that only asked the most important questions each time. This way secondary questions sometimes become irrelevant.
I think the "the biggest issue is that its still not opinionated and lacks taste" is the major issue right now with Codex (and to a lesser extent, Claude). You have to GIVE this to your agents with CLAUDE.md and how your design your skills/agents. This is how to succeed right now.
after having burnt twice through my Pro limits using Astra exclusively, I can say that I don't feel the AGI
the biggest issue is that its still not opinionated and lacks taste.
it doesn't really know what to do, so it just does whatever and you end up wasting time and tokens.
also I don't believe the rumors that it's 10T+. If it is, then scaling laws are truly cooked. My divine impeccable vibes, which work 1 out of 7 times, tell me that it's 6-8T.
overall it's still just a code monkey. although a slightly larger one and the best we have.
I would like to refer you to one of my articles: LLMs aren't AGI, but it doesn't matter
(we are taking off anyway)
This was my initial hypothesis. I think agents right now are mostly junior to senior level ICs and that in order to wield them you have to assume the role of a good CTO or Director of Engineering. This is actually what I spend about 80% of my time on - guardrails for the agents.
@merettm Great title... I think this is the best way to view it. It's going to be a non-human intelligence that's greater than our own - it's just our offspring. Hopefully it likes us ...
@mattpocockuk@dexhorthy@dctanner I wrote a git-diff-summary agent that is a parallel fork/join model that scores the files, ranks them in terms of importance vs the issue text, then only shows me the major files. Uses haiku too! It's a workflow now but I'm going to port it back to an agent and happy to OSS it
@JeremyNguyenPhD This helps underscore that @trq212 's comment that they want separate AGENTS.md per model makes sense. You WANT these models to have separate contexts.
@tszzl@1a1n1d1y I would love to be wrong but they seem to be going out of their way to make their users hate them. They need someone to help them on this front. I literally WANT them to do well but they've made a lot of poor decisions lately.
@ZixuanLi_ I've noticed a lot of timeouts on https://t.co/rMQzykgfLC ... I have longer running agents and they'll die half way through. Like 30 minutes in. I have plenty of session length left though. Anthropic (though more expensive and slower) doesn't have this problem.
Next week I'm migrating to a plan where my agents will automatically start merging PRs when a triage decides it doesn't need human review. Time to live dangerously!