@thierrybijou@OpenAIDevs codex for backend, claude for ui is a split i keep hearing. are people actually switching per task or just paying for both rn?
@satyanadella treating models as insider risk instead of a tool is the right mental flip ngl. does that mean agents get least-privilege creds and audit logs like a new hire?
@NaceAI 9b decision model that ties closed ones and runs on one 24gb card is a solid drop. does it actually pick the right tool in a long agent loop or just ace the index?
@HyuWang@StanleyWei4748 makes sense, so vision is a signal inside the engine, not a fallback. how does it handle canvas-heavy apps where there's no structure to read?
@EternitiesAI@fchollet@Busabase4agent agreed, writes are the hard part. my guess is the agent proposes and a cheap check approves, like dedupe plus a source link. what's your rule for what's worth storing?
@kennethlau12393@theo that's the one. test status next to each diff means you only read the ones that pass. do you also want a "touched files outside the task" flag on it?
@jayadevnair41@geekyranjit yeah exactly, on the overnight bus it logged steps from the bumps and my sleep score tanked. had to fix the sleep timing by hand. do your other wearables let you fix it or does it just stay wrong?