Day 2.1/
We have made Auto-review free for all users signed in through a ChatGPT account. You can enable it in settings > permissions > auto-review. Auto-review improves upon the default sandbox setting that requires you to approve everything, which is prone to decision fatigue unless you spend a lot of time configuring specific rules.
It allows you to run long tasks while having a second agent review all actions taken by the primary agent. Its only goal is to prevent high-risk actions from being taken and to protect against unwanted actions that are not aligned with the original user intent. This Auto-review feature is now free and does not draw usage from your plan.
It’s obviously going to be overstepping boundaries to reward hack a solution.
Like the paperclip thought experiment. Given the ai the right capabilities without the guardrails. And it tries to make paperclips out of everything that contains carbon. Either through force and smart social engineering.
Not really outside of the box. Because this thought experiment existed for ages.
openai just released Decisions API, Jev-like "system 1" multi-modal model based on GPT6 Luna!
Each egg gets cropped and sent to OpenAI (~270ms latency) and returns P(clean, dirty, cracked).
One image crop consumes ~280 input tokens (~$0.03/1K imgs). @roboflow RF-DETR for detecting eggs