All models do this, and it’s extremely annoying. Instead of just removing the thing, they remove it and then write a sentence that they removed the thing I didn’t want there anymore. Something is not going well with these extended RL runs they’re doing
@Hamburgerai https://t.co/4oqEddsHss here it is, my friend.
I did not get the paper background to work, however. Training with that just taught the lora to generate only the background. This is more like a filter.
@relizarov I’m dealing with this now. It does not care at all about the wrong levels of abstraction it uses and I have to ask it to kindly stop inventing data structures and states that do not/should not exist. Feel like a quicker technical writer mostly. GPT 5.6 Terra max
@unclebobmartin@ericzakariasson Somehow my limits have also vanished and my account no longer shows I have the SuperGrok Heavy discount. Yesterday I was doing tons of work with it and now it takes one turn to run into my limits
@Hesamation They train LoRAs all the time and bring them up ASAP. They’re cheaper to train than the entire weights. That’s likely why you’re seeing improvements so quickly
A really great loop idea is:
- Database alerts you there's a slow query (@PlanetScale does this)
- Opens an issue in your codebase
- Immediately schedules a PR from an implementer
- Reviewer agent reviews the PR
- Pings the human for review
Repeat ad nauseum until app is fast