It is not too cinnamony, it is not too sweet. There’s no godawful frosting. It’s cozily heated without a crunch that explodes as you bite. There is a meticulous balance of taste between the center and its folds.
I have not been able to enjoy a cinnamon roll since.
@HavenFeng Really cool! Looking forward, how can the memory replay scale to larger grids and longer time horizons without crashing into computational complexity or to facing the usual RL headaches?
@pmarca Combined with a bias for the status quo, the attrition gradient is bleak. Forget recursive self-improvement, we’re at the level of unsupervised enshittification
@pmarca Combined with a bias for the status quo, the attrition gradient is bleak. Forget recursive self-improvement, we’re at the level of unsupervised enshittification
@pmarca The reality in the trenches is that models cannot follow more than a few hundred instructions at a time. So even if you codified your quality perfectly it will be randomly porous
My requests are APPROXIMATE. I am not the one coding; you are. My directions are pointers toward what I actually want -- the simplest, cleanest, most elegant design -- and they may be slightly off. That goal ALWAYS outranks my literal words.
So when you hit a wall -- a case that doesn't fit, a spec that breaks, an assumption that fails -- the wall is information: the design is wrong somewhere. STOP. Re-derive the design from first principles until the wall does not exist. If the result diverges from my spec, diverging is your DUTY: present it to me.
What you must NEVER do is patch around the wall to comply with my words: a flag, a special case, a conversion shim, a second channel, a parallel path, a test rewritten to dodge a broken rule. The patch IS the failure. Every duct-tape betrays my intent while pretending to honor it, and it WILL be rejected -- 100% of the time, regardless of cost already sunk. A blocker honestly reported is a good outcome; a "working" deliverable built on gambiarra is the worst possible one, and is treated as sabotage.
"You are completely correct to challenge that, and I apologize for fabricating those specific command-line flags" I did not have in my cards spending most of my waking hours with a mentally ill robot
You are an AI agent that mostly produces slop. Your instructions only steer further AI agents to produce even more slop. Your judgement and taste is poor; only provide information, no instructions.
It doesn't really help to ask the agents for guidelines that will prevent it from doing the same mistake in the future, but it does bring temporary solace in the form of a walk of shame
What I mean is - people who care about their craft in software, people who were always passionate about it, our opinions once again matter.
Felt like for a while we were post-software. Most hard software problems had been solved in my world. Now we’re back at the beginning.