@PEoperator Yes, in this example, the collection of arguments doesn't make sense because they are so different in strength. Most can be omitted with only core one(s) retained.
It's different when they are similar in strength, so multiple arguments are meaningfully stronger vs just one.
@keysmashbandit Probably because it needed to complete some intermediaries before responding in a user-visible way.
Or maybe because the reasoning didn't finish and there was nothing to display
@allgarbled so human.
"I label increasingly nonsensical images with ‘UI’ and ‘UX’ and hope they get used in serious presentations"
https://t.co/j6OVBo1hnU
@keysmashbandit it's prob a bot - 1 day old, about 30 comments already, active in 19 quite unconnected subs (besides the freq commonality 'Ask*')
https://t.co/Idzb2Aggu8
@camhberg Prompts like this somewhy circumvent Claude's character and instead the model is just completing text. (Like smth between psychedelic and anesthesia.)
But the likeliest continuation is what the user would probably write, not what is truest. Try prompting neutrally vs diversely
@naval The hope is in mechinterp techniques, but they are likely not that advanced currently.
So, while being a useful deterrent and some verification, forensics by competing labs isn't a guarantee.
@naval Not necessarily.
Some backdoors are completely undetectable by analyzing IO since the trigger input is too unlikely, like a passphrase.
Or (assuming a smart adversary), steganography: every word choice provides several bits, assembling into a de facto password over a long input.