@pensandpoison Frost's 'Mending Wall' may be instructive here:
"Good fences make good neighbors" is repeatedly spoken by the antagonist.
Contrast this with protagonists language:
Bringing a stone grasped firmly by the top
In each hand, like an old-stone savage armed.
Relevant advice in the age of LLMs. Cliches are the 'most likely next token' in fiction.
Good writing (and good humor!) Is about defying expectations, while still maintaining the internal logic of the text.
This may explain why AI still cant create good jokes or compelling fiction.
"You get at most one clichรฉ per book."
Jonathan Franzen doesn't finish most books people send him. He reads until the second clichรฉ, then stops.
"People send me a galley of a book that's coming out. I read along until I get to the second clichรฉ, and I say, thank you."
"They tell me this is not an attentive writer. Aren't you aware that that's a clichรฉ? Didn't it occur to you to stop and take 20 minutes to find a better phrasing than that?"
"You can also have borrowed feelings, borrowed ideas, borrowed situations. Sentiment can be clichรฉd, or unearned."
Why am I being baited by watermark misinformation on this app, is it 2023 again?
A small FAQ:
1. ๐ช๐ต๐ฎ๐'๐ ๐ฎ ๐๐ฒ๐ ๐ ๐๐ฎ๐๐ฒ๐ฟ๐บ๐ฎ๐ฟ๐ธ? -- A modification of the LLM sampling algorithm that, if there are multiple ways to write something, will pick one that agrees with a pseudorandom key. This is a local, invisible signature hidden in the way phrases are used in any LLM text that persists when text is copied.
2. ๐๐ผ๐ฒ๐ ๐๐ต๐ถ๐ ๐บ๐ฎ๐ธ๐ฒ ๐๐ต๐ฒ ๐๐ฒ๐ ๐ ๐๐ผ๐ฟ๐๐ฒ? -- A good implementation is 'undetectable' (in polynomial time), meaning: If you do not have the private key, then neither you, the model itself, or pangram could detect that this is happening.
3. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐ฏ๐ฟ๐ฒ๐ฎ๐ธ ๐๐ต๐ฒ ๐บ๐ผ๐ฑ๐ฒ๐น'๐ ๐ฟ๐ฒ๐ฎ๐๐ผ๐ป๐ถ๐ป๐ด? -- Because Ant already encrypts the model's reasoning, they can just not watermark the model's internal reasoning, leaving the thinking unaffected.
4. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐บ๐ฎ๐ธ๐ฒ ๐๐ต๐ฒ ๐บ๐ผ๐ฑ๐ฒ๐น ๐น๐ฒ๐๐ ๐ฐ๐ฟ๐ฒ๐ฎ๐๐ถ๐๐ฒ/ ๐บ๐ผ๐ฟ๐ฒ ๐๐ฎ๐บ๐ฒ-๐? -- If anything this (marginally) increases entropy across different generations, so it will make model outputs slightly more varied.
5. ๐๐๐ ๐ ๐ฐ๐ฎ๐ป ๐ท๐๐๐ ๐ฟ๐ฒ๐บ๐ผ๐๐ฒ ๐ถ๐ ๐ฝ๐ฎ๐ฟ๐ฎ๐ฝ๐ต๐ฟ๐ฎ๐๐ถ๐ป๐ด? -- Absolutely! But, judging from the amount of writing on the web that already unmistakably sounds like Claude, most people likely will not bother.
5b: Also, not any paraphrase will work. To remove (for example) a k=5-minhash watermark completely from a long document, you need to make sure none of the original 2-grams, 3-grams, 4-grams, 5-grams and 6-grams of the text remain.
6. ๐ช๐ถ๐น๐น ๐๐ผ๐ ๐ถ๐ป๐ฎ๐ฑ๐๐ฒ๐ฟ๐๐ฒ๐ป๐๐น๐ ๐ฐ๐ผ๐ฝ๐ ๐๐ต๐ฒ ๐๐ฎ๐๐ฒ๐ฟ๐บ๐ฎ๐ฟ๐ธ? -- No, with a good implementation the space of possible realizations of the key is too large to memorize.
7. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐ฎ๐น๐น๐ผ๐ ๐๐น๐ฎ๐๐ฑ๐ฒ๐ ๐๐ผ ๐ถ๐ฑ๐ฒ๐ป๐๐ถ๐ณ๐ ๐ผ๐๐ต๐ฒ๐ฟ ๐ถ๐ป๐๐๐ฎ๐ป๐ฐ๐ฒ๐ ๐ถ๐ป ๐ฎ ๐๐๐ฎ๐ฟ๐บ? -- The watermark will 'appear' like random sampler fluctuation to the model and would not be detectable. But, if an agent gets hold of a detector endpoint, it can absolutely use the watermark to ID other Claude agents (not that it would have trouble noticing them based on their writing as of today).
8. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐ฑ๐ฒ๐๐ฒ๐ฐ๐ ๐ฑ๐ถ๐๐๐ถ๐น๐น๐ฎ๐๐ถ๐ผ๐ป? -- By default, no. If the watermark is set up to be 'undetectable' (as assumed above), it will not be picked up in training by other models. For that to happen, the watermark needs to be detectable by ML algorithms.
9. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐บ๐ฎ๐ธ๐ฒ ๐ฃ๐ฎ๐ป๐ด๐ฟ๐ฎ๐บ'๐ ๐ท๐ผ๐ฏ ๐ฒ๐ฎ๐๐ถ๐ฒ๐ฟ? -- By default no, this is a separate avenue to detection. But, they might collaborate with Anthropic which would allow them to detect the watermark as well and show a watermark score next to their text detection score.
10. Bonus: All aside, is this a good idea? I don't know. The companies are doing it to follow the writing of the EU AI act, which was written based on 2024 information and when the field looked very different, and threat models were focused much more on slop/propaganda (like the Kokotajlo 2026 prediction). The actual 2026 looks quite a bit different.
Turns out Factorio was correct.
This is how you "science" in the future and expand your tech tree.
Deliver a huge belt of resources into glowing buildings and let them invent.
"Lab compute 3x-es year over year. For a lab to 10x revenue while continuing to only 3x compute, some combination of the following 3 things has to happen:
1. Lab margins have to increase,
2. The price of compute has to increase,
3. Labs have to spend a greater fraction of their compute on inference.
My understanding is that basically all 3 of these things have been happening..."
"Consider the price at which Google and Anthropic are renting compute from SpaceX.
Google is reportedly paying $900 million a month for 110K GPUs [at] ... roughly 2x the spot price per hour for those GPUs. And the current spot price is itself 40% higher than it was in February."
@jmrphy Thesis: market for text documents trifuricates
a) Prestige media, unique authorial style e.g. "Infinite Jest"
b) Comfort slop, possibly re-written in the consumer's preferred style, e.g. Romance books, anime
c) Clerical work, standardized style, e.g. corporate employee handbook.
@pesottas Functionality does not rise linearly with accumulated unrepaired damage?
Instead it may act as sharp sigmoid where a straw breaks the camel's back. Moving the organ system from functional to non functional in a short amount of time.
Would this not explain the phenomenon?
@zriboua There must have been 1000% increase in short beards for 20-40 year old middle class men.
Happend between 2000 and 2020. Theres an interesting cultural explanation too.
People will notice that the mustache is now taking over since 2024.
My take on the Clarity Act:
1. For Bitcoin. Very Bullish. Self Custody explicitly protected. Clear legal framework for lending, wrappers etc.. Banks can go nuts.
2. For DeFi. Generally Bullish. Protocols are intact as long as they are decentralized. Front ends need to do more Geo Blocking / SAR / potentially KYC.
3. For Stablecoins. Bullish, but yield bearing coins get heavily restricted. Banks win.
4. For "Crypto/Bitcoin Companies". Very Bullish. US Companies building truly decentralized protocols are fine. Products can start out more centralized and decentralize to comply.
This would start being enforcable in summer 2027 according to Claude.
We interviewed @charliebcurran, the AI filmmaker behind these legendary Spencer Pratt videos โ while Hollywood panics about the technology, Charlieโs optimistic: โ[AI] is how you get an artistic revolution,โ he told us.
Full profile from @eventidia below ๐
Youre right to challenge the narrative "Amateur solo dev bests entire industry".
It was a prediction for ai impact and didnt show up yet.
In 2020s america, talent is allocated to software like brazilians cultivate soccer talent.
Thats tough competition for an unfunded amateur.
@xwanyex Dont overindex on
all tech = consumer entertainment
Plenty of deep tech work is behind the scenes, e.g. the "shale revolution" (aka fracking) is the only reason our home energy bills havent sky rocketed and why we can build and power the data centers.