Aquest dilluns, 27 de juliol, es compleixen 750 anys de la mort de Jaume I, que és enterrat a Poblet. Aquests dies se succeeixen els homenatges i les activitats en record del rei, com el que va fer ahir la Muixeranga de la Marina Alta
New (shorter) lecture! Over-optimization, foundations of reward hacking, sycophancy, verbosity, etc. In recording this, I realized that rubrics are going to be prone to overopt in a way like reward models, where RLVR is its own thing. This is mostly fundamentals, history, and reflections!
00:00 Intro & Why We Care About Over-optimization
04:20 Part 1: Over-optimization & Goodhart's Law
09:57 Signatures of Over-Optimization & Misalignment
14:47 Part 2: Beyond "Just Style"
20:13 Llama 4 & Gaming the Leaderboards
23:04 Course Recap & Conclusion
Primarily on Chapter 14.
Estàs barrejant conceptes que no són contraris al nacionalisme.
La interculturalitat és convivència i interacció entre cultures. La diversitat és l’existència de diferències. El cosmopolitisme és una identitat oberta al món. L’internacionalisme és cooperació entre nacions. La pluralitat i l’heterogeneïtat només descriuen una societat diversa. I el globalisme, de fet, sovint tendeix a uniformitzar més que no pas a preservar diferències.
Cap d’aquests conceptes elimina la necessitat de comunitats polítiques, llengües, institucions i marcs nacionals. Al contrari: la diversitat real només es manté quan les cultures tenen prou poder, territori i institucions per reproduir-se.
Sense això, no tens diversitat: tens cultures petites dissolent-se dins la llengua, el mercat i l’Estat dominants. Defensar que el català pugui continuar existint no és anar contra la diversitat. És fer possible que la diversitat no sigui només decoració folklòrica.
@arqueoleg@NoelHuguet@jordioriolserra Una correlació de -0.32 és una correlació feble quasi nul·la, que explica només un 10% de la variabilitat. Jo no diria que amb això puguis predir amb alta precisió. Posant altres variables i mirar de fer lag-correlation? Segurament. On has agafat les dades/gràfic?
Teresa Comellas, amb 9,878, és la millor nota de la Prova d’Accés a la Universitat de les Balears.
El Govern concedeix enguany 20 premis de 600 euros als millors qualificats. Ara, tenen un estiu per davant i més a prop poder estudiar el que realment volen.
↘️ https://t.co/s77sT1pOWx
In nearly 5 years of modern generative ai, this is the first book I’m seeing with a super high level of coverage and comprehension.
> language modelling
> inference optimisation
> RL and its methods
> system scaling
> applied concepts like agentic ai, rag, memory
> environments and benchmarking
These fields have a subtle boundary differentiating them, but ultimately overlap in modern applications. Agents require system scaling, memory needs inference optimisation, rl requires understanding of environments and benchmarks.
For the first time in my exp, all in one place. Found this on paperswithcode[.]co
Claude Tag is a Trojan horse. Not because Anthropic is doing anything evil. Because the incentives are obvious.
Day one, this looks like a great feature: tag Claude in Slack, let it follow the thread, remember context, connect to tools, break down tasks, chase work, and act like a teammate.
But that is exactly the problem. The moment your AI vendor becomes a shared coworker, it stops being just a model provider. It starts becoming the place where work is interpreted, remembered, routed, and eventually executed.
That is not model lock-in. That is context lock-in. You are now renting your company back from them.
Models can be swapped. Agents can be copied. But the memory of how your company actually works is much harder, maybe impossible, to move: the Slack scar tissue, the exception paths, the customer promises, the unfinished threads, the weird workflows, the implicit owners, the “we tried that in Q2 and it failed” knowledge.
Once that lives inside one vendor’s agent layer, you are not renting intelligence anymore. You are renting your company’s operating memory.
And the pricing model makes it even more dangerous. A human coworker has a salary. Claude has unbounded tokenized activity. The more work moves through it, the more the vendor captures not just IT spend, but labor spend.
This is the enterprise bargain people will regret: Convenience now, and rapid decent into dependency.
The right architecture is simple: rent the best intelligence from whoever is best this month. OpenAI, Anthropic, Gemini, open source, whatever. But own the context layer.
Your company memory should be inspectable, permissioned, portable, and model-neutral. It should not be buried inside the same vendor that sells you the intelligence and the workflow surface.
Claude Tag is useful. That is why it is dangerous. Rent the intelligence, but own the context. Or, regret later.
I've just published a new blog post explaining some graph basics, Graph Neural Networks (GNNS), and Graph Attention Networks, including GAT and GATV2. Check it!
https://t.co/ZKNdruFjor
Ens entrevisten al pòdcast 🎙 de La Renaixença de @som3cat d'@enPeyu! El nostre company @jordimash parla de la història de l'associació, els seus orígens i la tecnologia de l'època, els serveis que oferim, el voluntariat... i de xiclets 🍭! https://t.co/HIkMPc0sP6
Aquesta entrevista no té un únic protagonista: és un recorregut a la tasca que ha dut a terme Softcatalà i a totes les persones, presents i passades, que l'han fet possible.
Molt content de poder representar la comunitat de Softcatalà i tenir l'oportunitat de guanyar un fuet 😀
Would you like to join the research effort on JEPA and World Models easily?
After a full year of hard work, we’re excited to finally release stable-worldmodel:
an open-source, scalable platform built to accelerate JEPA & World Model research!
📄: https://t.co/gnxGvens5A
Fita espectacular pel català! A part del text de la Viquipèdia, sabíeu que l'historial d'edicions de Viquipèdia té un valor també molt important per entrenar IAs? Per exemple, es fan corpus de correcció gramatical com WikEd Error a partir de l'historial d'edicions. Tot s'aprofita