@Pinperepette A proposito di roba vecchia: ho iniziato a guardare Rick and Morty, che esiste da oltre 10 anni. Non mi capacito di come non gli abbia mai dato un’occhiata prima. Se ti trovi nella fortunata condizione di non averlo ancora visto, te lo consiglio.
We’re making better intelligence easier to access in ChatGPT for everyone:
- GPT-5.6 Sol now powers both Instant and deep reasoning for Plus & Pro users, delivering more factual, focused responses.
- Free & Go users get unlimited text chats with GPT-5.6 Luna starting tomorrow.
La nuova versione di DeepSeek 4 Flash mantiene la stessa architettura e le stesse dimensioni, ma offre prestazioni superiori. Se i test sul campo confermeranno i benchmark, progetti come DwarfStar di @antirez renderanno sempre più giustificato l'investimento in hardware.
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!
🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇
🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex!
Check out the configuration details in our official API docs: https://t.co/smCwQZMeiq
Claude Opus 5 from @AnthropicAI is the new SOTA on ARC-AGI-3: 30.2%
The previous high score (7.8%) was set by GPT-5.6 Sol (Max)
Throughout our analysis, we observed novel behavior that allows Opus 5 to solve previously unbeaten environments, outperforming Fable
Chissà se tra dieci anni, aprendo X mentre sorseggiamo un caffè, leggeremo: "GPT-14.5 ha dimostrato l' ipotesi di Riemann". O ancora più scioccante: "GPT-14.5 ha trovato un controesempio all'ipotesi di Riemann". A quel punto chissà quanto l'AI avrà già trasformato le nostre vite.
Today we're releasing Laguna S 2.1, our most capable model to date.
It's a 118B total parameter Mixture-of-Experts model with 8B activated per token, a context window of up to 1M tokens, and thinking and no-thinking modes.
Capable enough to hold its own against models many times its size. Small enough to run on a single @NVIDIAAI DGX Spark.
Laguna S 2.1 is fully open under OpenMDW-1.1, with weights available today on @huggingface
https://t.co/xxGeAgo35R
Fortunatamente devo usarlo solo un paio di volte l'anno.
Che poi @antirez è da un po' che prova a dire una cosa molto semplice: il software scadente esiste da ben prima dell'avvento dell'AI.
Stamattina, mentre cercavo di caricare una serie di documenti su un sito istituzionale lentissimo e pieno di bug, pensavo che persino l'AI più scema farebbe di meglio. Ma mooooolto meglio.
Sono abbastanza fiducioso che nei prossimi mesi raggiungeremo una densità d’intelligenza per milione di parametri sufficiente a far girare sul nostro PC, anche con hardware “normale”, modelli capaci di fare molte cose utili.
Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone.
Bonsai 27B is the new multimodal flagship of the Bonsai family. Based on Qwen3.6 27B, it brings a new capability tier to local AI: multi-step reasoning, structured tool use, long-context workflows, and coherent agentic loops.
Until now, models in this class have been impractical to deploy locally. A 27B model occupies roughly 54 GB in 16-bit precision, and even a strong 4-bit build is around 18GB - too large for a phone and for most laptops.
Bonsai 27B changes that.
It comes in two variants:
• Ternary Bonsai 27B: 5.9 GB, 1.71 effective bits per weight, optimized for laptop-class quality.
• 1-bit Bonsai 27B: 3.9 GB, 1.125 effective bits per weight, optimized for phone-class footprint.
Everything is open-sourced today under the Apache 2.0 license.
Nelle prossime settimane vedremo Kimi 3 (imminente?), DeepSeek 4 definitivo e GLM 5.5.
Sono davvero curioso di capire quanto saranno migliorati e come si posizioneranno rispetto ai modelli americani.
Esiste un modo per mantenere i modelli AI aperti, seguendo il percorso del software open source? Oppure il costo enorme della potenza di calcolo li rende dipendenti dalle scelte degli Stati che ospitano laboratori e infrastrutture?
Domanda decisiva per il futuro dell’AI aperta.