Speculative decoding is essentially using dumber model to predict what smarter model would generate. If in your workflow you get huge decoding boost from speculative decoding, your smart model is basically overqualified and you could use a smaller, faster and dumber one.
@MiaAI_lab Used that as a work horse on Halo Strix, then switched to 3.8 27B, never returned back. The newer is so much more reliable and requires much less oversight. Yes it's slow, but for local tasks that's not an issue if you can let it run unattended for hours and it will finish tasks
We promised open weights for Qwen3.8. Now, time to meet them! 🎉
⚡ Qwen3.8-27B:
- A native multimodal dense model. With just 27B parameters, it outperforms Qwen3.7-Plus overall and shines in real-world coding & office workflows.
- 262K native context, easily extendable to 1M tokens via YaRN.
- Built for builders. Highly efficient, high-quality, and licensed under Apache 2.0.
🚀 The open weights for Qwen3.8-2.4T-A95B (Max-level) have also been released recently.
Whether you're shipping lightweight applications with Qwen3.8-27B locally or building agents with Qwen3.8-2.4T-A95B, they're yours now!
Download, deploy, and build something we haven't imagined yet. 👀👇
- Hugging Face:
https://t.co/4kaAcqYEVj
- ModelScope:
https://t.co/eRIMZCGkhC
@max_katz@yulia_navalnaya Можно уточнить, мы сейчас в какой точке кацеворота находимся:
— ФБК растратил весь политический капитал и никому не интересен или
— Кремль так боится ФБК, что одно видео Юлии Навальной заставляет АП снять Яблоко с выборов и полностью поменять их концепцию?
CONFIRMED: Qwen3.8-2.4T and Qwen3.8-27B will be released with open weights in five days, at 10:00 AM on August 12 (UTC+8)! 🥳🎉
For the first time, Qwen will open the weights of a Max-class model: Qwen3.8-2.4T-A95B. Other upcoming models in the Qwen3.8 series include Qwen3.8-27B, which offers flagship-level intelligence.
More models from the Qwen3.8 series will also be released later on separate pages! 😻
Юлия Леонидовна, машина сильнее, мощнее, она сделана из стали и алюминия, и передвигается гораздо быстрее, чем Ульяна. Во многих аспектах более технически совершенна.
Очевидно, здесь вина Ульяны, что она не сопоставила ресурсы и опираясь на эфемерные правила дорожного движения, которые якобы должны её защищать, допустила это столкновение.
@i935100@adagamov То есть есть возможность увеличить количество ежедневных ракет на Киев в 100 раз, но Путин этим не пользуется из-за логики эскалации? Верится с трудом
@i935100@adagamov Так оно и так каждый день летит, в чем разница? Логика эскалации предусматривает неиспользование какого то вооружения, чтобы этим можно было угрожать.