Lets be real for a moment
this hobby is addicting, and half of you want another GPU because of a release you saw an hour ago
Me too
But this place is an echo chamber and the hype here is NOT a buying signal
Remember this
My new low-bit quant of Qwen3.8-Flash-Next.
2.125-bit experts
2.39 bits/weight across the transformer
92.0 GB download
~39 GiB in memory
95.5% of BF16’s score
A 180B-parameter model that runs on one DGX Spark.
The hugest of huge thanks to @LambdaAPI and the amazing @TheZachMueller for making a project this compute-hungry possible.
🔗 https://t.co/Gayvg7xECi
My new low-bit quant of Qwen3.8-Flash-Next.
2.125-bit experts
2.39 bits/weight across the transformer
92.0 GB download
~39 GiB in memory
95.5% of BF16’s score
A 180B-parameter model that runs on one DGX Spark.
The hugest of huge thanks to @LambdaAPI and the amazing @TheZachMueller for making a project this compute-hungry possible.
🔗 https://t.co/Gayvg7xECi