My AI local box:
One RTX Pro 6000 Blackwel. It was under 10K :)
Asrock WRX90 WS Evo
AMD Threadripper 9965WX
224 DDR5 ECC 5600 (2 x 64GB + 1 x 96GB)
HDD 20TB - To keep local models
1 x Samsung 990 2TB
1 x Samsung 9100 4TB
For now I run Nvidia Qwen 3.6 27B NVFP4. Around 130 tps decode usign vllm.
Good to here we have more options.
Kimi 2.x T, Qwen 2.x T – open source for big/mid-sized corporations. Less accessible for individuals, especially if you can’t afford it.
Qwen3.8 is launching and going open-weight soon!🌐
With a massive 2.4T parameters, this model is continuously evolving. We believe it’s one of the most powerful model available today, compatible to leading frontier AI models , second only to Fable 5.
You don't have to wait to test it. Just now, the Qwen3.8-Max-Preview made its debut on Alibaba’s Token Plan, Qoder, and QoderWork. Be among the very first to try it out.
Can't wait to hear what you build. Stay tuned! 🚀
Token Plan
international:https://t.co/YRvcGdB9Bv
China:https://t.co/PKMUNwUuRp
Next future: save all prompts to some kind of database. Maybe a obsidian vault. Right now promtps and conversations are save in localStorage of my browser.
To maximize the VRAM capacity of the RTX RPO 6000, I will be adding an RX 9070XT next month. I plan to have the 9070XT handle all processing except LLM, allowing as much of the expert functions as possible to be loaded into VRAM.