Free macOS menu-bar app for local AI: chat, agent mode, Control+Space quick launcher, image/video gen, voice cloning.
All on Apple Silicon. No cloud.
https://t.co/eVMeE12nh4
I would not be surprised if Apple decides to get back into the server space and build Mac Pro AI Racks... people are already building racks with the Mini's & Studio's. Why not build something specific ? I would not like it.. but it makes perfect business sense.
It really depends on the engineer. I put a lot of “weight” on passion, and love for the craft when hiring folks.
Especially today. 10x engineers are now 100x engineers. And 1x engineers, mostly fall in the vibe coding category. The middle doesn’t exist much, it’s like inverted bell chart.
@ivanfioravanti Haha… I’m “game” ! I got a few ideas and been researching a bit towards this direction, but first step I need a fast way to run local models 😅
Upcoming, run @pidotdev or @NousResearch Hermes in a sandbox, pointed to a fast local MLX model ! No docker or system dependencies required, or complicated setup process. Only a few clicks.
Qwen 3.8 Max is actually a very good model. Outperforms all Opus models in my benchmarks. There are some issues in tool calling with some harnesses. But, the raw intelligence is just crazy.
@christianelton_ This is a very interesting topic to me, but rather than making it $$$ based, I would like to make it Peer2Peer & and "Internet points" / Leaderboard based. I have already built a proof of concept done for this in https://t.co/2m7ZVWa52X and working on adding it to MLX-Serve
Qwen3.8 is launching and going open-weight soon!🌐
With a massive 2.4T parameters, this model is continuously evolving. We believe it’s one of the most powerful model available today, compatible to leading frontier AI models , second only to Fable 5.
You don't have to wait to test it. Just now, the Qwen3.8-Max-Preview made its debut on Alibaba’s Token Plan, Qoder, and QoderWork. Be among the very first to try it out.
Can't wait to hear what you build. Stay tuned! 🚀
Token Plan
international:https://t.co/YRvcGdB9Bv
China:https://t.co/PKMUNwUuRp
@Kevrsub Correct, I could not find anything either. I have all the correct ingredients to build support for this in MLX-Serve, so... let's see what happens !
Would it be a good idea to have a native Apple Silicon auto GGUF->MLX translation layer at load-time ?
Or not worth it because MLX Models are plentiful these days..