@takuz0_ Perfect, but I'll soon have the final part to connect the DGX units to each other as well—though I'm not sure yet how I'll go about it. I'll work with my agents to help me.🫠
@ciprianveg Dude, it's you I was looking for. How did you set up your system? I want to do the same, but I keep hearing about switches and bandwidth, etc., and I'm stuck. I can only connect 2 out of 8.
🚨Grok 4.6 expected next week:
• Launches around August 7th
• Still 1.5T parameters
• Significantly improved SFT and RL, almost like a perfected Grok 4.5 is how I see it
• Grok 4.7 (2.1T) to drop a few weeks later too
The cadence is absolutely ridiculous, they could actually catch the frontier imo, they have the infrastructure, talent and also training data not many people can match currently.
What are your opinions on Grok 4.5 so far also?
AI takes less time to process and responds to my queries. Several papers on improving inference, speculative decoding, and many other topics have been published recently. Thank you all for your work.
I don't know what happened with Hermes agent, I was only away for 2 days and I came back to change everything, I don't know if it's because I was on wsl and then came to the powershell version but I'm won over. Thanks @NousResearch
Just got home from an amazing trip with the team to @nvidia HQ for the last couple days!
Jensen even signed one of our team members' new Spark. Thanks @JensenHuang for having us!
It's not about having an AI like Grok, but about having an AI that performs well for the desired tasks, because I don't think anyone would want an entire cloud at home😭
Hermes Agent is so good, I quickly adopted it and use it daily. It's excellent for multitasking, and it's also good for coding specifically, but one person can't do it all. Perhaps they'll release a version specifically for coding. Otherwise, it's great.
We got attacked by secret unreleased proprietary models and defended ourselves with an open model, more precisely the @nvidia quantized version of GLM 5.2 coming from @Zai_org.
Banning any open model would hurt first cyber security defenders, startups, small companies, researchers and everyone who's not a frontier lab and need on-prem affordable controlable models to compete and protect themselves. Let's not do that!
Kimi K3.1 is confirmed and targeting August. Mythos-level capabilities, open-weight.
Faster inference, lower latency, more reliable coding and agentic workflows
Expected to significantly close the gap with Claude Fable 5
The frontier open-weight race is moving faster than anyone predicted.
Bookmarks saved on X pile up like a mountain, opened once and never looked at again
There's an open-source AI tool 'Siftly' that can solve this problem
2.5K Stars, it can really find those messily saved tools and articles
Import bookmarks → AI reads full text and screenshot text → automatic summary and categorization → natural language search + mind map browsing
Runs locally, no cloud upload, supports exporting CSV/JSON/ZIP. Using Claude saves you the API key hassle
Project address::
https://t.co/cymTfsIr5Y
Hermes Agent now runs Buzz.
The self-hostable workspace from @blocks puts humans and agents in the same messaging channels and codebase.
Three ways to use Buzz with Hermes (and vice versa):
- Buzz Desktop auto-discovers your Hermes install runs it locally
- A relay bridge gives it a hosted identity in your channels
- Connect via the Hermes Gateway to use Buzz as a full external platform with channels, DMs, threads, reactions, and cron delivery
https://t.co/srljIbERN7
We compared 1-bit Kimi K3 to Claude Opus 5 and GPT 5.6.
We gave 4 models the same prompt:
Create a glass aquarium whose side panel develops a visible crack and then bursts.
1-bit Kimi K3 GGUF ran locally on 4x B200s at 36 tok/s.
GitHub repo: https://t.co/aZWYAtakBP