Hermes Agent is now dramatically more efficient, especially for smaller/weaker/local models!
With the help of @nvidia's Nemo Relay and several other strategies Hermes was able to identify a ton of optimizations beyond just saving tool execution time or memory - but also turns needed to complete tasks, schema improvements to reduce context load, and token efficiency gains by tracing through wasted turns and tool errors across all 250,000 conversations I've had with my Hermes.
All of the below are now in Hermes Agent, update to start saving now or wait until tomorrow for the full version update release.
Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with @NVIDIA
✅ A fully open stack enterprises can own and customize
✅ Benchmark-leading performance
✅ Over 10x lower inference costs
Blog: https://t.co/9QumeM3897
Video: https://t.co/KmTKS8r2fk
"LLMs are mismanaged geniuses" is a common sentiment these days.
Basically: We're not even scratching the surface of current model intelligence because we don't know how to harness them most effectively.
But we know that harnesses will be like the operating systems of future companies.
LangChain - the absolute legends - are doing great work to try and start managing our geniuses a little bit better.
It's incredible to see the lift they got with Nemotron 3 Ultra from their work, and this is only just starting to scratch the surface!
Teach an agent a workflow once. Have it remember after every rebuild.
This tutorial shows how to deploy @nousresearch Hermes Agent with NVIDIA NemoClaw and OpenShell, connect it to Slack, Outlook, GitHub, and NVIDIA developer forums, then turn a chat correction into a reusable skill.
Private data stays behind runtime policies. Learned skills persist across deployments.
We have worked with @nvidia to integrate their official Agent Skills catalog into the Hermes Skills Hub.
These skills teach your agent how to use CUDA-X libraries, Omniverse and Physical AI workflows, NeMo training and inference tools, and other platform components.
We are excited to share that NVIDIA NemoClaw now runs on Olares OS!
You can now run your entire AI agent stack locally and securely on devices ranging from the Olares One desktop to NVIDIA DGX Spark systems. @NVIDIA_AI_PC
Learn more:
https://t.co/FgHZTeiI1U