Today, we are open-sourcing the Innate OS, the intuitive agentic OS for general-purpose robots
🎮 Try it today in our simulator and make the smartest robot
⬇️ Here is a demo with two robots collaborating
We are entering the physical agents era, but building for mobile robots is still reserved for PhDs. We designed the Innate OS for intuitiveness, so that everyone can start automating their lives.
Innate OS runs on our <$1k open-source robot or on your computer. It takes just a minute to build your first agent in our simulator.
I told them they should make it part of the company's DNA to be really nice. Startups often carry some mark of their origins as they expand into other markets, and this would be a great thing to carry with you.
Last year, we moved into our offices in heart of FiDi. This was a foundational moment for us, and now we're inviting the community to join us for our office warming. Space is limited, apply below:
Last year, we moved into our offices in heart of FiDi. This was a foundational moment for us, and now we're inviting the community to join us for our office warming. Space is limited, apply below:
2025 wrapped:
It was a tough year being steeped in the "trough of disillusionment" with GenAI. Enterprises exhibited lower confidence than ever that AI could actually be a force multiplier for their orgs. But what I'm most proud of isn't what's in this picture.
Of course, a rapid growth in our F500 deployments and our public launch getting to the front page of HN, is great, but what I find truly exceptional is the team we've built.
We brought on star operational and engineering team members that allowed us to handle complex application environments and bureaucratic internal processes that would've typically taken a team 5-10x our size.
Our most notable accomplishment this year is breaking through the noise in what I glibly characterize as a sea of slop - amongst countless evals, observability tools, fine-tuning platforms etc. our message of removing the friction between regulated industries and non-deterministic AI resonated at the highest levels.
Over the last 3 weeks, the team @ctgtinc has been pulling late nights + long weekends.
We are not a billion dollar AI company, we are a small team <15 with huge conviction that there’s a need to deploy AI safely in domains like finance, law, and support.
Cyril and the team at CTGT are productizing mechanistic interpretability. They make it possible to edit the behavior of LLMs to add safety policy guarantees without retraining, in a way that is much more reliable than simple prompting.
We just launched @CTGTInc's Mentat, an OpenAI-compatible API that gives enterprises deterministic control over LLM behavior.
Benchmarks showed clear gains in accuracy, truthfulness, and hallucination prevention.
We built it after seeing models ignore correct information or produce misconceptions because of the internal patterns they rely on during generation.
Prompts and RAG could not correct that.
Mentat uses our policy engine to govern behavior at the feature level in real time.
It aligns output with an organization’s rules and trusted information and resolves inaccuracies before they reach the end user.
This reduces hallucinations and keeps answers consistent without fine tuning or fragile prompt stacks.
Full breakdown here: https://t.co/DmmsB2ATNK
Today we launched Mentat, an OpenAI-compatible API that gives builders deterministic control over LLM behavior.
CTGT-governed models now deliver frontier-level reliability.
In our evaluations, GPT-120B-OSS reached 96.5% hallucination reduction on HaluEval, surpassing frontier models like Gemini 3 and Claude 4.5 Opus.
All of this is powered by CTGT’s Policy Engine, which evaluates content in real time against an organization’s rules and trusted information.
We are building the infrastructure that lets companies move from “AI that usually works” to AI they can trust every time.
The quality and density of talent in the age of AI is unprecedented. There's a lot more slop, but high agency people are able to go farther than ever before.
no bbq and parties, this is how july 4th looked for our cto at 5am this morning
when a fortune 20 co drops a feature request, you have to skip out on everything your friends are doing
people talk a lot about the highs of being a founder, but skip over the quieter moments
other companies salivate at getting awards like this.
we don't. we know they're meaningless for our goal of pushing the frontier of AI in high-risk applications. the only thing that matters is driving real results for our customers.
“There is a narrow way through.”
Only @CTGTInc knows what's "actually working in enterprise AI"
Everything else is low RoI, smoke and mirrors, or a combination of both
https://t.co/OpO3lhHJt4
never seen a group of more cracked students than at our AI startup school afters
2 IOI gold medalists, 3 IMO gold/silver medalists, Putnam fellows…unreal
thanks @ycombinator and @GradientVC for helping put together an amazing night