Getting ready to launch one of our most requested feature: iOS & Android simulators in the cloud
Here's @joshlebed showing Niteshift:
- making changes to an iOS app & backend API
- testing it all e2e in the cloud
- 15 second iteration loop (much faster soon!)
Top teams like @standardbots, @ListenLabs, @elicitorg, and @AmbrookAg run Niteshift against the hardest stacks in software: robot simulators, data-heavy agents, and hours-long eval loops
And since it runs in the cloud, anyone – engineers, PMs, and designers – can work in code
We're launching @niteshiftdev – the full-stack cloud for coding agents
Verification is the new bottleneck.
Software teams can now define their dev environment and verification tools once. Then run any frontier agent in the cloud: Claude Code, Codex, or OpenCode
.@GentraceAI is continuing to grow.
We are looking for senior SWEs in NYC and SF who want to build tools to help the world's most advanced companies make their generative AI systems reliable and predictable.
We are now 10 strong, with more customers and usage than before.
DM @dougsafreno, @virtuallyvivek or me if interested:
Most engineers approach LLM-as-a-judge all wrong.
The usual high-level metrics like hallucination or safety rarely tell you if your app actually works as intended.
Let’s talk about why that’s a problem and how to fix it: