Last month the Vals AI team hit the road for campus recruiting. We spent time at @UCBerkeley, @MIT, @GeorgiaTech, and @UofIllinois.
The students we met are already running evals, building agents, and asking sharp questions about the future of AI.
We will be at:
@Stanford on October 6th
@Princeton on October 29th
Can't wait to meet you all!
The model appeared defeated, spending the next several hours doing essentially nothing but farming potatoes. Stream viewers noticed and complained that Astra needed to "pick up the pace." Astra also seemed to become paranoid about creepers: "GREEN tall thing ahead was SUGARCANE, NOT creeper!" At times it was hard on itself: "you can screw up and drop things"; "do NOT waste another night chasing dark pink pixels" (its own disparaging wording for pigs); "our last tool stupidly ended Slabsselected/rightUP and 15sec thinking killed us".
Gov. Shapiro just signed an execute order establishing stringent AI data center regulations, focusing on energy, water and community burdens. But measurements of the environmental footprint of AI are incomplete outside the context of how the technology is used. Our new Environmental Impacts Report ties environmental impact directly to the tasks for which AI is being used.
Today we're announcing our $40M Series A at a $400M valuation, led by @a16z , with participation from existing investors @8vc, @pearvc, and @BloombergBeta and new investors @HRTVentures and @nextladder.
Alongside the fundraise, three more announcements:
- Vals Smith: is now generally available. Anyone can create a custom coding benchmark from any GitHub repo with 120 free credits to get started.
- Frontier Risk Benchmarks: We are releasing the RSI Index in collaboration with @CoreWeave and just launched ReverseEngBench, a new cyber benchmark built with Columbia University, Tufts University, UC Berkeley, and UCLA. We are also sharing our initial work in mental health, with more to come across environmental impact, military, and biosecurity.
- Website + Vals Index 2.0: We completely rebuilt the Vals website and have released Vals Index 2.0, with coverage of more of the economy.
Our revenue has already grown 8x compared to all of 2025. Our customer base doubled and the team tripled in 6 months. Our results have been cited in model cards from OpenAI, Anthropic, Google, Meta, and xAI.
The AI economy runs on self-reported grades. When a model ships, the scores come from the company that built it. No other trillion-dollar industry works this way. Finance has ratings agencies. Medicine has the FDA. AI has vibes and vendor benchmarks. We built Vals to be the independent evaluation layer the industry is missing.
Check out our new website and try Vals Smith!