So excited to be officially launching 83 Sciences at
@ycombinator! The future of materials discovery starts with the lab notebooks collecting dust in file cabinets!
The next industrial era will be built on materials discoveries and insights hiding in the 90% of experimental data that never gets published.
Innovate faster, publish more, do more with your data with 83 Sciences
https://t.co/uUnoLgZj3r
We've created a really unique environment to execute on the scope and ambitions of our program. If you're passionate about working full-stack on robotics, please building with us!
I’ve started a new company with @philhchen!
Phil built frontier LLMs across research & engineering at OpenAI, DeepMind, and Scale.
I was shipping AI experiments at Ramp Labs.
We've been heads down building personalized AI coworkers for every business.
We’re growing our team of researchers, designers, and IMO gold medalists. Reach out if you're interested!
excited to share that after nearly 5 amazing years at @scale_AI, last year I joined @OpenAI
one of the many things which made this a special opportunity was the potential openai has to drive widespread enterprise adoption of AI—and frontier is an important step in that direction
Introducing OpenAI Frontier—a new platform that helps enterprises build, deploy, and manage AI coworkers that can do real work. https://t.co/4W0adQzSZ1
Meet Corridor: the security layer for AI coding. Now generally available.
@CorridorSecure is the first security tool that moves at the speed you build – enforcing security guardrails in real time.
Get two weeks free, with plans starting at $20/month.
We’re introducing SEAL Showdown, the AI leaderboard that actually captures real preferences, powered by a platform used by real people.
Public benchmarks today rely on contrived tasks or narrow user groups. That leaves us guessing which models are actually preferred by people.
SEAL Showdown changes that.
Model performance can now be segmented by demographics and domains. Rankings aren’t just a single global average, they can be broken down by region, profession, education level, age and more, giving a nuanced view of how models work for different people.
SEAL Showdown sets a new bar, because AI should be judged by how well it works for everyone.
AI revolutionized coding. Security hasn't kept up–until now.
Introducing @CorridorSecure: the future of secure coding.
We just raised a $5.4M seed from @Conviction and hired @alexstamos. Corridor is trusted by leading companies like @cursor_ai–and we’re just getting started. 🧵
AI improvement at the labs has evolved from a guessing game to one driven by precision-targeted fixes to identified failure modes.
@scale_AI's SEAL research benchmarks and evaluation platform is making this possible.
Check out coverage by @willknight
https://t.co/lB3V6ozU7A
🎉 Excited to share more about the work we’ve been doing at @scale_AI to help AI labs:
✅ evaluate model performance
✅ analyze weaknesses
✅ drive targeted improvements
Thank you to @willknight@WIRED for covering!