Our goal is to make agents solving real-world problems safe and reliable. Founded by AI/ML engineers and researchers from Meta, Databricks, NVIDIA, and OpenAI.
Can agents learn from past successes — without fine-tuning?
In our latest blog at Tile Labs, we test a simple idea: giving agents examples of successful past runs before they act. No weight updates. No retraining. Just better context.
What we found:
• More selective agents → fewer errors
• Smaller models improve the most
• Frontier models gain better calibration
Read more: https://t.co/7IvWA4f6AU
🚀 Our first blog is live: Benchmarking AI Agents using RL Environments
We evaluate frontier models and uncover a key insight — the gap between completing a workflow and completing it correctly is larger than you think.
Read more: https://t.co/Gh23BjKNjZ
We’re excited to introduce Tile Labs.
A research lab focused on advancing model intelligence — exploring how AI systems become more capable, adaptive, and autonomous.
More to come soon.
[email protected]