Across all agentic tools and all employees at Uber, weekly active users have grown 7x and weekly agent requests have grown 9.4x.
Despite that growth, total AI spend has relatively stabilized and cost per token has decreased.
In our latest blog, we break down the architecture and operational choices driving AI efficiency at scale, including better prompt caching, access to 1,000+ internal and third-party MCP servers through a single gateway, and an AI Context Graph spanning 30+ internal systems.
The result? Cost per 1,000 requests for a given frontier model has fallen 34% from its peak, and cost per session is down 52%.
Learn more ⬇️
Messaging helps our members connect with each other and opportunity. Our team has been working on improving the search experience within messaging with a new backend called InSearch. https://t.co/TM49LTMwM6