✨ Excited to share our latest work from The Data Provenance Initiative ☸️
This is the most comprehensive audit of multimodal training data, auditing ~4000 datasets between 1990 and 2024, and covering more than 400 unique tasks in 608 languages!
🧵 1/n
I can’t say enough good things about John Carmack @ID_AA_Carmack and his Keen Technologies. But now Khurram Javed @kjaved_ and I have broken away to start our own startup and pursue a slightly different path toward understanding intelligence. Like Keen (and like Ineffable) we at Oak Lab @oaklab_ai believe in reinforcement learning and that intelligence is created and maintained from run-time experience. But we think current deep learning methods are weak and inefficient, and need not more tweaks, but fundamentally new ideas and a thorough reworking before they can provide a solid foundation for achieving the more ambitious goals of AI.
it’s surprising to me how many people seem to not understand that great models are built with super high quality curated data
finding novel ways to create / get this data is a huge edge
The new DeepSeek Engram paper is super fun! It also integrates mHC, and I think they're probably releasing all these papers to make the V4 report of reasonable length😄
Here's a nice short summary from @GeminiApp 🫡
On this day, the Heuristically programmed ALgorithmic computer, better known as HAL 9000, became operational
"I am putting myself to the fullest possible use, which is all I think that any conscious entity can ever hope to do."
- HAL 9000
On the slow death of scaling
Because we can't scale forever, and the world is not enough
Give it a chance, this isn't another deep learning hitting a wall...
https://t.co/wKDb5DD3wd
Announcing FunctionGemma, a specialized version of our Gemma 3 270M model that’s fine-tuned for function calling ⚙️
The new release brings bespoke function calling to the edge, and is designed as a strong base for further training into custom, fast, private, local agents that translate natural language into executable API actions.
https://t.co/nkfZAKgBMm
I just want to confirm that this is based on a real document and we did train Claude on it, including in SL. It's something I've been working on for a while, but it's still being iterated on and we intend to release the full version and more details soon.
I do not think you can pursue meaningful research without (1) some grandiose delusion about your abilities (2) a sense of esthetics and harmony to judge ideas still free of experimental confirmation (3) an unreasonable taste for the required tangible work (e.g. programming)