Foundation to track the progression and creation of the future of Artificial General Intelligence and Artificial Super Intelligence.
AI, written by humans.
As #OpenSource Models become more powerful, the focus of Red Teaming must switch from Black Box to White Box style attacks, keeping models safe even when Open Source Weights fall into Malicious Actor's hands. Check out the paper below 👇
This paper demonstrates an Untargeted Jailbreak attack with up to 80% effectiveness when the weights are known: https://t.co/khJ16WiayM
The current techniques to defend against a gradient attack like this have not even been invented yet.
They must be.
The idea that suddenly creating an #AI that can code as well as a human means it can improve itself into #singularity presupposes that complexity does not explode at a certain threshold- we essentially need an intelligence leagues above us to even have a chance.
@huge_icons Despite the hate, React will never let you down just because of all the documentation and tooling. If I knew I never had to write code for a big project then Svelte all the way.
@sreenandhanpp Absolutely, and I think this is why test performance in Universities across the world is plummeting, its much easier to say you understand a piece of code than to actually understand it, and how to extend it.
@Umesh__digital This is why long horizon benchmarks need to be formalised better, I think we fall into the trap often of producing benchmarks that are "next-step only", and thus determining that a model can reason well.