@AndrewYNg I want to learn more about llm and how to structure/extract knowledge (more or less) automatically by having them crawl a lot of unstructured data.
Actually, gradient descent can be seen as attention that applies beyond the model's context length! Let me explain why 🧵 👇 (1/N)
Ref:
https://t.co/BXQvCV60pa
https://t.co/i5lte2kuMW
🧵This was another crazy week in AI. Google made a $300M investment in Anthropic, Microsoft is integrating OpenAI into their products, and Apple said AI would impact every product they have.
Here's a summary of what happened this week 👇
Models such as Stable Diffusion are trained on copyrighted, trademarked, private, and sensitive images.
Yet, our new paper shows that diffusion models memorize images from their training data and emit them at generation time.
Paper: https://t.co/LQuTtAskJ9
👇[1/9]