Tauch mit uns ins Thema künstliche Intelligenz 🤖 versus Kreativität 🎨
Wir sprechen mit Thomas Bürgin über 𝙥𝙤𝙨𝙞𝙩𝙞𝙫𝙚 𝙪𝙣𝙙 𝙣𝙚𝙜𝙖𝙩𝙞𝙫𝙚 𝙀𝙧𝙛𝙖𝙝𝙧𝙪𝙣𝙜𝙚𝙣 𝙢𝙞𝙩 𝙆𝙄-𝙏𝙤𝙤𝙡𝙨.
Die Folge auf YouTube: https://t.co/ehL4WEzuXB
Google DeepMind's CEO just stunned 60 Minutes viewers.
The Nobel prize winner revealed:
• An AI that can see and understand in real-time
• A plan to end ALL diseases in 10 years
• Exactly when AI will surpass human intelligence
Here are his 4 most jaw-dropping insights:🧵
These demos all show examples of “multimodal prompting” — giving Gemini combinations of different modalities and having Gemini respond. Here’s how we made them — and some ideas for your own multimodal prompts. https://t.co/9Dh1EFitKM
The UX for generating and manipulating media with AI is going to change dramatically because the media itself is changing.
Right now it's mostly text2image and image2video. But 3D is coming up really fast.
In the next 12 months we'll be generating entire 3D scenes in Luma or Midjourney.
You'll be creating NeRFs or Gaussian Splats, not images. Those will serve as our "digital sets".
You'll be able to navigate through 3D scenes and adjust your angle, perspective, depth of field, focal length, etc.
Inside those scenes is where we'll capture photos and shoot videos.
Imagine going image2nerf, moving your camera around a 3D scene, keyframing out a camera path.
Imagine segmenting objects and characters within those scenes, prompting action and VFX.
Imagine tracing movement paths or controlling character movement and expressions using your front facing camera and AR.
Luma is ushering in a new medium in the form of NeRFs/Gaussian Splats. Midjourney is heading in that direction now as well.
They are innovating on the media itself, and the features are going to be built for that new medium, not for what we see today.
Initially it will look like a static 3D image you can adjust your camera angle and set the depth of field.
After that we'll be generating full 3D scenes with movement and VFX.
Both of these tools are VERY well positioned to be the entry point for the masses to begin exploring 3D.
This is the beginning of OS1 and Her.
With the glasses getting sight next year. The ability to have your own personal assistant always with you will be mind bending.
ChatGPT can now see, hear, and speak. Rolling out over next two weeks, Plus users will be able to have voice conversations with ChatGPT (iOS & Android) and to include images in conversations (all platforms).
https://t.co/uNZjgbR5Bm
Imagine you’re scrolling down the X app on visionOS
You see your friend is about to dive off a hot air balloon
You click, and now you’re skydiving right with him
Or a friend at a TS concert, or in Japan
I think this is one of the most compelling use cases of social media + VR🔥