Hello Community!
Just recently dropped my first collection of NFT
portraits of a businessman on opensea
Much love!
hope you like it 🖤🖤
https://t.co/8Hqwigu6R6
#nft#ntfcommunity#cryptoart#opensea
TikTok presents Boximator!
This method can generate rich and controllable motions for image-to-video generations by drawing box constraints and motion paths onto an image and combining it with a prompt:
"A girl in red is covering her face with a skull."
10 crazy examples:
I spent all day learning Metahuman Animator and managed to create a digital version of myself. It's using one of our new Scan Store Metahuman identities and textures (my face), captured using an iPhone. Can't wait to finish these and get them out there.
The Stable Video Diffusion model just dropped 🔥
The new model supports:
– Text-to-Video
– Image-to-Video
– 14 or 25 frames at 576 x 1024
– Multi-View Generation
– Frame Interpolation
– 3D Scene Understanding
– Camera Control via LoRA
Paper: https://t.co/viLhcbQ3oN
Code: https://t.co/J5SvRDJnD9
SVD model: https://t.co/5ipFQyXHoU
SVD-XT model: https://t.co/GCsqmptnrE
My new favourite 3D creation workflow 😍
1. Text-to-image with Stable Diffusion XL (on TPU!)
2. Image to 3D with Gaussian Splatting
3. Endless possibilities with the output 3D model
🔽 Links below
The famed Stanford Smallville is officially open-source!
25 AI agents inhabit a digital Westworld, unaware that they are living in a simulation. They go to work, gossip, organize socials, make new friends, and even fall in love. Each has unique personality and backstory.
Smallville is among the most inspiring AI agent experiments in 2023. We often talk about a single LLM's emergent abilities, but multi-agent emergence could be way more complex and fascinating at scale. A population of AI can play out the evolution of an entire civilization.
Endless new possibilities ahead. Gaming will be the first to feel the impact.
Github: https://t.co/xUll7KaaTp
Paper: https://t.co/PMDQysrOz9
Authors: @joon_s_pk@joseph_c_obrien@carriejcai@merrierm@percyliang@msbernst
Let's talk numbers, the original 2D video is 7.3Mb; the needed data to play it in this 3D volumetric way, which allows you to change the camera position and focus realtime is just an extra 11Mb.