๐PAVAS, a framework for generating physically plausible audio from video, by integrating physics estimation at #CVPR2026!
Led by our intern Hyun-Bin Oh (https://t.co/VpKVdmSFqN), in collaboration with @kamitsutoshi, @Tae_Hyun_Oh, and @mittu1204.
๐ง&๐: https://t.co/PDsr1MSerq
๐๐๐๐ง๐, ๐ผ๐๐ฟ ๐ฟ๐ฒ๐ฐ๐ฒ๐ป๐ ๐ฝ๐ฟ๐ผ๐ท๐ฒ๐ฐ๐ ๐ผ๐ป ๐ต๐ถ๐ด๐ต-๐ณ๐ถ๐ฑ๐ฒ๐น๐ถ๐๐ ๐ฏ๐ ๐๐ฎ๐๐๐๐ถ๐ฎ๐ป ๐ฎ๐๐ฎ๐๐ฎ๐ฟ ๐๐๐ป๐๐ต๐ฒ๐๐ถ๐ ๐ต๐ฎ๐ ๐ฏ๐ฒ๐ฒ๐ป ๐ฎ๐ฐ๐ฐ๐ฒ๐ฝ๐๐ฒ๐ฑ ๐๐ผ ๐๐ฉ๐ฃ๐ฅ ๐ฎ๐ฌ๐ฎ๐ฒ!
๐ก๐ง๐;๐๐ฅ
We study a mutually reinforcing synergy of 2D & 3D face priors (generative & data priors) to tackle the longstanding challenge in monocular synthesis of animatable, photorealistic 3D face avatar โ Balancing between in-the-wild generalization and efficient synthesis.
๐๐๐๐ง๐: ๐๐ณ๐ณ๐ถ๐ฐ๐ถ๐ฒ๐ป๐ ๐๐ฎ๐๐๐๐ถ๐ฎ๐ป ๐๐ฒ๐ฎ๐ฑ ๐๐๐ฎ๐๐ฎ๐ฟ ๐ณ๐ฟ๐ผ๐บ ๐ฎ ๐ ๐ผ๐ป๐ผ๐ฐ๐๐น๐ฎ๐ฟ ๐ฉ๐ถ๐ฑ๐ฒ๐ผ ๐๐ถ๐ฎ ๐๐ฒ๐ฎ๐ฟ๐ป๐ฒ๐ฑ ๐๐ป๐ถ๐๐ถ๐ฎ๐น๐ถ๐๐ฎ๐๐ถ๐ผ๐ป ๐ฎ๐ป๐ฑ ๐ง๐๐๐-๐๐ถ๐บ๐ฒ ๐๐ฒ๐ป๐ฒ๐ฟ๐ฎ๐๐ถ๐๐ฒ ๐๐ฑ๐ฎ๐ฝ๐๐ฎ๐๐ถ๐ผ๐ป
๐ ๐ฃ๐ฟ๐ผ๐ท๐ฒ๐ฐ๐ ๐ฝ๐ฎ๐ด๐ฒ: https://t.co/gsbkL5lYeC
๐ ๐ฃ๐ฎ๐ฝ๐ฒ๐ฟ: https://t.co/sJGGGHRQ7p
ELITE was a project that I've worked on with all my heart, and thanks to my collaborators' tremendous efforts and their patience. Huge congrats to my coauthors: Lee Hyoseok, Subin Park, @GerardPonsMoll1 , @Tae_Hyun_Oh
Now that we have building blocks for "๐ฅ๐๐ค๐ฉ๐ค๐ง๐๐๐ก๐๐จ๐ฉ๐๐" and "๐จ๐ค๐๐๐๐ก" world simulations. What'll be the next steps? ๐ง
The work is led by the amazing Sungbin Kim https://t.co/IMN9ZyYhPs, and collaborated with Jeongsoo Choi, Joon Son Chung, @Tae_Hyun_Oh, David Harwath
Checkout https://t.co/LEBteX4JSz for more samples, and the forthcoming code and model!
The work is led by the amazing Sungbin Kim https://t.co/IMN9ZyYhPs, and collaborated with Jeongsoo Choi, Joon Son Chung, @Tae_Hyun_Oh, David Harwath
Checkout https://t.co/LEBteX4JSz for more samples, and the forthcoming code and model!
Atย #ICLR2025, we will present "NeuFace: A Large-Scale 3D Face Mesh Video Dataset via Neural Re-parameterized Optimization."
๐ Hall 3 + Hall 2B #69
๐ Thu, Apr 24, 3:00โ5:30 pm Singapore Time
I'd really like to meet & discuss with fellow researchers. Let's connect!
(1/3)
Please come and see this IJCV Special Issue on Audio-Visual Generation. (Deadline: 15th April, 2025)
Fast process, conference extension, and potential opportunities to be presented at an ICCV workshop!
Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration
Kim Jun-Seong, GeonU Kim, @ug___k, Yu-Chiang Frank Wang, @choe_jaesung, @Tae_Hyun_Oh
tl;dr: distil language knowledges into 3DGS
https://t.co/U329pxVh8t
@taiyasaki That was the most frustrating part as an AC. Invite emergency reviewer(ER)->wait->decline->invite->wait->decline->...
Even those ERs seem to have availabilities in the slot.
Also, the CVPR openreview setting doesn't show emails. That's another hassle to find emails separately
๐จ Call for Papers: Special Issue on Audio-Visual Generation ๐ฅ๐ต
We, guest editors, prepared an exciting new Special Issue on Audio-Visual Generation! ๐ The International Journal of Computer Vision (IJCV) is now accepting submissions.
Please find the call for papers below!
A new special issue on Audio-Visual Generation is open for submissions in the International Journal of Computer Vision. For more information on the special issue & how to submit, access the CfP here: https://t.co/s1zBhS1Hnu
A new special issue on Audio-Visual Generation is open for submissions in the International Journal of Computer Vision. For more information on the special issue & how to submit, access the CfP here: https://t.co/s1zBhS1Hnu
AI can restyle large 3D scenes based on reference images!
FPRF is able to stylize NeRF scenes with multiple reference images without additional optimization while preserving multi-view appearance consistency.
https://t.co/nZMSVhJ5Rd
๐ Our paper, "Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields", has been accepted to AAAI 2024.
TLDR; "Photorealistically Stylizable" city-level NeRF in a "Feed-Forward" manner!
Paper: https://t.co/aVZwu97gTe
Page: https://t.co/bUDCjmq5as
This work is a nice collaboration with GeonU Kim and @Tae_Hyun_Oh.
GeonU, the 1st author, was just a 1st semester M.S. student when he started this project! His great efforts and productivity led to this amazing result ๐
Page: https://t.co/bUDCjmq5as
Our expressive 3D talking head generation work will be presented in this WACV 2024 (6th, Jan.).
Our lab is making progress in understanding social expressions. The first step is laughing. Among many expressions of humans, laughing is one of the most important social signals.
Check out LaughTalk, our new #WACV2024 paper!
Project page: https://t.co/8YAWyPkdfP
Arxiv: https://t.co/ezQxTXCdXp
In this work, we propose a 3D talking head model, LaughTalk, that can simultaneously express speech๐ฃ๏ธ and laughter๐
Check out LaughTalk, our new #WACV2024 paper!
Project page: https://t.co/8YAWyPkdfP
Arxiv: https://t.co/ezQxTXCdXp
In this work, we propose a 3D talking head model, LaughTalk, that can simultaneously express speech๐ฃ๏ธ and laughter๐
๐ Exciting News! Our latest work, archived on December 18, 2023, introduces a groundbreaking Text-to-Texture synthesis method.
https://t.co/HNSB8OAzSh
https://t.co/DU8dujGB4A
Here are some key highlights: