🏆 Honored that the AIST Team ranked 1st in DCASE 2026 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes (S5).
Our system combines TF-Locoformer and ATST-Frame in an end-to-end iterative architecture.
https://t.co/9tDHgmFgul
I'll be attending #NeurIPS2025 and presenting our paper (Shallow Flow Matching https://t.co/PRvCaB2ssq) instead of the first author.
If you're interested in speech synthesis technologies, come to San Diego poster session 6 on Friday afternoon!