@nasrinmmm Startups have a very specific market niche and are mainly based on existing Open Source(OS) technologies with a lot of potential that still need some refinement. And this is what we are experiencing today. Lots of new OS DL models that can be used for a wide variety of apps.
There is so much AI research momentum, I cannot remotely keep up. The pace is unreal.
There are gems and prizes in GitHub repos everywhere—mostly unnoticed.
Research Friday - One TTS Alignment To Rule Them All (https://t.co/YKEibn13OV) Easy but powerful idea. Once you read it, it seems obvious that you can use CTC for alignment learning. Especially, when it's a paradigm for STT models... How haven't we used it before for TTS? Smart!
Research Friday - CLAP : Learning Audio Concepts From Natural Language Supervision (https://t.co/oUcC3iWX8t) I have already read about CLIP but I am curious about how they approach it for audio.
@cyrta@csteinmetz1 I have never been on that situation. It sounds like there is no trivial solution. Maybe instead of doing it after the talk try to make it live with headsets. Although it does not sound
comfortable at all.
Reading this post reminds me how difficult is to keep up to date the sota... Many interesting works... Little time... And an incredible world out there. Overwhelming and beautiful at the same time!
If you are looking for your niche in AI, here are five topics to avoid 😉
https://t.co/NIFuslBR1c
Very nice to see deep generative models (diffusion models) to be listed out.