A new Gradium TTS model is in public beta. It reads phone numbers, emails, IBANs and time expressions correctly, natively.
Test it via the API, send feedback on your production cases, get 1M credits.
Prompt for your coding agent to start in seconds: https://t.co/gAK4fmO7uJ
Gradbot is Gradium's prototyping framework for building voice agents.
This week, @BhosalePratim sat down with cofounders @neilzegh and @lmazare to walk through the architecture and what matters when picking a harness.
https://t.co/IqI63dMlfD
Amazing work lead by @BhosalePratim, vibe code your own voice agent in minutes and benefit from asynchronous function calling: your LLM can trigger some tool calls while the bot is still talking with you! 🤖 🚀
Happy that our paper "Incorporating End-to-End Speech Recognition Models for Sentiment Analysis" with @EgorLakomkin got accepted to #ICRA2019
https://t.co/gf4JvjsFb2
The code for our #emnlp2018 paper for collecting&filtering automatically speech samples from YouTube is here:https://t.co/hiJujgAoBy Few improvements are coming! paper: https://t.co/FlhJ7T6hYQ
Our paper "Automatic Dataset Construction for Speech Recognition from YouTube Videos" got accepted @emnlp2018! Constructive and helpful reviews - now working on improving the paper for the final! #emnlp2018
Yay! Our paper "On the Robustness of Speech Emotion Recognition for Human-Robot Interaction with Deep Neural Networks" got accepted at @iros_2018@ma_zamani#iros2018 https://t.co/fl55RjjsNI