"Voice-pro has started supporting F5-TTS, CosyVoice and kokoro TTS."
A robust alternative to ElevenLabs, Voice-Pro empowers podcasters, developers, and creators with advanced voice solutions.
https://t.co/oGL4JzPg0x
🚀 Just discovered the genius behind "Voice Cloning and Multilingual TTS in One Click"! This innovative tool transforms our interaction with AI voice technology. The future of communications is here—imagine the possibilities! 🌍💬 Check it out on GitHub https://t.co/urk5MXriEW
Github 👨🔧: Comprehensive Gradio WebUI for audio processing, powered by Whisper engines (Whisper, Faster-Whisper, Whisper-Timestamped). Features Voice Changer(RVC), zero-shot Voice Cloning (E2, F5-TTS), YouTube downloading, vocal isolation(UVR5), Text-to-Speech (Edge-TTS), and multi-language translation. Perfect for content creators and developers.
Helps you process multimedia content with AI-powered tools for YouTube video handling, voice separation, speech recognition, translation, and text-to-speech. It supports over 100 languages for speech and translation, features zero-shot voice cloning, and offers real-time translation capabilities. The tool provides a WebUI for integrated workflows and batch processing, targeting content creators and researchers needing advanced audio and video manipulation.
-------------
What it offers:
🔊 AI-powered multimedia processing for YouTube videos, voice separation, speech recognition, translation, and text-to-speech
🎤 Zero-shot voice cloning using F5-TTS & E2-TTS for creating unique voices
🎥 YouTube video download and audio extraction in various formats
🔇 Professional vocal isolation with UVR5 technology for noise removal
📢 Multilingual text-to-speech with Edge-TTS supporting 400+ voices and 100+ languages
🌍 Instant translation across more than 100 languages for subtitles and dubbing
🔥 AI cover creation using RVC technology for voice modulation and AI voice integration
💻 User-friendly WebUI with dedicated tabs for studio workflows, subtitle generation, translation, speech synthesis, and live translation
🚀 One-click installation and update scripts for easy setup and maintenance
🛠️ Batch processing for handling large volumes of multimedia files
Voice-Pro is a Gradio WebUI for audio processing powered by Whisper engines. Key features include Voice Changer, zero-shot Voice Cloning, Text-to-Speech (Edge-TTS, F5-TTS), and multi-language translation. Perfect for content creators and developers.
https://t.co/sstwiIFBXg
Create a podcast featuring Mark Zuckerberg and Elon Musk! Imagine the conversations these two giants will have. Utilize Voice-Pro's collection of over 50 celebrity voices! With just a 15-second clip, you can achieve amazing voice replication.
https://t.co/sstwiIFBXg
@ThorstenVoice Voice-pro now supports F5-TTS.
With just the script, you can create a podcast using celeb voices.
For more details, please check GitHub:
https://t.co/sstwiIFBXg
@DotCSV Voice-pro now supports F5-TTS.
With just the script, you can create a podcast using celeb voices.
For more details, please check GitHub:
https://t.co/sstwiIFBXg