Let's go! MetaVoice 1B 🔉
> 1.2B parameter model.
> Trained on 100K hours of data.
> Supports zero-shot voice cloning.
> Short & long-form synthesis.
> Emotional speech.
> Best part: Apache 2.0 licensed. 🔥
Powered by a simple yet robust architecture:
> Encodec (Multi-Band Diffusion) and GPT + Encoder Transformer LM.
> DeepFilterNet to clear up MBD artefacts.
Synthesised: "Have you heard about this new TTS model called MetaVoice."
Google just made an incredible AI video breakthrough with its latest diffusion model, Lumiere.
2024 is going to be a massive year for AI video, mark my words.
Here's what separates Lumiere from other AI video models:
📊 Cube.js: An Open Source Analytics Framework - https://t.co/tdY00m86ZP (For building things like internal BI tools or customer dashboards. There's a guide to its use here.)