Meta AI Releases WavFlow, a Multimodal Audio Generation Framework That Produces Synchronized High-Fidelity Audio from Video and Text
Meta AI unveils WavFlow, a groundbreaking multimodal framework that generates synchronized, high-fidelity audio directly from video and text inputs, matching top latent-based models on major benchmarks while making its codebase publicly available on GitHub.