In another paper, they present the MultiTalk dataset, comprising over 420 hours of talking videos in 20 languages scrapped entirely from YouTube, which they used to enhance 3D talking head generation. (11/23)
In another paper, they present the MultiTalk dataset, comprising over 420 hours of talking videos in 20 languages scrapped entirely from YouTube, which they used to enhance 3D talking head generation. (11/23)
Given a multi-stream audio input, a reference image and a prompt, MultiTalk generates a video containing interactions following the prompt, with consistent lip motions aligned with the audio.
Given a multi-stream audio input, a reference image and a prompt, MultiTalk generates a video containing interactions following the prompt, with consistent lip motions aligned with the audio.
Created using these cutting-edge AI tools: [Midjourney, Dreamnia, Flux Kontext, MultiTalk, Omnihuman, Skyreel, Suno].
Check it out and let me know what you think! ❤️
🧵1/2
youtu.be/JAvzyjMmrvM?...
Created using these cutting-edge AI tools: [Midjourney, Dreamnia, Flux Kontext, MultiTalk, Omnihuman, Skyreel, Suno].
Check it out and let me know what you think! ❤️
🧵1/2
youtu.be/JAvzyjMmrvM?...
AI tools used: Midjourney, Dreamnia, Nano-banana, Kling, MultiTalk, Omnihuman, Skyreels, Suno.
youtu.be/NP6462xhCiw
AI tools used: Midjourney, Dreamnia, Nano-banana, Kling, MultiTalk, Omnihuman, Skyreels, Suno.
youtu.be/NP6462xhCiw
project page: meigen-ai.github.io/multi-talk
code (coming soon): github.com/MeiGen-AI/M...
project page: meigen-ai.github.io/multi-talk
code (coming soon): github.com/MeiGen-AI/M...
For the lip-sync stuff there's a solution for Windows:
MultiTalk Wan2GP (1.3GB model)
But even if ported to Mac and highly performant it's maybe still too limited to use it on a Mac with 10 GB VRAM: maybe we get 480px, 25/s?.
For the lip-sync stuff there's a solution for Windows:
MultiTalk Wan2GP (1.3GB model)
But even if ported to Mac and highly performant it's maybe still too limited to use it on a Mac with 10 GB VRAM: maybe we get 480px, 25/s?.
👉 It animates each character, even in cartoons or singing
👉Open source
#MultiTalk #AI #AIvideo #opensource #GenAI
👉 It animates each character, even in cartoons or singing
👉Open source
#MultiTalk #AI #AIvideo #opensource #GenAI
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
https://arxiv.org/abs/2406.14272
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
https://arxiv.org/abs/2406.14272
MultiTalk: Introspective and Extrospective Dialogue for Human-Environment-LLM Alignment
https://arxiv.org/abs/2409.16455
MultiTalk: Introspective and Extrospective Dialogue for Human-Environment-LLM Alignment
https://arxiv.org/abs/2409.16455
MultiTalk, a novel framework for audio-driven multi-person conversational video generation.
MultiTalk, a novel framework for audio-driven multi-person conversational video generation.