Alibaba's latest audio-driven portrait video generation framework, EMO, can create videos of any duration based on input audio. Developed by the Alibaba Intelligent Computing Research Institute team, this expressive video generation technology represents a significant improvement over previous AI video generation methods, though it also has the drawback of being time-consuming. The team, including Bo Liefeng, detailed the technical route and features of EMO in their paper. This new technology marks a breakthrough in the AI field, sparking great anticipation for future developments.
Alibaba Launches Audio-Driven AI Video Generator EMO
10.7K views
Related

Alibaba Launches New Qoder: Upgraded from AI Programming Tool to Intelligent Agent Workstation for Everyone
21.9K
Jack Ma and Management Continuously Increase Holdings by Over 8 Billion Hong Kong Dollars, Alibaba's 8 Billion Hong Kong Dollar Placement Funds Fully Invest in AI
10.9K
Alibaba Invests 8 Billion HKD in AI: Nearly 3 Times Over-subscribed, Joseph Tsai and Wu Yongming Also Invest
12.1K

AliVideo Large Model Wan3.0 Officially Launched, Fully Achieving Ultra-Long Generation and Multi-Dimensional Consistency
14.9K

Alibaba Open Sources Qwen-UI-Agent: The Real-World Multimodal Foundation Model Arrives
33.5K

Mobile-First Baseline Far Exceeds GPT and Claude! Alibaba Launches Qwen-UI-Agent, Marking a New Era for GUI Agents
17.2K
