ByteDance Develops OmniHuman, an AI Framework That Can Generate Realistic Videos of Humans

OmniHuman can generate realistic videos from a single human image and motion signals such as audio or video.

Advertisement
Written by Akash Dutta, Edited by Siddharth Suvarna | Updated: 7 February 2025 14:31 IST
Highlights
  • OmniHuman can generate full-body videos
  • The AI system was trained on 18,700 hours of human video data
  • It is a research work and the model is not available in the public domain

OmniHuman can match lip movement and gestures with speech or music

Photo Credit: Unsplash/Markus Winkler

ByteDance, the company behind TikTok, recently shared its research on a new artificial intelligence (AI) framework. Dubbed OmniHuman, it is a video-generation framework that can create realistic human videos with full-body movement and lip-syncing. The researchers stated that it requires a human image along with motion signals such as video or audio to generate output. Several demonstration videos generated using the AI model have also been shared, showcasing the realism of the final output. Notably, the company stated that the AI model is available in the public domain.

OmniHuman Can Generate Realistic Human Videos

The researchers shared several demonstrations and detailed the framework on its website. It is an end-to-end system that was built using a novel multimodality motion conditioning mixed training strategy, the post claimed. While the researchers did not share any benchmark metrics, they claimed that the AI model “significantly outperforms existing methods.”

Advertisement

OmniHuman can generate videos using an image of the person and a motion signal. Motion signals can be audio only, video only or a combination of audio and video. The AI model can generate realistic videos based on text prompts. These videos can be full-body where the limbs, facial expressions, and lip movement can be synced with the audio or music playing in the background. OmniHuman can generate videos in different aspect ratios, allowing flexibility to users.

OmniHuman output example
Photo Credit: OmniHuman

 

The use of motion signals is a novel technique, which the company is calling omni-conditions training. With this, the AI model is trained on different modalities, including text, image, audio, and video. Researchers said this allowed the model to learn mixed conditioning which overcame the scarcity of high-quality data.

Advertisement

Notably, the model was trained on 18,700 hours of human video data. The details about the training process have been documented in a paper published in the online pre-print journal arXiv.

The company also shared several demonstrations of videos generated using the model, and the results appear to be highly realistic with natural body movements, hand gestures, and lip movements. Such realism has also raised concerns about deepfakes. However, the company has specified that the AI model is currently not available to be downloaded, and there is no service people can use to access its capabilities.

 

Get your daily dose of tech news, reviews, and insights, in under 80 characters on Gadgets 360 Turbo. Connect with fellow tech lovers on our Forum. Follow us on X, Facebook, WhatsApp, Threads and Google News for instant updates. Catch all the action on our YouTube channel.

Advertisement
Popular Mobile Brands
  1. Flipkart Big Billion Days Sale 2026: iPhone 17, iPhone 15 Deals Revealed
  2. Vivo V80 Could Be More Expensive Than Its Predecessor, Leak Suggests
  3. Lumio Aura 5 With 4.5-Inch Aluminium Woofers Debuts in India: See Price
  4. These Xiaomi and Redmi Smartphones Could Launch Soon in China
  5. Mivi One 5G First Impressions
  6. Two New The Last of Us Games Are in Early Stages of Development
  7. Lava Shark 2 Pro 5G With a 6,000mAh Battery Arrives in India at This Price
  8. Pixel Buds App Update Brings Gemini Voice Controls
  9. Nothing Phone (4a), (4b) Now Available in New Variants in India: Prices
  10. Oppo K14 Plus Launched in India: Price, Specifications
  1. Amazon Fire TV Stick 4K 2026 Design Renders Leaked Ahead of Launch
  2. Oppo K14 Plus Launched In India With Dimensity 7360 Max SoC, 8,000mAh Battery: Price, Specifications
  3. Lava Shark 2 Pro 5G Launched in India With 6,000mAh Battery, 50-Megapixel Rear Camera: Price, Specifications
  4. Lumio Aura 5 Launched in India With Dual 4.5-Inch Aluminium Woofers, 200W Peak Output: Price, Specifications
  5. Naughty Dog Confirms Two New The Last of Us Games, Says Will Fully Reveal Intergalactic in 2027
  6. Vivo V80 Price in India, Storage Configurations Tipped Ahead of Launch
  7. Google Pixel Buds Pro 2, Pixel Buds 2a Get New Gemini Voice Controls With Latest App Update
  8. Apple Ordered to Pay $5.7 Billion Over Haptics Patents Used in iPhone, Apple Watch
  9. Flipkart Products Get Buy Button in Google Gemini and AI Mode Test: Report
  10. Samsung Galaxy Z Fold 8, Galaxy S25 Series Tipped to Receive Price Hikes After Galaxy A56 and Galaxy A36
Download Our Apps
Available in Hindi
© Copyright Red Pixels Ventures Limited 2026. All rights reserved.