AI model · Video
OmniHuman
ByteDance’s OmniHuman generates a synced video from a single reference image and an audio track.
OmniHuman
VideoHavincy- Provider
- ByteDance
- Type
- Video
- Cost
- From 155 credits
Overview
What is OmniHuman?
ByteDance’s OmniHuman generates a synced video from a single reference image and an audio track. Trained on 18,700 hours of human motion, it makes audio the main control signal: lips, expressions and movements follow the voice.
In Havincy, upload a portrait and a voice to get a video of a person speaking, with no motion-capture equipment. The length follows the audio.
Key strengths
- Talking portrait
- Audio drive
- Présentation
Specs
OmniHuman: specifications
- Publisher
- ByteDance
- Generation modes
- talking avatar (image + voice)
- Available since
- July 27, 2025
- License
- Commercial use allowed
- Cost in Havincy
- From 155 credits
Specs taken from the publisher’s public model page. See the documentation on fal.ai
Use cases
What you can create with OmniHuman
Talking head
A person addressing the camera from a photo.
Educational videos
Explanations led by a virtual presenter.
Social media
To-camera formats for TikTok and Reels.
Presentations
An animated speaker for your product pages or demos.
Prompt examples
Prompts to try with OmniHuman
Paste an example into Havincy, then adapt the subject, style and format to your project.
Make this portrait speak with this 20-second voiceover.
Animate our advisor’s photo so he explains our offer.
Create a virtual presenter from this image and this audio script.
Use cases
Compatible skills
Use this model with the following Havincy skills.
How it works
How to create with this model
- Describe the desired result in a Havincy conversation.
- Add a reference when the style, product, or character needs to be preserved.
- Generate, then ask for a variation or improvement without starting over.
FAQ
Frequently asked questions about OmniHuman
How much does OmniHuman cost?
On fal, $0.14 per second of video. In Havincy, about 155 credits for 8 seconds.
OmniHuman or OmniHuman 1.5?
OmniHuman 1.5 is more expressive and offers 1080p. OmniHuman remains a good, slightly cheaper choice for classic talking heads.
Can I use a generated image?
Yes, a photo or a portrait generated in Havincy works, as long as the face is clearly visible.
Ready to create with OmniHuman?
Describe your idea in Havincy, then refine the result in the conversation.