BD
ByteDance · Avatars
New

OmniHuman 1.5

Your photo speaks or sings to your voice: lips, facial expressions, and gestures

up to 30 s1080p
30 gen / second of videoAvailable from plan Premium
Try OmniHuman 1.5
Overview

What it can do OmniHuman 1.5

OmniHuman 1.5 by ByteDance animates photos to match your audio: the person speaks or sings with precise lip sync, head movements, hand gestures, and facial expressions. Send a front-facing portrait and a voice recording or song up to 30 seconds, and the clip will be the same length. Great for greetings, music videos, vlogs, and expert content.

Best suited for

BloggingGreetingsClips
  • ✓Lips in sync with audio
  • ✓Songs and speech up to 30 seconds
  • ✓Head and hands move
  • ✓Lips synced precisely to audio: speech, song, rap
  • ✓Head, hands, and facial expressions move — not just the mouth
  • ✓Audio up to 30 seconds — the video will be just as long
How to write prompts

Formula for a good prompt for OmniHuman 1.5

  1. Photofront view, even lighting
  2. Textwhat to say
  3. Tonehow to say
  4. Lengthup to a minute

Tips specifically for this model

  • Send a clear front-facing portrait and voice audio. Prompt is optional: you can specify an emotion or gesture. Price depends on audio length.
  • A clean recording without background music gives the most precise lip-sync.
  • Description is optional: use it to set emotion and gestures.
✕ Bad

say something about the product

✓ OK

Text: «Our service saves one hour a day — I will show how in one minute». Tone: friendly, confident.

Examples

Ready-made prompts for high quality results

Copy the prompt, insert your details — and get a clip matching the sample.

01

Greeting

Speaks warmly with a smile, waves hand at the end.
Result: Photo sends greetings in your voice Try it →
02

Clip

Sings emotionally, sways to the beat, closes eyes on high notes.
Result: Photo sings your song Try it →
03

Tested in AskMeAI studio

Speaks emotionally, gestures with hands, smiles at the end.
Result: the prompt used to test the model in the bot Try it →
What not to do

Mistakes that ruin the result

In OmniHuman 1.5

  • Do not use photos with multiple faces — you need just one person.
  • Do not use dark or blurry shots — lips will sync worse with the sound.
  • Multiple faces in the photo — only one person needed
  • Dark and blurry shots

General rules for avatars

  • Do not use another person’s photo or voice without their consent.
  • Do not pick profile photos, sunglasses, or hands covering the face.
  • Do not write text longer than one minute per clip.
Specifications

OmniHuman 1.5 in numbers

Section
Avatars
Developer
ByteDance
Duration
2–30 s
Quality
1080p
Sound
no sound
Price
30 gen / second of video
Available from plan
Premium
Status
works on the website and in the bot
Price

How many costs

30 gen / second of video
PlanPriceEnough for
Freefree—
Premium450 ★ · 1 monthup to 15 videos
Premium1 200 ★ · 3 monthsup to 45 videos
Premium2 250 ★ · 6 monthsup to 90 videos
Premium4 200 ★ · 12 monthsup to 180 videos
All plans →

Questions About OmniHuman 1.5

Did not find an answer? Message us, real people reply.

Contact support
How much does OmniHuman 1.5 cost?

30 generations per second of video. Generations are AskMeAI internal currency: they are included in Premium subscription and sold in one-off packs.

Can I try OmniHuman 1.5 for free?

OmniHuman 1.5 is available with Premium subscription and one-time generation packs. Try photo effects and chat with free models at no cost.

How to get the best result in OmniHuman 1.5?

Send a clear front-facing portrait and voice audio. Prompt is optional: you can specify an emotion or gesture. Price depends on audio length. Use the formula: photo → text → tone → length.

Where does OmniHuman 1.5 work?

On the askmeai.ru website studio and in the @GPT5Telegbot Telegram bot — shared balance.