Upload a portrait
Shoulders up, neutral expression.
数字人口播
A portrait and an audio file become a presenter: head movement, blinks, gestures and lip-sync are all driven by the sound. For product clips, intros, or characters who need to speak to camera.
一张脸、一段音频,出一个会说话的人。
How it works
Shoulders up, neutral expression.
Clean voice, little background noise.
Pick a calmer or livelier gesture style if offered.
You give
You get
Under the hood
Each model lists its own price and content policy. The spicy route is MiniMax H3; other models follow their vendor’s rules.
FAQ
Typically up to 30–60 seconds per run depending on the model; chain runs for longer scripts.
Yes — generate the portrait with a text-to-image model first.
No. Avatar models apply their vendor policy; keep the content within it.
Enter an invite code to open the tool in Chat, or join the waitlist from this page. We open seats in batches and prioritise the tools people ask for most.
Related tools
Invite code opens the tool. No code — join the waitlist; the most-requested tools open first.
Early access
Talking avatar is invite-only for now. Enter a code, or leave your email and we will count the request.