avatar
ST
SadTalker
Fast and affordable talking head generation from a single photo. Good for testing and quick iterations. Expression control available.
Try SadTalker
Generating withSadTalker5c per generation
Created with SadTalker
Features
- Fast Generation
- Expression Control
- Budget Friendly
- Single Photo Input
Specifications
- Resolution
- 256x256 to 512x512
- Input
- Photo + Audio
- Audio Limit
- 30s
- Output
- MP4 Video
Input Requirements
Portrait Photo*
image upload
Front-facing portrait photo
Audio File*
audio upload
Speech audio (max 30s)
Expression Scale(optional)
slider
Still Mode (head only)(optional)
checkbox
Preprocess(optional)
select
Related Models
OmniHuman v1.5
Photo + Audio to talking avatar
from 2 credits · $0.32-$9.60
Kling Avatar v2
Versatile lip sync for any character
from 2 credits · $0.23-$13.80
Sync-3 Lipsync
Video dubbing with 4K lip sync
from 2 credits · $0.27-$16.01
Hunyuan Avatar
Talking and singing, up to 120s
Fabric 1.0
Photo + audio talking avatar
from 1 credits · $0.16/s+
Infini Talk
Audio-driven talking avatar
from 4 credits · $0.40/s+
Wan 2.2 S2V
Speech-to-video from photo + audio
from 3 credits · $0.50-$3.00
Frequently Asked Questions
How much does SadTalker cost?
SadTalker costs 5 credits per generation (~$1.00). You get 10 free credits every day to try it.
Can I use SadTalker outputs commercially?
Yes, all content generated with SadTalker on Arteza comes with a commercial license.
What file format does SadTalker output?
MP4 video files with lip-synced audio.