Skip to main content
Generates realistic talking avatar videos from a face image and audio. Ideal for lip-synced speaking videos.

Request

Required: prompt, audio, face_image
Optional: frames_per_second, num_frames, resolution
Reference inputs take asset ids from POST /v1/uploads or an earlier output’s asset_id. Prices and allowed values come from the live catalog; GET /v1/models/wan-2.2-infinitetalk-fal is always current.