infery
← All models

EchoMimic V3

echomimic-v3

Video editby Echomimic V3

EchoMimic V3 generates a talking avatar model from a picture, audio and text prompt.

Details

Accepts
image + audio

Pricing

Price
25 cr / second

Prices in credits (1 credit = $0.01).

Data schema

Input

FieldTypeDescription
seedThe seed to use for the video generation.
promptrequiredstringThe prompt to use for the video generation.
audio_urlrequiredstringThe URL of the audio to use as a reference for the video generation.
image_urlrequiredstringThe URL of the image to use as a reference for the video generation.
guidance_scalenumberThe guidance scale to use for the video generation.
negative_promptstringThe negative prompt to use for the video generation.
audio_guidance_scalenumberThe audio guidance scale to use for the video generation.
num_frames_per_generationintegerThe number of frames to generate at once.

Output

FieldTypeDescription
videoThe generated video file.