Video
AI Avatar Multi (Text)
Credit-based pricing. Supports image-to-video.
Output
video
Resolution
480p, 720p
Inputs
- Start image (required)
Example prompt
Upload a portrait photo, type what each speaker says, pick voices — AI Avatar Multi generates a two-person talking video with lip-syncTry this prompt →
LLM-ready
API & LLM schema
Exact request contract for this model. Agents can fetch it from /api/v1/models?id=fal-ai/ai-avatar/multi-text.
POST
/api/v1/generate16 fields · 4 required| Field | Type | Requirement | Contract |
|---|---|---|---|
model_id | constant | Required | ArtEmotion model identifier. |
extra | object | Optional | Model-specific settings may also be nested here. |
max_credits | number | Optional | Reject before submission if the estimated list price exceeds this cap. · Range: 1–… |
webhook_url | string | Optional | Format: uri |
webhook_secret | string | Optional | Optional model input. |
folder_id | string | Optional | Optional model input. |
prompt | string | Optional | Optional model input. |
acceleration | string | Optional | The acceleration level to use for generation. · Allowed: none, regular, high · Default: regular |
voice2 | string | Optional | The second person's voice to use for speech generation · Allowed: Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill · Default: Roger |
second_text_input | string | Required | The text input to guide video generation. |
seed | integer | Optional | Random seed for reproducibility. If None, a random seed is chosen. · Default: 81 · Range: 0–4294967295 |
voice1 | string | Optional | The first person's voice to use for speech generation · Allowed: Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill · Default: Sarah |
num_frames | number | Optional | Number of frames to generate (41–241). Frames above 81 are billed at 1.25×. · Default: 191 · Range: 41–241 |
first_text_input | string | Required | The text input to guide video generation. |
resolution | string | Optional | Allowed: 480p, 720p |
image_url | string | Required | Format: uri |
Minimal request example
{
"model_id": "fal-ai/ai-avatar/multi-text",
"second_text_input": "your_second_text_input",
"first_text_input": "your_first_text_input",
"image_url": "https://example.com/image"
}Raw JSON Schema
{
"$schema": "https://json-schema.org/draft/2020-12/schema",
"$id": "https://www.artemotion.ai/api/v1/models?id=fal-ai%2Fai-avatar%2Fmulti-text",
"title": "AI Avatar Multi (Text) generation request",
"description": "Request body accepted by POST /api/v1/generate for fal-ai/ai-avatar/multi-text.",
"type": "object",
"properties": {
"model_id": {
"type": "string",
"const": "fal-ai/ai-avatar/multi-text",
"description": "ArtEmotion model identifier."
},
"extra": {
"type": "object",
"additionalProperties": true,
"description": "Model-specific settings may also be nested here."
},
"max_credits": {
"type": "number",
"minimum": 1,
"description": "Reject before submission if the estimated list price exceeds this cap."
},
"webhook_url": {
"type": "string",
"format": "uri",
"maxLength": 2048
},
"webhook_secret": {
"type": "string",
"maxLength": 512
},
"folder_id": {
"type": "string"
},
"prompt": {
"type": "string"
},
"acceleration": {
"title": "Acceleration",
"description": "The acceleration level to use for generation.",
"default": "regular",
"type": "string",
"enum": [
"none",
"regular",
"high"
]
},
"voice2": {
"title": "Voice2",
"description": "The second person's voice to use for speech generation",
"default": "Roger",
"type": "string",
"enum": [
"Aria",
"Roger",
"Sarah",
"Laura",
"Charlie",
"George",
"Callum",
"River",
"Liam",
"Charlotte",
"Alice",
"Matilda",
"Will",
"Jessica",
"Eric",
"Chris",
"Brian",
"Daniel",
"Lily",
"Bill"
]
},
"second_text_input": {
"title": "Second Text Input",
"description": "The text input to guide video generation.",
"type": "string"
},
"seed": {
"title": "Seed",
"description": "Random seed for reproducibility. If None, a random seed is chosen.",
"default": 81,
"type": "integer",
"minimum": 0,
"maximum": 4294967295
},
"voice1": {
"title": "Voice1",
"description": "The first person's voice to use for speech generation",
"default": "Sarah",
"type": "string",
"enum": [
"Aria",
"Roger",
"Sarah",
"Laura",
"Charlie",
"George",
"Callum",
"River",
"Liam",
"Charlotte",
"Alice",
"Matilda",
"Will",
"Jessica",
"Eric",
"Chris",
"Brian",
"Daniel",
"Lily",
"Bill"
]
},
"num_frames": {
"title": "Num Frames",
"description": "Number of frames to generate (41–241). Frames above 81 are billed at 1.25×.",
"default": 191,
"type": "number",
"minimum": 41,
"maximum": 241,
"multipleOf": 1
},
"first_text_input": {
"title": "First Text Input",
"description": "The text input to guide video generation.",
"type": "string"
},
"resolution": {
"type": "string",
"enum": [
"480p",
"720p"
]
},
"image_url": {
"type": "string",
"format": "uri"
}
},
"required": [
"model_id",
"second_text_input",
"first_text_input",
"image_url"
],
"additionalProperties": false
}FAQ
How much does AI Avatar Multi (Text) cost on ArtEmotion?
Credit-based pricing. You pay in ArtEmotion credits — every plan and top-up converts USD to credits at a fixed rate.
Do I get my credits back if AI Avatar Multi (Text) fails?
Yes — failed generations are never charged. The credits are released back to your balance automatically.
Can I call AI Avatar Multi (Text) from the API?
Yes. Use POST /api/v1/generate with model_id: "fal-ai/ai-avatar/multi-text". See the API reference for the full schema.
Where are my generations stored?
Every output is saved to your personal Library. You can export or delete everything any time from Privacy & deletion.
Ready to generate with AI Avatar Multi (Text)?
Start now →No commitment. New accounts get 50 free credits on signup. See pricing.