Skip to main content
ArtEmotion
Video

LTX-2.3 22B (T2V)

Credit-based pricing. Supports audio generation.

Output
video
Aspect ratio
16:9, 9:16, 1:1
Resolution
1080p, 1440p, 2160p
Duration
3s, 4s, 5s, 6s, 7s, 8s, 10s

Inputs

  • Generates audio

Example prompt

Aurora borealis rippling in vivid curtains of emerald and violet over a frozen Nordic lake at midnight, ice crystals glinting, lone pine silhouettes against the dancing sky
Try this prompt →
LLM-ready

API & LLM schema

Exact request contract for this model. Agents can fetch it from /api/v1/models?id=fal-ai/ltx-2.3-22b/text-to-video.

POST/api/v1/generate38 fields · 2 required
FieldTypeRequirementContract
model_idconstantRequiredArtEmotion model identifier.
extraobjectOptionalModel-specific settings may also be nested here.
max_creditsnumberOptionalReject before submission if the estimated list price exceeds this cap. · Range: 1–…
webhook_urlstringOptionalFormat: uri
webhook_secretstringOptionalOptional model input.
folder_idstringOptionalOptional model input.
promptstringRequiredOptional model input.
generate_audiobooleanOptionalWhether to generate audio for the video. · Default: true
audio_stg_scalenumberOptionalThe Spatiotemporal Guidance (STG) scale for the audio. Higher values result in more consistent and focused audio content. · Default: 0 · Range: 0–20
seedintegerOptionalThe seed for the random number generator.
use_multiscalebooleanOptionalWhether to use multi-scale generation. If True, the model will generate the video at a smaller scale first, then use the smaller video to guide the ge · Default: true
audio_modality_scalenumberOptionalThe modality scale for the audio. Controls the ratio between video and audio modalities. · Default: 3 · Range: 0–10
use_restart_samplingbooleanOptionalWhether to use restart sampling. This will inject a small amount of noise during each denoising step, which can help improve the quality of the genera · Default: false
schedulerstringOptionalThe scheduler to use. · Allowed: ltx2, linear_quadratic, beta · Default: ltx2
distill_lora_first_pass_scalenumberOptionalThe scale of the distill LoRA to use for the first pass. Set to 0 to disable. · Default: 0.2 · Range: 0–1
gradient_estimation_gammanumberOptionalThe gamma of gradient estimation during denoising. Set to 0 to disable. · Default: 2 · Range: 0–10
distill_lora_second_pass_scalenumberOptionalThe scale of the distill LoRA to use for the second and subsequent passes. · Default: 0.5 · Range: 0–1
num_framesnumberOptionalThe number of frames to generate. · Default: 121 · Range: 9–481
accelerationstringOptionalThe acceleration level to use. · Allowed: none, regular, high, full · Default: regular
video_output_typestringOptionalThe output type of the generated video. · Allowed: X264 (.mp4), VP9 (.webm), PRORES4444 (.mov), GIF (.gif) · Default: X264 (.mp4)
video_rescaling_scalenumberOptionalThe rescaling scale for the video. Controls the ratio between classifier-free guidance and spatiotemporal guidance. · Default: 0.7 · Range: 0–1
camera_lora_scalenumberOptionalThe scale of the camera LoRA to use. This allows you to control the camera movement of the generated video more accurately than just prompting the mod · Default: 1 · Range: 0–1
camera_lorastringOptionalThe camera LoRA to use. This allows you to control the camera movement of the generated video more accurately than just prompting the model to move th · Allowed: dolly_in, dolly_out, dolly_left, dolly_right, jib_up, jib_down, static, none · Default: none
audio_rescaling_scalenumberOptionalThe rescaling scale for the audio. Controls the ratio between classifier-free guidance and spatiotemporal guidance. · Default: 0.7 · Range: 0–1
video_modality_scalenumberOptionalThe modality scale for the video. Controls the ratio between video and audio modalities. · Default: 3 · Range: 0–10
video_qualitystringOptionalThe quality of the generated video. · Allowed: low, medium, high, maximum · Default: high
video_write_modestringOptionalThe write mode of the generated video. · Allowed: fast, balanced, small · Default: balanced
video_stg_scalenumberOptionalThe Spatiotemporal Guidance (STG) scale for the video. Higher values result in more consistent and focused video content. · Default: 0 · Range: 0–20
video_cfg_scalenumberOptionalThe Classifier-Free Guidance (CFG) scale for the video. Higher values result in more consistent and focused video content. · Default: 3 · Range: 1–20
num_inference_stepsnumberOptionalThe number of inference steps to use. · Default: 40 · Range: 8–50
negative_promptstringOptionalThe negative prompt to generate the video from. · Default: news broadcast, 3d animation, computer graphics, pc game, console game, video game, cartoon, childish, watermark, logo, text, on screen text, subtitles, titles, signature, slowmo, static
enable_prompt_expansionbooleanOptionalWhether to enable prompt expansion. · Default: true
enable_safety_checkerbooleanOptionalRun the safety checker to block unsafe content. · Default: true
fpsnumberOptionalThe frames per second of the generated video. · Default: 24 · Range: 1–60
audio_cfg_scalenumberOptionalThe Classifier-Free Guidance (CFG) scale for the audio. Higher values result in more consistent and focused audio content. · Default: 7 · Range: 1–20
aspect_ratiostringOptionalAllowed: 16:9, 9:16, 1:1
resolutionstringOptionalAllowed: 1080p, 1440p, 2160p
durationnumberOptionalAllowed: 3, 4, 5, 6, 7, 8, 10
Minimal request example
{
  "model_id": "fal-ai/ltx-2.3-22b/text-to-video",
  "prompt": "Aurora borealis rippling in vivid curtains of emerald and violet over a frozen Nordic lake at midnight, ice crystals glinting, lone pine silhouettes against the dancing sky"
}
Raw JSON Schema
{
  "$schema": "https://json-schema.org/draft/2020-12/schema",
  "$id": "https://www.artemotion.ai/api/v1/models?id=fal-ai%2Fltx-2.3-22b%2Ftext-to-video",
  "title": "LTX-2.3 22B (T2V) generation request",
  "description": "Request body accepted by POST /api/v1/generate for fal-ai/ltx-2.3-22b/text-to-video.",
  "type": "object",
  "properties": {
    "model_id": {
      "type": "string",
      "const": "fal-ai/ltx-2.3-22b/text-to-video",
      "description": "ArtEmotion model identifier."
    },
    "extra": {
      "type": "object",
      "additionalProperties": true,
      "description": "Model-specific settings may also be nested here."
    },
    "max_credits": {
      "type": "number",
      "minimum": 1,
      "description": "Reject before submission if the estimated list price exceeds this cap."
    },
    "webhook_url": {
      "type": "string",
      "format": "uri",
      "maxLength": 2048
    },
    "webhook_secret": {
      "type": "string",
      "maxLength": 512
    },
    "folder_id": {
      "type": "string"
    },
    "prompt": {
      "type": "string"
    },
    "generate_audio": {
      "title": "Generate Audio",
      "description": "Whether to generate audio for the video.",
      "default": true,
      "type": "boolean"
    },
    "audio_stg_scale": {
      "title": "Audio Stg Scale",
      "description": "The Spatiotemporal Guidance (STG) scale for the audio. Higher values result in more consistent and focused audio content.",
      "default": 0,
      "type": "number",
      "minimum": 0,
      "maximum": 20
    },
    "seed": {
      "title": "Seed",
      "description": "The seed for the random number generator.",
      "type": "integer"
    },
    "use_multiscale": {
      "title": "Use Multiscale",
      "description": "Whether to use multi-scale generation. If True, the model will generate the video at a smaller scale first, then use the smaller video to guide the ge",
      "default": true,
      "type": "boolean"
    },
    "audio_modality_scale": {
      "title": "Audio Modality Scale",
      "description": "The modality scale for the audio. Controls the ratio between video and audio modalities.",
      "default": 3,
      "type": "number",
      "minimum": 0,
      "maximum": 10
    },
    "use_restart_sampling": {
      "title": "Use Restart Sampling",
      "description": "Whether to use restart sampling. This will inject a small amount of noise during each denoising step, which can help improve the quality of the genera",
      "default": false,
      "type": "boolean"
    },
    "scheduler": {
      "title": "Scheduler",
      "description": "The scheduler to use.",
      "default": "ltx2",
      "type": "string",
      "enum": [
        "ltx2",
        "linear_quadratic",
        "beta"
      ]
    },
    "distill_lora_first_pass_scale": {
      "title": "Distill Lora First Pass Scale",
      "description": "The scale of the distill LoRA to use for the first pass. Set to 0 to disable.",
      "default": 0.2,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "gradient_estimation_gamma": {
      "title": "Gradient Estimation Gamma",
      "description": "The gamma of gradient estimation during denoising. Set to 0 to disable.",
      "default": 2,
      "type": "number",
      "minimum": 0,
      "maximum": 10
    },
    "distill_lora_second_pass_scale": {
      "title": "Distill Lora Second Pass Scale",
      "description": "The scale of the distill LoRA to use for the second and subsequent passes.",
      "default": 0.5,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "num_frames": {
      "title": "Num Frames",
      "description": "The number of frames to generate.",
      "default": 121,
      "type": "number",
      "minimum": 9,
      "maximum": 481,
      "multipleOf": 1
    },
    "acceleration": {
      "title": "Acceleration",
      "description": "The acceleration level to use.",
      "default": "regular",
      "type": "string",
      "enum": [
        "none",
        "regular",
        "high",
        "full"
      ]
    },
    "video_output_type": {
      "title": "Video Output Type",
      "description": "The output type of the generated video.",
      "default": "X264 (.mp4)",
      "type": "string",
      "enum": [
        "X264 (.mp4)",
        "VP9 (.webm)",
        "PRORES4444 (.mov)",
        "GIF (.gif)"
      ]
    },
    "video_rescaling_scale": {
      "title": "Video Rescaling Scale",
      "description": "The rescaling scale for the video. Controls the ratio between classifier-free guidance and spatiotemporal guidance.",
      "default": 0.7,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "camera_lora_scale": {
      "title": "Camera LoRA Scale",
      "description": "The scale of the camera LoRA to use. This allows you to control the camera movement of the generated video more accurately than just prompting the mod",
      "default": 1,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "camera_lora": {
      "title": "Camera LoRA",
      "description": "The camera LoRA to use. This allows you to control the camera movement of the generated video more accurately than just prompting the model to move th",
      "default": "none",
      "type": "string",
      "enum": [
        "dolly_in",
        "dolly_out",
        "dolly_left",
        "dolly_right",
        "jib_up",
        "jib_down",
        "static",
        "none"
      ]
    },
    "audio_rescaling_scale": {
      "title": "Audio Rescaling Scale",
      "description": "The rescaling scale for the audio. Controls the ratio between classifier-free guidance and spatiotemporal guidance.",
      "default": 0.7,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "video_modality_scale": {
      "title": "Video Modality Scale",
      "description": "The modality scale for the video. Controls the ratio between video and audio modalities.",
      "default": 3,
      "type": "number",
      "minimum": 0,
      "maximum": 10
    },
    "video_quality": {
      "title": "Video Quality",
      "description": "The quality of the generated video.",
      "default": "high",
      "type": "string",
      "enum": [
        "low",
        "medium",
        "high",
        "maximum"
      ]
    },
    "video_write_mode": {
      "title": "Video Write Mode",
      "description": "The write mode of the generated video.",
      "default": "balanced",
      "type": "string",
      "enum": [
        "fast",
        "balanced",
        "small"
      ]
    },
    "video_stg_scale": {
      "title": "Video Stg Scale",
      "description": "The Spatiotemporal Guidance (STG) scale for the video. Higher values result in more consistent and focused video content.",
      "default": 0,
      "type": "number",
      "minimum": 0,
      "maximum": 20
    },
    "video_cfg_scale": {
      "title": "Video Cfg Scale",
      "description": "The Classifier-Free Guidance (CFG) scale for the video. Higher values result in more consistent and focused video content.",
      "default": 3,
      "type": "number",
      "minimum": 1,
      "maximum": 20
    },
    "num_inference_steps": {
      "title": "Num Inference Steps",
      "description": "The number of inference steps to use.",
      "default": 40,
      "type": "number",
      "minimum": 8,
      "maximum": 50,
      "multipleOf": 1
    },
    "negative_prompt": {
      "title": "Negative Prompt",
      "description": "The negative prompt to generate the video from.",
      "default": "news broadcast, 3d animation, computer graphics, pc game, console game, video game, cartoon, childish, watermark, logo, text, on screen text, subtitles, titles, signature, slowmo, static",
      "type": "string"
    },
    "enable_prompt_expansion": {
      "title": "Enable Prompt Expansion",
      "description": "Whether to enable prompt expansion.",
      "default": true,
      "type": "boolean"
    },
    "enable_safety_checker": {
      "title": "Enable Safety Checker",
      "description": "Run the safety checker to block unsafe content.",
      "default": true,
      "type": "boolean"
    },
    "fps": {
      "title": "Fps",
      "description": "The frames per second of the generated video.",
      "default": 24,
      "type": "number",
      "minimum": 1,
      "maximum": 60
    },
    "audio_cfg_scale": {
      "title": "Audio Cfg Scale",
      "description": "The Classifier-Free Guidance (CFG) scale for the audio. Higher values result in more consistent and focused audio content.",
      "default": 7,
      "type": "number",
      "minimum": 1,
      "maximum": 20
    },
    "aspect_ratio": {
      "type": "string",
      "enum": [
        "16:9",
        "9:16",
        "1:1"
      ]
    },
    "resolution": {
      "type": "string",
      "enum": [
        "1080p",
        "1440p",
        "2160p"
      ]
    },
    "duration": {
      "type": "number",
      "enum": [
        3,
        4,
        5,
        6,
        7,
        8,
        10
      ]
    }
  },
  "required": [
    "model_id",
    "prompt"
  ],
  "additionalProperties": false
}

FAQ

How much does LTX-2.3 22B (T2V) cost on ArtEmotion?

Credit-based pricing. You pay in ArtEmotion credits — every plan and top-up converts USD to credits at a fixed rate.

Do I get my credits back if LTX-2.3 22B (T2V) fails?

Yes — failed generations are never charged. The credits are released back to your balance automatically.

Can I call LTX-2.3 22B (T2V) from the API?

Yes. Use POST /api/v1/generate with model_id: "fal-ai/ltx-2.3-22b/text-to-video". See the API reference for the full schema.

Where are my generations stored?

Every output is saved to your personal Library. You can export or delete everything any time from Privacy & deletion.

Ready to generate with LTX-2.3 22B (T2V)?

Start now →

No commitment. New accounts get 50 free credits on signup. See pricing.