Text → video · Image → video · Audio → video
VideoPopularFast

LTX-2.3 22B

Lightricks' audio-driven video model for text, image, audio, image+audio, and video-to-video workflows.

No GPU. No setup. Cancel anytime.

About

LTX-2.3 22B is Lightricks' audio-driven video model: production-ready text-to-video and image-to-video with native dialogue and sound generated in the same pass.

Distilled builds run the current eight-step schedule for fast, cheap generation. Dev builds run the full 20-step schedule for maximum coherence and quality at higher latency.

The family also exposes audio-to-video, image+audio-to-video, and video-to-video model IDs, including ControlNet-assisted V2V for canny, pose, depth, detailer, outpaint, and inpaint workflows.

Variants

LTX-2.3 22B ships in 10 workflow and quality variants. Pick the mode that fits, or start with the default.

Variant Details From
T2V Distilled Fast
ltx23-22b-fp8_t2v_distilled
Workflow: Text-to-video with native dialogue and sound
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 8 default · 4–12
13.8 Spark · $0.069 Create →
I2V Distilled Fast
ltx23-22b-fp8_i2v_distilled
Workflow: Image-to-video from start, end, or start+end frames
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 8 default · 4–12
13.8 Spark · $0.069 Create →
A2V Distilled Fast
ltx23-22b-fp8_a2v_distilled
Workflow: Audio-to-video from prompt plus uploaded audio
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 8 default · 4–12
13.8 Spark · $0.069 Create →
IA2V Distilled Fast
ltx23-22b-fp8_ia2v_distilled
Workflow: Image+audio-to-video for audio-reactive animation
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 8 default · 4–12
13.8 Spark · $0.069 Create →
V2V Distilled Fast
ltx23-22b-fp8_v2v_distilled
Workflow: Video-to-video with ControlNet, outpaint, and inpaint modes
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (25 default)
Audio: native dialogue & sound
Steps: 8 default · 4–12
13.8 Spark · $0.069 Create →
T2V Dev Quality
ltx23-22b-fp8_t2v_dev
Workflow: Text-to-video with native dialogue and sound
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 30 default · 15–50
34.6 Spark · $0.17 Create →
I2V Dev Quality
ltx23-22b-fp8_i2v_dev
Workflow: Image-to-video from start, end, or start+end frames
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 30 default · 15–50
34.6 Spark · $0.17 Create →
A2V Dev Quality
ltx23-22b-fp8_a2v_dev
Workflow: Audio-to-video from prompt plus uploaded audio
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 30 default · 15–50
34.6 Spark · $0.17 Create →
IA2V Dev Quality
ltx23-22b-fp8_ia2v_dev
Workflow: Image+audio-to-video for audio-reactive animation
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (24 default)
Audio: native dialogue & sound
Steps: 30 default · 15–50
34.6 Spark · $0.17 Create →
V2V Dev Quality
ltx23-22b-fp8_v2v_dev
Workflow: Video-to-video with ControlNet, outpaint, and inpaint modes
Duration: 4 – 20 s
Resolution: 640–3840 px · 1–60 fps (25 default)
Audio: native dialogue & sound
Steps: 30 default · 10–50
34.6 Spark · $0.17 Create →

Prompting tips

LTX-2.3 is audio-driven and cinematic — the strongest prompts describe four things. Use the prompt field’s wand button to auto-expand a short idea into this shape.

  • What — the subject, action, and setting.
  • Feel — style, mood, and genre.
  • Camera — lens, angle, and how it moves.
  • Time — motion type, speed, and frame-rate intent.

LoRAs

LTX-2.3 has grown a real LoRA ecosystem in Sogni. TalkVid-3K powers voice identity for Sogni Chat Personas and is available in Sogni Web Voice Transfer. The Lightricks control LoRAs and community LoRAs are integrated in Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent, with more Sogni apps coming soon.

LoRA What it does Where it is integrated
Union Control
Lightricks IC-LoRA
Canny edge, depth-map, and pose-skeleton control in a single adapter — steer structure, camera moves, and body motion from a reference video. Integrated in Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent for video → video canny, depth, pose, and detailer control workflows.
In/Outpainting
Lightricks IC-LoRA
Masked video editing — remove or replace anything in the shot, or extend the frame beyond its original borders. Integrated in Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent for video → video inpaint and outpaint workflows.
TalkVid-3K
ID-LoRA · research
Identity-preserving speech — a reference image and voice clip drive likeness and vocal identity together in a single generative pass. Powers the Personas feature in Sogni Chat and the Voice Transfer controls in Sogni Web.
Transition
ValiantCat · community
First-frame → last-frame morph transitions with unusually good motion continuity. Trigger word: zhuanchang. Integrated into Sogni 360 and Sogni Photobooth to automagically enhance video outputs. Also available through Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent.

The LoRA catalog is integrated server-side by workflow and app surface; it is not silently attached to ordinary LTX renders. Pick the Sogni surface that exposes the workflow you need, and Sogni handles the adapter details behind it.

Pricing

Use pay-as-you-go Spark packs for each clip (1 Spark = $0.005), or choose a flat-rate Sogni plan for credit-free fair-use generation in the app.

Variant / configuration Spark USD
Distilled variants · 5 s · 1280 × 720 · 8 steps 13.8 Spark $0.069
Distilled variants · 10 s · 1280 × 720 · 8 steps 27.6 Spark $0.14
Distilled variants · 5 s · 1920 × 1088 · 8 steps (scales with pixels) 31.4 Spark $0.16
Dev variants · 5 s · 1280 × 720 · 20 steps 34.6 Spark $0.17
Dev variants · 10 s · 1280 × 720 · 20 steps 69.2 Spark $0.35

API

One Sogni API key reaches every model on the Supernet — call LTX-2.3 22B with the exact model id for the variant you want.

import { SogniClient } from '@sogni-ai/sogni-client';

const client = await SogniClient.createInstance({
  appId: crypto.randomUUID(),
  apiKey: process.env.SOGNI_API_KEY,
  network: 'fast',
});

const project = await client.projects.create({
  type: 'video',
  modelId: 'ltx23-22b-fp8_t2v_distilled',
  positivePrompt: 'a slothicorn surfing a wave of liquid paint, slow push-in, cinematic',
  numberOfMedia: 1,
  duration: 5,
});

const [url] = await project.waitForCompletion();
console.log(url); // result link — download within 24h
curl https://api.sogni.ai/v1/creative-agent/workflows \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $SOGNI_API_KEY" \
  -d '{
    "input": {
      "title": "First render",
      "steps": [{
        "id": "step1",
        "toolName": "generate_video",
        "arguments": { "prompt": "a slothicorn surfing a wave of liquid paint, slow push-in, cinematic", "videoModel": "ltx23", "duration": 5 }
      }]
    },
    "confirm_cost": true
  }'

Example uses ltx23-22b-fp8_t2v_distilled. Use any exact model id from the Variants table above. Full reference at docs.sogni.ai.

Why run it on Sogni

Subscriptions or Spark

Use a flat monthly plan for credit-free fair-use generation, or buy Spark packs when pay-as-you-go fits better. Both run on the same creator-owned GPU network.

Unlimited plans

One flat price in the app. Generate under fair use without a per-image meter.

🧩

100+ models

Image, video, music, and language models in one workspace and one API key.

Pay-as-you-go Spark

Prefer pay-as-you-go? Call LTX-2.3 22B by id and pay with Spark packs.

🌐

Powered by people

Runs on a decentralized GPU network where workers share subscription revenue.

FAQ

LTX-2.3 22B on Sogni

Which LTX-2.3 variant should I use?

Start with the Distilled workflow that matches your input type: T2V, I2V, A2V, IA2V, or V2V. Distilled is the current eight-step default. Use the matching Dev model when you need maximum coherence and can spend the full 20-step schedule.

Start with LTX-2.3 22B today

Create in the app, or build with the API. Your call.