Lightricks' audio-driven video model for text, image, audio, image+audio, and video-to-video workflows.
No GPU. No setup. Cancel anytime.
LTX-2.3 22B is Lightricks' audio-driven video model: production-ready text-to-video and image-to-video with native dialogue and sound generated in the same pass.
Distilled builds run the current eight-step schedule for fast, cheap generation. Dev builds run the full 20-step schedule for maximum coherence and quality at higher latency.
The family also exposes audio-to-video, image+audio-to-video, and video-to-video model IDs, including ControlNet-assisted V2V for canny, pose, depth, detailer, outpaint, and inpaint workflows.
LTX-2.3 22B ships in 10 workflow and quality variants. Pick the mode that fits, or start with the default.
| Variant | Details | From | |
|---|---|---|---|
T2V Distilled Fastltx23-22b-fp8_t2v_distilled | Workflow: Text-to-video with native dialogue and sound Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 8 default · 4–12 | 13.8 Spark · $0.069 | Create → |
I2V Distilled Fastltx23-22b-fp8_i2v_distilled | Workflow: Image-to-video from start, end, or start+end frames Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 8 default · 4–12 | 13.8 Spark · $0.069 | Create → |
A2V Distilled Fastltx23-22b-fp8_a2v_distilled | Workflow: Audio-to-video from prompt plus uploaded audio Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 8 default · 4–12 | 13.8 Spark · $0.069 | Create → |
IA2V Distilled Fastltx23-22b-fp8_ia2v_distilled | Workflow: Image+audio-to-video for audio-reactive animation Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 8 default · 4–12 | 13.8 Spark · $0.069 | Create → |
V2V Distilled Fastltx23-22b-fp8_v2v_distilled | Workflow: Video-to-video with ControlNet, outpaint, and inpaint modes Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (25 default) Audio: native dialogue & sound Steps: 8 default · 4–12 | 13.8 Spark · $0.069 | Create → |
T2V Dev Qualityltx23-22b-fp8_t2v_dev | Workflow: Text-to-video with native dialogue and sound Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 30 default · 15–50 | 34.6 Spark · $0.17 | Create → |
I2V Dev Qualityltx23-22b-fp8_i2v_dev | Workflow: Image-to-video from start, end, or start+end frames Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 30 default · 15–50 | 34.6 Spark · $0.17 | Create → |
A2V Dev Qualityltx23-22b-fp8_a2v_dev | Workflow: Audio-to-video from prompt plus uploaded audio Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 30 default · 15–50 | 34.6 Spark · $0.17 | Create → |
IA2V Dev Qualityltx23-22b-fp8_ia2v_dev | Workflow: Image+audio-to-video for audio-reactive animation Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (24 default) Audio: native dialogue & sound Steps: 30 default · 15–50 | 34.6 Spark · $0.17 | Create → |
V2V Dev Qualityltx23-22b-fp8_v2v_dev | Workflow: Video-to-video with ControlNet, outpaint, and inpaint modes Duration: 4 – 20 s Resolution: 640–3840 px · 1–60 fps (25 default) Audio: native dialogue & sound Steps: 30 default · 10–50 | 34.6 Spark · $0.17 | Create → |
LTX-2.3 is audio-driven and cinematic — the strongest prompts describe four things. Use the prompt field’s wand button to auto-expand a short idea into this shape.
LTX-2.3 has grown a real LoRA ecosystem in Sogni. TalkVid-3K powers voice identity for Sogni Chat Personas and is available in Sogni Web Voice Transfer. The Lightricks control LoRAs and community LoRAs are integrated in Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent, with more Sogni apps coming soon.
| LoRA | What it does | Where it is integrated |
|---|---|---|
| Union Control Lightricks IC-LoRA | Canny edge, depth-map, and pose-skeleton control in a single adapter — steer structure, camera moves, and body motion from a reference video. | Integrated in Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent for video → video canny, depth, pose, and detailer control workflows. |
| In/Outpainting Lightricks IC-LoRA | Masked video editing — remove or replace anything in the shot, or extend the frame beyond its original borders. | Integrated in Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent for video → video inpaint and outpaint workflows. |
| TalkVid-3K ID-LoRA · research | Identity-preserving speech — a reference image and voice clip drive likeness and vocal identity together in a single generative pass. | Powers the Personas feature in Sogni Chat and the Voice Transfer controls in Sogni Web. |
| Transition ValiantCat · community | First-frame → last-frame morph transitions with unusually good motion continuity. Trigger word: zhuanchang. | Integrated into Sogni 360 and Sogni Photobooth to automagically enhance video outputs. Also available through Sogni Chat, the Sogni Client SDK, and Sogni Creative Agent. |
The LoRA catalog is integrated server-side by workflow and app surface; it is not silently attached to ordinary LTX renders. Pick the Sogni surface that exposes the workflow you need, and Sogni handles the adapter details behind it.
Use pay-as-you-go Spark packs for each clip (1 Spark = $0.005), or choose a flat-rate Sogni plan for credit-free fair-use generation in the app.
| Variant / configuration | Spark | USD |
|---|---|---|
| Distilled variants · 5 s · 1280 × 720 · 8 steps | 13.8 Spark | $0.069 |
| Distilled variants · 10 s · 1280 × 720 · 8 steps | 27.6 Spark | $0.14 |
| Distilled variants · 5 s · 1920 × 1088 · 8 steps (scales with pixels) | 31.4 Spark | $0.16 |
| Dev variants · 5 s · 1280 × 720 · 20 steps | 34.6 Spark | $0.17 |
| Dev variants · 10 s · 1280 × 720 · 20 steps | 69.2 Spark | $0.35 |
One Sogni API key reaches every model on the Supernet — call LTX-2.3 22B with the exact model id for the variant you want.
import { SogniClient } from '@sogni-ai/sogni-client';
const client = await SogniClient.createInstance({
appId: crypto.randomUUID(),
apiKey: process.env.SOGNI_API_KEY,
network: 'fast',
});
const project = await client.projects.create({
type: 'video',
modelId: 'ltx23-22b-fp8_t2v_distilled',
positivePrompt: 'a slothicorn surfing a wave of liquid paint, slow push-in, cinematic',
numberOfMedia: 1,
duration: 5,
});
const [url] = await project.waitForCompletion();
console.log(url); // result link — download within 24h curl https://api.sogni.ai/v1/creative-agent/workflows \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $SOGNI_API_KEY" \
-d '{
"input": {
"title": "First render",
"steps": [{
"id": "step1",
"toolName": "generate_video",
"arguments": { "prompt": "a slothicorn surfing a wave of liquid paint, slow push-in, cinematic", "videoModel": "ltx23", "duration": 5 }
}]
},
"confirm_cost": true
}' Example uses ltx23-22b-fp8_t2v_distilled. Use any exact model id from the Variants table above. Full reference at docs.sogni.ai.
Real generations from Sogni. Hit Use prompt to open the app with LTX-2.3 22B selected and the full prompt preloaded.
Use a flat monthly plan for credit-free fair-use generation, or buy Spark packs when pay-as-you-go fits better. Both run on the same creator-owned GPU network.
One flat price in the app. Generate under fair use without a per-image meter.
Image, video, music, and language models in one workspace and one API key.
Prefer pay-as-you-go? Call LTX-2.3 22B by id and pay with Spark packs.
Runs on a decentralized GPU network where workers share subscription revenue.
Start with the Distilled workflow that matches your input type: T2V, I2V, A2V, IA2V, or V2V. Distilled is the current eight-step default. Use the matching Dev model when you need maximum coherence and can spend the full 20-step schedule.
Create in the app, or build with the API. Your call.