MODELS · VIDEO · LIGHTRICKS

LTX-2.3.

Multimodal video generation with native synchronized audio, text, image, or audio-driven input, and resolution up to 4K.

Text to videoImage to videoAudio to video

Free to start · 100+ models on one account · cancel anytime

About LTX-2.3

LTX-2.3 is Lightricks' multimodal video generation model, producing synchronized video and audio in a single pass rather than layering sound on afterward. It runs in three modes: type a text prompt to generate video from scratch, upload a starting image to animate it into motion, or drive the output directly from an audio clip so the visuals sync to the sound, with clip length following the audio. Text and image mode clips run 6 to 10 seconds, generated at resolutions from 1080p up to full 4K, in either landscape or portrait orientation. Native audio generation is on by default in the text and image modes, so dialogue and ambient sound arrive built into the render — no separate audio pass needed. It's built for production-ready output: temporal stability and motion coherence hold up across longer clips and higher resolutions, suited to creative and marketing pipelines that need polished results fast.

  • Text to video
  • Image to video
  • Audio to video

How to use LTX-2.3 on getvivix

STEP 1
Sign in & open Studio

Create a free getvivix account — no card required.

STEP 2
Pick LTX-2.3

Choose LTX-2.3 from the model list and set your options.

STEP 3
Generate & download

Enter your prompt or upload your input, hit generate, then download in full quality.

LTX-2.3 — frequently asked

What is LTX-2.3?

LTX-2.3 is one of 100+ AI models available on getvivix. LTX-2.3 is Lightricks' multimodal video generation model, producing synchronized video and audio in a single pass rather than layering sound on afterward. It runs in three modes: type a text prompt to generate video from scratch, upload a starting image to animate it into motion, or drive the output directly from an audio clip so the visuals sync to the sound, with clip length following the audio. Text and image mode clips run 6 to 10 seconds, generated at resolutions from 1080p up to full 4K, in either landscape or portrait orientation. Native audio generation is on by default in the text and image modes, so dialogue and ambient sound arrive built into the render — no separate audio pass needed. It's built for production-ready output: temporal stability and motion coherence hold up across longer clips and higher resolutions, suited to creative and marketing pipelines that need polished results fast.

How do I use LTX-2.3 on getvivix?

Sign in to getvivix and open the Studio, pick LTX-2.3 from the model list, enter your prompt (or upload your input), and generate — then download the result in full quality.

Is LTX-2.3 free to try?

Yes — getvivix has a free tier, so you can try LTX-2.3 without a card. Sign up and start generating right away, alongside 100+ other AI models on one account.

What can LTX-2.3 do?

LTX-2.3 supports text to video, image to video, audio to video. It runs on getvivix alongside 100+ other frontier AI models, all from one account.

More video models

Browse all 100+ models