[

LTX Trainer

]

Introducing the new LTX Trainer

Train LTX on your characters, styles, workflows, and IP. Build models that learn what matters to your business. Designed for high-control generative pipelines.

LTX-2.3 Examples Video

Turn workflows into capabilities

Most AI video tools generate content. LTX Trainer lets you fine tune the model itself.

Train LoRAs and IC-LoRAs on your own data to create reusable capabilities that persist across projects. Instead of rebuilding the same workflow again and again, teach the model once and reuse that knowledge everywhere.

//

What's new

  • Unified conditioning

    Train across video, audio, cross-modal, and reference-conditioned workflows from a single framework.
  • Composable conditioning

    Mix references, masks, first frames, audio, video, prefixes, and suffixes to compose custom training workflows.
  • Agentic setup

    Use Claude or other LLMs to configure your LoRA and IC-LoRA training runs. You bring the dataset. The agent handles the config.

Built for Professionals

Built for professionals

Custom models

Train on the data that matters to your business. Create capabilities tailored to your products, characters, styles, workflows, and production requirements.

Consistent outputs

Build persistent visual identities instead of relying on one-off generations. Maintain consistency across campaigns, productions, products, and teams.

LTX-2.3 Better Prompt Adherence

Scalable production

Turn successful creative workflows into reusable model capabilities. Reduce repetitive work and scale production more efficiently.

LTX-2.3 Better Image To Video

New Training Capabilities

Video

  • Resolutions: 1080p, 1440p, 4K
  • FPS: 24 /25 / 48 / 50
  • Duration: up to 20 seconds
  • Lower compute load and faster render times

Audio

  • Text-to-Audio (T2A)
  • Audio Extension
  • Audio Inpainting

Cross-Modal

  • Audio-to-Video (A2V)
  • Video-to-Audio (V2A)

Reference Conditioning (IC-LoRAs)

  • Video-to-Video (V2V)
  • Audio-to-Audio (A2A)
  • Audio+Video-to-Audio+Video (AV2AV)

Built for Professionals

New training capabilities

Video

  • Resolutions: 1080p, 1440p, 4K
  • FPS: 24 /25 / 48 / 50
  • Duration: up to 20 seconds
  • Lower compute load and faster render times

Audio

  • Text-to-Audio (T2A)
  • Audio Extension
  • Audio Inpainting

Cross-Modal

  • Audio-to-Video (A2V)
  • Video-to-Audio (V2A)

Reference Conditioning (IC-LoRAs)

  • Video-to-Video (V2V)
  • Audio-to-Audio (A2A)
  • Audio+Video-to-Audio+Video (AV2AV)

Designed for real-world deployment

Production-ready LoRA training for enterprise teams building scalable, controllable video generation workflows.

Builders

Product teams, AI startups, and developers building AI-powered video features. Add production-grade video generation as a product capability, not a research project. One API, production-ready results, and no custom orchestration.

Producers at scale

Brands, agencies, and creative teams producing high volumes of content. Turn existing assets into video at scale. Faster iteration, lower production cost, and more output from what you already have.

On-prem operators

Teams that require full control over deployment and data. Run video generation in your own environment. On-premises, no cloud dependency, and full infrastructure ownership.

Platform teams

Platforms powering creative tools with multiple AI models. Upgrade your video output with a best-in-class engine. Improve generation quality, retain users, and differentiate with a model built for production, not prototypes.