Generative video built for VFX production

LTX-2.3 gives VFX teams controllable video generation: pre-trained camera control LoRAs, IC-LoRA video-to-video transforms, and high-fidelity pre-vis at 4K and 50fps. Open-source weights on HuggingFace, at 1/5th to 1/10th the compute cost of leading models.

Try LTX-2.3
//

Key Capabilities

  • Pre-trained camera control LoRAs

    Seven pre-trained camera movement LoRAs available on HuggingFace: dolly in, dolly out, dolly left, dolly right, jib up, jib down, static. Load directly, no training or fine-tuning required.
  • IC-LoRA video-to-video transforms

    Apply trained IC-LoRA adapters for video-to-video transforms: deblurring, colorization, pose and depth conditioning, and style transfer. Pre-trained adapters available: IC-LoRA Union Control
  • 4K/50fps pre-visualization

    Generate pre-vis sequences from text or image prompts at up to 4K resolution and 50fps. Up to 20 seconds per generation on Fast, 10 seconds on Pro. Used to explore shots, sequences, and visu

Shot pre-visualization

VFX supervisors and directors exploring camera angles, shot design, and visual approaches before committing to production. LTX-2.3 generates 4K clips from text or image prompts in fast iteration cycles. Generate multiple creative options in parallel. Traditional pre-vis pipelines require dedicated artists and long turnarounds. This doesn't.

Controlled camera movement in generated sequences

Apply camera control LoRAs to lock specific camera behavior into any generation. Choose from seven pre-trained options: dolly in, dolly out, dolly left, dolly right, jib up, jib down, static. Weights are on HuggingFace. Load directly, no training needed. Works in ComfyUI, via API, and in the open-source pipeline.

Video-to-video style, depth, and motion transforms

Apply IC-LoRA Union Control or IC-LoRA Motion Track Control to transform existing footage: alter visual style, apply depth conditioning, or guide motion. For custom, branded, or proprietary transforms, train a custom IC-LoRA using ltx-trainer on paired reference-target video sequences. Deploy on-prem.

High-volume generation for production pipelines

Run large-volume generation via LTX either on-premise or with per-second pricing. No seat licenses. Integrate directly into your production pipeline. Open-source option for on-premise deployment. At 1/5th to 1/10th the compute cost of leading models, the savings compound fast at production volume.

Designed for real-world deployment

Production-ready AI video model for VFX teams building scalable, controllable video generation workflows.

Builders

Product teams, AI startups, and developers building AI-powered video features. Add production-grade video generation as a product capability, not a research project. One API, production-ready results, and no custom orchestration.

Producers at scale

Brands, agencies, and creative teams producing high volumes of content. Turn existing assets into video at scale. Faster iteration, lower production cost, and more output from what you already have.

On-prem operators

Teams that require full control over deployment and data. Run video generation in your own environment. On-premises, no cloud dependency, and full infrastructure ownership.

Platform teams

Platforms powering creative tools with multiple AI models. Upgrade your video output with a best-in-class engine. Improve generation quality, retain users, and differentiate with a model built for production, not prototypes.

How LTX-2.3 works for VFX

Input

Technical characteristics:

  • Text prompt or reference image (text-to-video or image-to-video)
  • Optional: pre-trained camera control LoRA, load from HuggingFace, no training required
    • Options: dolly in, dolly out, dolly left, dolly right, jib up, jib down, static
  • Optional: IC-LoRA adapter weights
    • Pre-trained: IC-LoRA Union Control, IC-LoRA Motion Track Control
    • Custom: train via ltx-trainer on paired reference-target video sequences
  • Pipeline: TI2VidTwoStagesPipeline / TI2VidTwoStagesHQPipeline (4K output), ICLoraPipeline (IC-LoRA transforms), DistilledPipeline (8-step fastest)
  • API model: ltx-2-fast (up to 20s, lower cost) or ltx-2-pro (up to 10s, highest quality)

Output

Technical characteristics:

  • Video up to 4K resolution, up to 50fps, up to 20 seconds on Fast or 10 seconds on Pro
  • IC-LoRA transforms: transformed video preserving temporal consistency
  • Open-source weights on HuggingFace, export, self-host, or share
  • Available via LTX API (REST) or open-source on 80GB+ VRAM GPU (32GB with FP8 quantization)