- LTX Studio offers five image models (Nano Banana Pro, Nano Banana 2, FLUX.2 Pro, Z-Image, ChatGPT Images 2.0) and seven video models (LTX-2.3 Pro, LTX-2 Pro, Kling 2.6 Pro, Kling 3.0 Pro, Seedance 2.0, Veo 2, Veo 3.1)
- The right model depends on your output type, creative goal, and how much control you need
- Image models differ in speed, precision, and stylistic range — video models differ in realism, audio, and generation length
- Most workflows benefit from pairing models: generate an image first, then bring it to life with video
Pick the wrong model in LTX Studio and a five-minute shot becomes a two-hour rerun. Pick the right one and it lands on the first or second pass. LTX Studio integrates twelve production-grade image and video models in one workspace, and the difference between the fastest and slowest teams is knowing which one to reach for. This guide answers that question, model by model.
What Models Are Available on LTX Studio?
LTX Studio's model lineup covers both image and video generation, giving creators access to leading third-party models alongside Lightricks' own LTX-2.3 — all inside a single creative workspace.
Video models:
- LTX-2.3 Pro — Lightricks' own open-source model, delivering native portrait video, synchronized audio, and tight LTX Studio integration
- LTX-2 Pro — Previous-generation Lightricks open-source model, available in a pro quality tier
- Kling 2.6 Pro / Kling 3.0 Pro — Kuaishou's cinematic video generation models, with Kling 3.0 Pro adding multi-shot sequences up to 15 seconds
- Seedance 2.0 — built for natural human motion and dynamic scenes with minimal subject drift
- Veo 3.1 — Google's flagship video model, offering dual keyframe control, native synchronized audio, and exceptional visual realism
- Veo 2 — Google's previous-generation video model, delivering strong realism and visual quality
Image models:
- Nano Banana Pro — Pro-tier Gemini image model for final production assets
- Nano Banana 2 — Gemini Flash image model combining speed with strong subject consistency
- FLUX.2 Pro — Black Forest Labs' high-resolution diffusion model with HEX code color matching
- Z-Image — Alibaba's Tongyi Lab speed-optimized model for photorealistic output on all tiers
- ChatGPT Images 2.0 — OpenAI's GPT-4o-based model, strongest for text inside images and complex layouts
Which Video Model Should You Use?
Video model selection in LTX Studio comes down to four questions. Do you need synchronized audio? How long does the clip need to be? Is the shot driven by human performance or by camera and environment? And how many iterations do you expect before locking the final?
LTX-2.3 — Best for Fast Iteration, Portrait Video, and Open Workflow
LTX-2.3 is Lightricks' latest open-source video model and the default for fast iteration inside LTX Studio. As a 22B-parameter multimodal model, it generates video from text, image, audio, or video inputs and supports native portrait video — which matters for TikTok, Reels, and Shorts work where cropping a landscape generation loses composition. LTX-2.3 Pro is the higher-quality tier of the same model, tuned for final delivery rather than draft passes.
LTX-2.3 integrates tightly with LTX Studio's storyboard generator, Elements system, and timeline editor, making it the natural starting point for most productions. For AI image-to-video workflows, generate a still with Nano Banana 2 or FLUX.2, then animate it directly within the same session.
Use LTX-2.3 when:
- You need fast generation for rapid iteration and concepting
- Your content is portrait-format or mobile-first
- You want the tightest native integration with LTX Studio's storyboarding and Elements features
- You're prototyping or developing a high volume of draft scenes quickly
- You want an open-source model you can run locally or fine-tune
Kling 3.0 Pro — Best for Multi-Shot Cinematic Storytelling
Kling 3.0 Pro is Kuaishou's newest video model and the strongest choice in LTX Studio for cinematic, multi-shot work. It generates sequences up to fifteen seconds and holds subject consistency across camera angle changes, which is why it has become the default for brand films and product hero videos. Kling 2.6 Pro remains the pick for detail-preserving product and fashion shots where fine texture matters more than shot length.
Use Kling when:
- You need cinematic visual quality with strong motion consistency
- Your project involves product shots, fashion content, or brand films
- You want multi-shot sequences with consistent characters (Kling 3.0 Pro)
Seedance 2.0 — Best for Human Motion and Fluid Camera Work
Seedance 2.0 is the model to pick when a person needs to walk, gesture, dance, or interact with another person on screen. Its human dynamics feel physically credible in a way most competing video models struggle to match, and it holds temporal consistency through the full clip without the subject drift that appears on longer generations elsewhere. Seedance also handles dynamic backgrounds and fluid camera movement well.
Use Seedance 2.0 when:
- Your video features people walking, dancing, gesturing, or interacting and the motion needs to look physically natural
- You're generating dynamic scenes with multiple moving elements or complex background activity
- Camera movement is a key part of the shot — tracking, panning, or push-in moves that need to stay smooth
Veo 3.1 — Best for Dialogue, Realism, and Audio-Critical Content
Veo 3.1 is Google's flagship video model and the only model in LTX Studio that generates synchronized native audio in a single pass — voice, lip-sync, ambient sound, and effects all come out of the same generation. It also introduced dual keyframe control, letting you define the start and end frame of a clip. Veo 3.1 is available on Pro and Enterprise plans.
Use Veo 3.1 when:
- Your video includes dialogue, voiceover, or lip-sync that needs to feel natural
- You need maximum visual realism for a commercial, film, or brand campaign
- You want precise control over the start and end frame of each shot
Quick reference: Video model comparison
| Model | Best for | Standout feature |
|---|---|---|
| LTX-2.3 / LTX-2.3 Pro | Rapid iteration and vertical content | Native portrait video, open-source, tightest LTX Studio integration |
| LTX-2 Pro | Predictable open-weight production | Previous-generation open-source model with transparent weights |
| Kling 2.6 Pro | Product and fashion detail | Fine-texture retention on hero shots |
| Kling 3.0 Pro | Multi-shot cinematic storytelling | Fifteen-second sequences with cross-angle consistency |
| Seedance 2.0 | Human motion and fluid camera work | Physically credible body dynamics, low subject drift |
| Veo 3.1 | Dialogue and audio-first content | Native synchronized audio and dual keyframe control |
Which Image Model Should You Use?
Image selection in LTX Studio follows a different logic than video. Speed matters more, iteration count is higher, and the model you pick for the tenth exploration is often different from the one you pick for the final hero asset.
Nano Banana 2 — Best for Speed, Iteration, and Subject Consistency
Nano Banana 2 is Google's newest image model, built on the Gemini 3.1 Flash architecture. It generates images up to 4K and maintains subject consistency across up to five characters and fourteen objects. It handles text rendering far more reliably than previous models, making it useful for any asset that includes signage, logos, or branded copy. Use Nano Banana 2 for storyboards, early-stage concept work, and any workflow where multiple characters need to stay visually consistent.
Use Nano Banana 2 when:
- You need to generate and iterate quickly without sacrificing quality
- Your project involves multiple characters that need to stay consistent across shots
- You're working on storyboards, concept development, or early-stage creative direction
Nano Banana Pro — Best for Final Production Assets
Nano Banana Pro is the Pro-tier version of the same Gemini image family. Where Nano Banana 2 optimizes for speed, Nano Banana Pro optimizes for output fidelity — delivering richer detail, more nuanced lighting, and stronger prompt coherence on complex scenes. The practical pattern: iterate with Nano Banana 2 until you have a shot you like, then regenerate the final with Nano Banana Pro.

FLUX.2 Pro — Best for Brand-Accurate Campaign Volume
FLUX.2 Pro from Black Forest Labs is built for production-scale image work where brand color is not negotiable. It accepts HEX code input for exact color matching, generates 2K images in under 10 seconds, and is the right choice for social assets, ad creative, or product visuals at volume.
Use FLUX.2 Pro when:
- You need pixel-perfect brand color consistency across a campaign
- You're generating social media assets, product visuals, or ad creatives at scale

Z-Image — Best for Photorealism on Any Plan Tier
Z-Image from Alibaba's Tongyi Lab is a speed-optimized model focused on photorealistic output with tight prompt adherence. It is available across all LTX Studio tiers, including the free plan, which makes it a useful default for teams still evaluating the platform.
Use Z-Image when:
- Photorealism is the priority
- You need a capable model on a free or entry-level plan
- Speed and prompt accuracy matter more than text rendering

ChatGPT Images 2.0 — Best for On-Image Copy and Structured Layouts
ChatGPT Images 2.0 is OpenAI's GPT-4o-based image model, and the one to reach for when the image itself contains readable copy or when the composition follows precise structural instructions. Most image models pattern-match to a style. ChatGPT Images 2.0 interprets structured prompts more literally, which is why it wins on ad creatives with headlines, product mockups with labels, and any visual where copy placement is load-bearing.
Use ChatGPT Images 2.0 when:
- Your image needs readable text — headlines, labels, signage, or on-image copy
- Your prompt is structurally complex: multiple objects, specific spatial relationships, or layered visual instructions
Quick reference: Image model comparison
| Model | Best for | Standout feature |
|---|---|---|
| Nano Banana 2 | Fast iteration and multi-character scenes | Speed with strong subject consistency |
| Nano Banana Pro | Final production assets and complex lighting | Pro-tier Gemini fidelity for hero images |
| FLUX.2 Pro | Brand-accurate campaign volume | HEX code input for exact color matching |
| Z-Image | Photorealism on any plan tier | Speed-optimized, available on all tiers |
| ChatGPT Images 2.0 | Copy-heavy visuals and structured layouts | GPT-4o reasoning for text and composition |
How to Combine Models in One Workflow
LTX Studio’s real value is not any single model. It is the workspace around them. Every model sits inside the same Gen Space, connects to the same Elements library for character and asset consistency, and feeds directly into the same AI storyboard generator and timeline editor.
The pattern most production teams settle into: start with a fast image model (Nano Banana 2 or Z-Image) to explore directions. Save the strongest results as Elements so characters and styles travel across the project. Move to Nano Banana Pro or FLUX.2 Pro for the hero once the direction is locked. Then animate with the video model that matches the shot: LTX-2.3 for volume, Kling 3.0 Pro for cinematic sequences, Seedance 2.0 for human motion, or Veo 3.1 for dialogue and audio.
Three rules make the decision easier in the moment:
Match the model to the deliverable, not to your habits. A talking-head ad wants Veo 3.1 even if LTX-2.3 gets you a draft faster.
Use fast models for concepting and precision models for finals. Nano Banana 2, Z-Image, and LTX-2.3 exist to burn through iterations. Nano Banana Pro, FLUX.2 Pro, Kling 3.0 Pro, and Veo 3.1 exist to lock the shot.
Test two models on the same prompt before committing. A ten-second A/B run at the start of a scene saves an hour of rerenders later.
Conclusion
LTX Studio integrates twelve image and video models in one Gen Space, so picking the right one is a decision problem, not an access problem. Use LTX-2.3 for fast video iteration, Kling 3.0 Pro for multi-shot cinematic sequences, Seedance 2.0 for human motion, and Veo 3.1 for dialogue and audio. Use Nano Banana 2 and Z-Image for concepting, Nano Banana Pro and FLUX.2 Pro for brand-accurate finals, and ChatGPT Images 2.0 when the visual carries readable copy.
Combine them in one workspace, save characters as Elements, and let each model do the job it was built for. Start generating on LTX Studio and build the model stack that fits your production.
