Awesome LTX-2

A curated list of models, text encoders, and tools for the LTX-2 video generation suite.

ltx-logo

Intro

▓ Apps & Tools

LTX2.3-Multifunctional

LTX2.3-Multifunctional is a desktop-optimized version of LTX that lowers GPU requirements and simplifies usage. It integrates all features including image-to-video, text-to-video, start/end frames, lip-sync, video enhancement, and image generation into a single application.

Key Features:

Lower GPU Requirements: Only needs 24GB VRAM (vs 32GB for standard desktop version)
All-in-One Interface: No complex ComfyUI workflows or error-prone nodes
Features: T2V, I2V, start/end frames, lip-sync, video enhancement, image generation, LoRA support
Multi-Frame Insertion: Two modes for generating long videos
Easy Setup: No third-party software required, just install LTX desktop

Downloads & Resources:

HuggingFace | GitHub | ComfyUI Node | Tutorial

▓ Models

LTX-2 models are available in various formats including full weights, transformers-only, and GGUF quantizations for efficient inference.

▣ Checkpoints

Lightricks/LTX-2 - Official repository.
Lightricks/LTX-2.3 - Official repository (latest version).
Drbaph - Quantization

Ver	Name	Size
2.3	`ltx-2.3-22b dev`	46.1 GB
2.3	`ltx-2.3-22b dev`	29.1 GB
2.3	`ltx-2.3-22b dev`	29.9 GB
2.3	`ltx-2.3-22b dev`	29.1 GB
2.3	`ltx-2.3-22b dev`	21.7 GB
2.3	`ltx-2.3-22b dev`	29.1 GB
2.3	`ltx-2.3-22b distilled`	46.1 GB
2.3	`ltx-2.3-22b distilled`	29.5 GB
2.3	`ltx-2.3-22b distilled`	29.9 GB
2.3	`ltx-2.3-22b distilled`	29.1 GB
2.3	`ltx-2.3-22b distilled`	17.6 GB
2.3	`ltx-2.3-22b distilled`	29.7 GB
2.3	`ltx-2.3-22b distilled 1.1`	46.1 GB

2	`ltx-2-19b dev`	43.3 GB
2	`ltx-2-19b dev`	27.1 GB
2	`ltx-2-19b dev`	20 GB
2	`ltx-2-19b distilled`	43.3 GB
2	`ltx-2-19b distilled`	27.1 GB
2	`ltx-2-19b distilled`	20 GB

Quantized to fp8_e5m2 to support older Triton with older Pytorch on 30 series GPUs. For WangGP in Pinokio

Ver	Name	Precision	Size	Download
2	`ltx-2-19b dev`		27.1 GB

silveroxides Quantizations (mxfp8)

Note: The mxfp8mixed quantization requires a custom fork of ComfyUI-Kitchen with mxfp8 support. Standard ComfyUI installations may not support this quantization format.

Model	Quant	Size	Download
`ltx-2.3-22b-dev`		29.2 GB
`ltx-2.3-22b-distilled`		29.1 GB
`ltx-2.3-22b-distilled`		29.2 GB
`ltx-2.3-22b-distilled`		29.7 GB

Distilled LoRA

Ver	Rank	Size	Download
2.3	`384`	7.61 GB	┊
2.3	`208`	4.97 GB
2.3	`159`	3.83 GB
2.3	`111`	2.74 GB
2.3	`105`	2.59 GB

2	`384`	7.67 GB
2	`242`	4.88 GB
2	`175`	3.58 GB
2	`175`	1.79 GB

▣ TenStrip Distilled LoRA Experiments

Experimental distilled LoRAs optimized for finetunes and I2V workflows. These LoRAs avoid the issues of the massive rank 384 official LoRA which can be counterproductive with conditioned inputs and finetunes.

LoRA	Rank	Size	Description
`ltx-2.3-22b-distilled-lora-1.1_fro90_ceil36`	36	739 MB	Compact LoRA with dynamic ceiling at 36
`ltx-2.3-22b-distilled-lora-1.1_fro90_ceil72_condsafe`	72	662 MB	Cond-safe version with cross-attention bridges, adaln/scale-shift tables, gate logits, and prompt scale-shift zeroed. Much better suited for I2V and input conditioned workflows. Can use 1.0 strength safely on first pass I2V.
`ltx-2.3-22b-distilled-lora-fro90_ceil72`	72	1.4 GB	Standard version with higher dynamic ceiling

Notes:

Lower rank LoRAs (72 and below) can be used at 1.0 strength safely for I2V first pass, with upscale pass at 0.4-0.5 strength
_ceil suffix indicates the dynamic ceiling during reranking
_condsafe suffix indicates cross-attention and other conditioning layers have been zeroed for better I2V compatibility
The official rank 384 LoRA can actively dampen conditioning signals in I2V workflows; cond_safe versions work much better

Download All LoRAs

Spatial Upscaler

Required for current two-stage pipeline implementations in this repository. Download to COMFYUI_ROOT_FOLDER/models/latent_upscale_models folder.

Ver	Name	Size
2.3	`spatial-upscaler x2 1.0`	996 MB
2.3	`spatial-upscaler x1.5 1.0`	1.09 GB

2	`spatial-upscaler x2 1.0`	1.05 GB

Temporal Upscaler

Required for current two-stage pipeline implementations in this repository. Download to COMFYUI_ROOT_FOLDER/models/latent_upscale_models folder.

Ver	Name	Size
2.3	`temporal-upscaler x2 1.0`	262 MB

2	`temporal-upscaler x2 1.0`	262 MB

▣ Merges

Custom merged models combining multiple control signals or specialized configurations.

Ver	Name	Description	Download
2.3	`ltx-2.3-22b-distilled-1.1-fused-union-control`	Merged model combining Canny, Depth, and Pose control signals for unified control

══════════════════════════════════

▣ Finetunes

Community finetuned models based on LTX-2.3 with specialized improvements and optimizations.

Model	Description
	High-performance LoRA-integrated checkpoint family based on LTX 2.3. Includes both distilled (4-step) and non-distilled variants (20-30 steps). Recommended sampler: Euler + Simple/Normal/Linear_Quadratic.
	I2V-optimized merge using layer scaled merges at different steps. Not a straight weight merge - behaves much nicer than standard LoRA loading and respects prompts better. Includes BF16 full checkpoint and fp8_mixed_learned quantized versions.
	Uncensored video generation model based on LTX 2.3 supporting T2V and I2V natively. Includes a built-in prompt enhancer. Merge base for 10Eros. Supports GGUF format.

══════════════════════════════════

▣ GGUF Quantized Models

These models are optimized for lower memory usage. Note that in ComfyUI, these are typically loaded as transformer-only models.

QuantStack

QuantStack LTX-2.3

Model	Size	Download
ltx-2.3-22b	12.4 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	14.7 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	14 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	17.8 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	16.7 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	19.4 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	18.5 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	21 GB	dev ┊ distilled ┊ distilled-1.1
ltx-2.3-22b	25.5 GB	dev ┊ distilled ┊ distilled-1.1

QuantStack LTX-2

Model	Quant	Size	Download
LTX-2-dev		8.03 GB
LTX-2-dev		10.3 GB
LTX-2-dev		9.57 GB
LTX-2-dev		13.4 GB
LTX-2-dev		12.3 GB
LTX-2-dev		15 GB
LTX-2-dev		14.2 GB
LTX-2-dev		16.6 GB
LTX-2-dev		21.1 GB

Unsloth

Unsloth LTX-2.3 GGUF

Model	Size	Download
ltx-2.3-22b	42 GB	dev ┊ distilled
ltx-2.3-22b	42 GB	dev ┊ distilled
ltx-2.3-22b	8.28 GB	dev ┊ distilled
ltx-2.3-22b	10.8 GB	dev ┊ distilled
ltx-2.3-22b	9.95 GB	dev ┊ distilled
ltx-2.3-22b	12.7 GB	dev ┊ distilled
ltx-2.3-22b	13.8 GB	dev ┊ distilled
ltx-2.3-22b	14.3 GB	dev ┊ distilled
ltx-2.3-22b	13.1 GB	dev ┊ distilled
ltx-2.3-22b	15.3 GB	dev ┊ distilled
ltx-2.3-22b	16.3 GB	dev ┊ distilled
ltx-2.3-22b	16.1 GB	dev ┊ distilled
ltx-2.3-22b	15.2 GB	dev ┊ distilled
ltx-2.3-22b	17.8 GB	dev ┊ distilled
ltx-2.3-22b	22.8 GB	dev ┊ distilled
ltx-2.3-22b	9.5 GB	dev ┊ distilled
ltx-2.3-22b	13.5 GB	dev ┊ distilled
ltx-2.3-22b	11.4 GB	dev ┊ distilled
ltx-2.3-22b	16.5 GB	dev ┊ distilled
ltx-2.3-22b	14.2 GB	dev ┊ distilled
ltx-2.3-22b	18.3 GB	dev ┊ distilled
ltx-2.3-22b	16.3 GB	dev ┊ distilled

Unsloth LTX-2.3 GGUF - Distilled 1.1

Model	Size	Download
ltx-2.3-22b	42 GB	distilled-1.1
ltx-2.3-22b	42 GB	distilled-1.1
ltx-2.3-22b	7.94 GB	distilled-1.1
ltx-2.3-22b	10.6 GB	distilled-1.1
ltx-2.3-22b	9.74 GB	distilled-1.1
ltx-2.3-22b	14.2 GB	distilled-1.1
ltx-2.3-22b	13 GB	distilled-1.1
ltx-2.3-22b	15.9 GB	distilled-1.1
ltx-2.3-22b	15 GB	distilled-1.1
ltx-2.3-22b	17.8 GB	distilled-1.1
ltx-2.3-22b	22.8 GB	distilled-1.1
ltx-2.3-22b	10.9 GB	distilled-1.1
ltx-2.3-22b	13.4 GB	distilled-1.1
ltx-2.3-22b	16.4 GB	distilled-1.1
ltx-2.3-22b	14.1 GB	distilled-1.1
ltx-2.3-22b	18.2 GB	distilled-1.1

Unsloth LTX-2 GGUF

Model	Quant	Size	Download
ltx-2-19b-dev		37.8 GB
ltx-2-19b-dev		37.8 GB
ltx-2-19b-dev		10.1 GB
ltx-2-19b-dev		11.6 GB
ltx-2-19b-dev		8.1 GB
ltx-2-19b-dev		10.7 GB
ltx-2-19b-dev		10.1 GB
ltx-2-19b-dev		9.47 GB
ltx-2-19b-dev		11.3 GB
ltx-2-19b-dev		12.3 GB
ltx-2-19b-dev		12.8 GB
ltx-2-19b-dev		11.9 GB
ltx-2-19b-dev		13.7 GB
ltx-2-19b-dev		14.6 GB
ltx-2-19b-dev		14.3 GB
ltx-2-19b-dev		13.6 GB
ltx-2-19b-dev		16 GB
ltx-2-19b-dev		20.4 GB

Vantage

Vantage AI GGUFs

Model	Quant	Size	Download
ltx-2-19b-dev		9.96 GB
ltx-2-19b-dev		9.28 GB
ltx-2-19b-dev		11.6 GB
ltx-2-19b-dev		12.4 GB
ltx-2-19b-dev		12.8 GB
ltx-2-19b-dev		11.8 GB
ltx-2-19b-dev		13.6 GB
ltx-2-19b-dev		14.5 GB
ltx-2-19b-dev		14.4 GB
ltx-2-19b-dev		13.5 GB
ltx-2-19b-dev		15.9 GB
ltx-2-19b-dev		20.4 GB
ltx-2-19b-distilled		9.96 GB
ltx-2-19b-distilled		9.28 GB
ltx-2-19b-distilled		11.6 GB
ltx-2-19b-distilled		12.4 GB
ltx-2-19b-distilled		12.8 GB
ltx-2-19b-distilled		11.8 GB
ltx-2-19b-distilled		13.6 GB
ltx-2-19b-distilled		14.5 GB
ltx-2-19b-distilled		14.4 GB
ltx-2-19b-distilled		13.5 GB
ltx-2-19b-distilled		15.9 GB
ltx-2-19b-distilled		20.4 GB

Special Quantization: PolarQuant Q5

LTX-2.3 (22B) — PolarQuant Q5 is a bit-packed quantization method using Hadamard-Rotated Lloyd-Max Quantization. It achieves optimal Gaussian weight quantization via Hadamard rotation, delivering near-lossless quality with significant size reduction.

Specification

Specification	Value
Parameters	22B
Transformer Blocks	48
Hidden Dimension	4096
Layers Quantized	1,347 (of 5,947 total tensors)

Compression Statistics:

Component	Original Size	PQ5 Packed	Reduction
Transformer (1,347 layers)	37 GB	4.6 GB	-88%
VAE + Skip (4,600 layers)	9.1 GB	9.1 GB	BF16 kept
Upscalers	1.3 GB	1.3 GB	BF16 kept
Total	46.2 GB	15 GB	-68%

Quality Metrics:

Cosine Similarity: 0.9986 (near-lossless)
Download Size: 15 GB
Beats torchao INT4 on perplexity (PPL)

Hardware Requirements:

GPU	VRAM	Status
A100 (80 GB)	80 GB	Full speed
A100 (40 GB)	40 GB	Recommended
RTX 4090 (24 GB)	24 GB	With offloading

Key Features:

Mixed precision approach: transformer heavily quantized (-88%) while VAE remains BF16
5-bit bit-packed representation (Q5)
50-65% smaller than original with zero quality loss
One-command setup with easy generation wrapper

Model	Size	Download
`LTX-2.3-22B-PolarQuant-Q5`	15 GB

Installation: pip install safetensors huggingface_hub scipy ArXiv Reference: 2603.29078

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ Text Encoders

LTX-2 requires Gemma-3-12b variants. LTX-2.3 uses text projection layers.

▣ Comfy-Org Optimized Encoders

Official and optimized versions for ComfyUI.

Model Name	Size	Download
`gemma_3_12B_it`	24.4 GB
`gemma_3_12B_it_fpmixed`	13.7 GB
`gemma_3_12B_it_fp8_scaled`	13.2 GB
`gemma_3_12B_it_fp4_mixed`	9.5 GB
`gemma_3_12B_it-int8tensormixed`	13.2 GB
`gemma_3_12B_it-int8mixedblockwise`	13.6 GB
`gemma_3_12B_it-int8mixedtensorwise`	14.1 GB
`gemma_3_12B_it-int8tensormixed`	13.2 GB
`text_projection_fp8`	1.16 GB

gemma_3_12B_it_fpmixed: Experimental quant. Should be better than the fp8 scaled
gemma_3_12B_it_fp4_mixed: 90% fp4 layers

Note: The mxfp8mixed quantization requires a custom fork of ComfyUI-Kitchen with mxfp8 support. Standard ComfyUI installations may not support this quantization format.

Gemma-3-12b Abliterated

Why Choose Abliterated Encoders?

Standard Gemma models often incorporate safety alignment that "sanitizes" or weakens specific concepts within prompt embeddings. Even when the model doesn't explicitly refuse a request, this internal filtering can dilute creative intent. For LTX-2 video generation, using a standard encoder often results in:

Reduced Prompt Adherence: Key stylistic or descriptive terms may be ignored or weakened.
Visual Softening: Visual intensity and fine details are often "muted" to fit generic safety profiles.
Concept Dilution: Complex or niche creative requests are subtly altered, leading to less faithful representations of your vision.

Abliteration bypasses these restrictive alignment layers, allowing the encoder to translate your prompts into embeddings with maximum fidelity. This ensures LTX-2 receives the most accurate and un-filtered instructions possible.

Gemma-3-12b-Abliterated

Fixed versions of the abliterated Gemma-3-12b-it model by FusionCow, modified specifically for compatibility with LTX-2. The original model

Model	Precision	Size	Download
`Gemma ablit fixed`		23.5 GB
`Gemma ablit fixed`		13.8 GB

Gemma 3 12B IT Heretic

Models by DreamFast

Safetensors

Model	Precision	Size	Download
`Gemma_3_12B_it Heretic`		23.5 GB
`Gemma_3_12B_it Heretic`		12.8 GB

GGUF

Size	Quality	Recommendation
22GB	Lossless	Reference, same as original
12GB	Excellent	Best quality quantization
9.0GB	Very Good	High quality, good compression
7.9GB	Good	Balanced quality/size
7.7GB	Good	Slightly smaller Q5
6.8GB	Good	Still useful
6.5GB	Decent	Smaller Q4 variant
5.6GB	Acceptable	For very low VRAM only

Sikaworld1990 Gemma-3-12b Abliterated

NVFP4 quantization variants by Sikaworld1990 optimized for Blackwell GPUs.

Model	Precision	Size	Download
`Gemma-3-12b QAT Abliterated FP4`		12.1 GB
`Gemma-3-12b QAT Abliterated FP4`		8.91 GB
`Gemma-3-12b HereticX Abliterated`		15 GB
`Gemma-3-12b High-Fidelity Abliterated`		14.1 GB

FP4-HF: High-fidelity mixed precision calibration
FP4-Pure: Pure FP4 quantization for maximum compression
HereticX: Uncensored variant with maximum prompt fidelity
High-Fidelity: Optimized for quality with better detail preservation

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ Separated Components

Separated LTX2 checkpoint by Kijai and Kijai for LTX-2.3. For alternative way to load the models in Comfy.

▣ Diffusion Models (Transformer Only)

Ver	Name	Size
2.3	`ltx-2.3-22b dev`	42 GB
2.3	`ltx-2.3-22b dev`	23.5 GB
2.3	`ltx-2.3-22b dev`	24.1 GB
2.3	`ltx-2.3-22b dev`	25 GB
2.3	`ltx-2.3-22b distilled`	42 GB
2.3	`ltx-2.3-22b distilled`	23.5 GB
2.3	`ltx-2.3-22b distilled v2`	23.2 GB
2.3	`ltx-2.3-22b distilled`	23.5 GB
2.3	`ltx-2.3-22b distilled` (experimental)	24.1 GB
2.3	`ltx-2.3-22b distilled 1.1`	42 GB
2.3	`ltx-2.3-22b distilled 1.1`	25.2 GB
2.3	`ltx-2.3-22b distilled 1.1` (experimental)	24.1 GB

2	`ltx-2-19b dev`	37.8 GB
2	`ltx-2-19b dev`	21.6 GB
2	`ltx-2-19b dev`	14.5 GB
2	`ltx-2-19b distilled`	37.8 GB
2	`ltx-2-19b distilled`	21.6 GB

[!NOTE]
input_scaled additionally have activation scaling, and are set to run with fp8 matmuls on supported hardware (roughly 40xx and later Nvidia GPUs).

▣ VAE (Video & Audio)

Ver	Component	Size	Download
2.3	`Video VAE`	1.45 GB	┊
2.3	`Audio VAE`	365 MB	┊

2	`Video VAE`	2.45 GB
2	`Audio VAE`	218 MB

▣ Embedding Connectors & Text Projection

Ver	Name	Size	Download
2.3	`Embeddings Connectors dev`	2.31 GB	┊
2.3	`Embeddings Connectors distilled`	2.31 GB

2	`Connector dev`	2.86 GB
2	`Connector distilled`	2.86 GB

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ LoRA

▣ Enchancer, special

Lightricks LTX-2.3
- LipDub IC-LoRA - Enables lip dubbing on top of LTX-2.3 for video dubbing via joint audio-visual diffusion (based on JustDubIt research)
OmerHagawa
- Greenscreen Avatar IC-LoRA - Greenscreen avatar IC-LoRA for vertical video
systms
- SYSTMS FLW IC-LoRA - Seamless shot-to-shot transitions IC-LoRA with trigger word FLW, uses gray frames (RGB 127,127,127) between clips
LTX-2.3-IC-LoRA-Colorizer by DoctorDiffusion (331 MB) - Colorize black and white videos
JUST-DUB-IT
Best-Face-Swap-Video
Image-to-Video Adapter LoRA
- Original by MachineDelusions
- siraxe variant - Stripped audio layers + rank64 compressed (2.62 GB, 655 MB rank64 bf16)
Lightricks LTX-2.3
- HDR - Enables 16-bit HDR video generation and converts SDR video to HDR using LogC3 transform for extended dynamic range
- Union Control - Unified IC-LoRA combining Canny + Depth + Pose control signals for multi-signal video generation conditioning
- Motion Track Control - Guides object motion using sparse point trajectories via colored spline overlays on reference videos
vrgamedevgirl84
- Enhancer1 - Video enhancement LoRA
- Enhancer2 - Second video enhancement LoRA
oumoumad
- IC luminance map
- LTX-2 IC-LoRA-Ungrade - Removes color grading and contrast from footage, returning neutral ungraded appearance
- LTX-2.3 IC-LoRA-Ungrade - LTX-2.3 version of color grading removal IC-LoRA
- IC-LoRA-Outpaint - Extends video canvas by generating new content in black regions (letterbox areas), filling with temporally consistent content
- IC-LoRA-ReFocus - Removes lens blur and restores focus to out-of-focus footage (lens blur only)
- IC-LoRA-Uncompress - Removes MP4 compression artifacts (blocking, banding, mosquito noise) and restores clean video
- IC-LoRA-MotionDeblur - Removes motion blur from footage
- IC-LoRA-Deinterlace - Removes interlacing artifacts from video
- FXIC LTX2 IC-LoRA - Flux-inspired IC-LoRA for LTX video transformation with multiple optimizer variants (adamw, prodigy, masked) at various training steps
- DeArchive LTX-2.3 - In-Context LoRA for restoring archive video (old B&W footage, low-res web rips, sepia-toned silent-era prints) into colored, high-definition modern cinematography (Rank 128, 5,000 steps)
Kijai
- LTX2-IC-LoRAs - IC-LoRA trained with the realisdance set
Cseti
- IC-LoRA-Cameraman v1 - Transfers camera movements (zoom, pan, tilt, orbit) from reference video to generated output
- IC-LoRA-EditRefVid v1 - Edit reference video IC-LoRA for editing existing videos using reference guidance
100percentrobot
- Audio-Reactive LORA - Generates audio-reactive videos with motion synchronized to musical elements (beats, rhythm)
LiconStudio
- VBVR-lora-I2V - Enhances video generation for complex reasoning tasks including multi-object interactions, physical causality, and spatial relationships
- VBVR-lora-I2V Special
TheBurgstall
- LTX-2.3-Skin-Hair - Refines skin texture and hair rendering, reduces plastic skin artifacts, improves specular highlights
- VR-360-Outpaint IC-LoRA - Outpaints standard widescreen footage into a full 360° equirectangular projection for immersive/VR viewing.
Nightfury16
- Staging IC-lora 512 - Staging IC-LoRA for video composition control (512 latent scale)
siraxe
- MergeGreen IC-lora - Maintains motion at start/end frames, use middle frames with RGB 0,191,0 (75% green fill) in IC-LoRA workflow
- TTM IC-lora - Makes cutouts cartoony and adds cartoony characters to video scenes, based on the TTM approach (use with Img To Video bypass + Add Video IC-LoRA Guide node)
Lightricks LTX-2
- Canny Control - Edge detection control for structural guidance
- Depth Control - Depth map conditioning for 3D spatial control
- Detailer - Enhances fine details and textures in generated videos
- Pose Control - Human pose estimation control for motion guidance

Upscaler LoRAs:

LTX 2.3 Upscale IC-LoRA by Zlikwid
- Generative refinement LoRA for upscaling lower-res or soft videos
- Works by bicubic upscaling first, then running through LTX 2.3 with this LoRA
- Use prompt: upscale
LTX2.3-ICEdit-Insight by JoyFox Lab
- Task-aware video restoration and editing model family
- Supports: Video Restoration, HD Enhancement, Watermark Removal, Subtitle Removal
Singularity LTX-2.3 OmniCine by WarmBloodAban
- Comprehensive optimizer for LTX2.3 I2V and First/Last Frame workflows
- Features: Limb Evolution, Shot Injection, Natural Expression, Physical Integrity, Cross-Style Potential
- Uses "Singularity" prompting framework with 7-block bilingual structure

▣ Styles

OmerHagawa
- LTX2 UME PixelArt LoRA - Pixel art style LoRA for LTX-2
Playtime-AI
- Commission LTX-2.3 MtF Transformation - MtF transformation style LoRA for LTX 2.3
Andro0s
- Pixar Toon Style LoRA - Pixar-style CGI toon cinematic look with trigger word P1x4r
Cseti
- Arcane-Jinx v1 - Style LoRA inspired by Arcane's Jinx character design
- ReStyle IC-LoRA - Image-guided style transfer IC-LoRA that re-renders videos in a target style while preserving original content and motion
lopho
- Gantz O v1.0.0 - Movie-style LoRA (654 MB, 10000 steps)
bionicman69
- Arnold Style - Arnold Style LoRA for LTX 2.3. Get to the choppa!
- Star Trek TNG Style - Star Trek: The Next Generation style LoRA for LTX 2.3
oumoumad
- LTX-2-19b-LoRA-SPROUT
- Clay Stop Motion
kabachuha
- Hydraulic press
- Cakeify
- Big Anime Breasts
- Eat
- Squish – One Hand Only
- POP! Inflatable Animation - Comically inflate and pop cartoon/anime characters into confetti and fabric scraps (I2V focused)
CRT Animation Terminal by lovis93 - Real late-80s/early-90s CRT monitor look with scanlines, phosphor glow, chromatic aberration, and dithering. Trigger word: crtanim,. Available in 4000 and 10000 training steps variants
vrgamedevgirl84 Style LoRAs
- ClayMationStyle - Clay animation style LoRA for LTX-2.3
- Wild West Style
- Paper Cut Out Style
- Post Apocalyptic Style
- Pixar Toon Style
- Luxe Sensual Style
- Soft Enhance Style
- Crisp Enhance Style
- Fantasy Puppet Style
- Fantasy Realism Style
- Fantasy Painterly Style
- Fantasy Anime Style
- Cozy Felt Style
- Clay Mation Style
- 90s Animation Style
Alissonerdx LTX-LoRAs Collection - Comprehensive collection including:
- Anime2Half-Real - Converts anime-style content to half-realistic aesthetic (4500 steps, rank64)
- Edit-Anything Global - Global editing LoRA variants (6000-9000 steps, rank128)
- Inpaint Masked R2V/T2V - Region-based inpainting LoRAs for masked video editing
- Real2Anime/Anime2Real - Style conversion LoRAs (rank64)
Nebsh
Squish
Yoshiaki Kawajiri Retro Anime - LoRA trained on Yoshiaki Kawajiri's distinctive retro anime art style
Playtime-AI
- DonaldTrump
- Rick_and_Morty - BETA LoRA for Rick and Morty animated style
- LTX-2.3-Wednesday_Addams
- LTX-2.3-Kermit_the_Frog
- LTX-2.3-Jenna_Coleman
TheBurgstall
- LTX-2.3-Body-Positivity
- LTX-2.3-Googly-Eyes
TheBurgstall (LTX-2)
Black Venom
Lightricks

▣ Special

Wan2.1 VAE Adapter
- Latent space adapter for converting between LTX-2 and Wan2.1 VAE representations
- latent_adapter_final.pt (447 MB)

▣ ID-LoRA (Identity-Driven In-Context LoRA)

ID-LoRA is a method that enables identity-preserving audio-video generation in a single model. It jointly generates a subject's appearance and voice, letting a text prompt, a reference image, and a short audio clip govern both modalities together. Built on top of LTX-2.3 (22B), it is the first method to personalize visual appearance and voice within a single generative pass.

Unlike cascaded pipelines that treat audio and video separately, ID-LoRA operates in a unified latent space where a single text prompt can simultaneously dictate the scene's visual content, environmental acoustics, and speaking style—while preserving the subject's vocal identity and visual likeness.

Key Features:

Text prompt controls the scene and content
Reference image preserves the subject's visual likeness
Short audio clip preserves the subject's vocal identity
Single unified generation pass for both appearance and voice

Available LoRAs for LTX-2.3:

LoRA	LoRA Rank	Size	Download
ID-LoRA-TalkVid-3K	128	1.1 GB	┊
ID-LoRA-CelebVHQ-3K	128	1.1 GB	┊

Resources:

Project Page | GitHub | Paper (arXiv: 2603.10256)

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ ComfyUI Nodes

▣ Custom Node Collections

10S-Comfy-nodes by TenStrip - Custom ComfyUI nodes for improving motion quality when working with LTX 2.3's combined audio/video latent pipeline. Includes Latent Cross Fade Auto Concat, Audio Latent Stretch, Latent Motion Sharpener, Latent Temporal Upsampler, Latent Motion Retime, and Latent Temporal Inpainter for clean 30fps output from 24fps sampled models.
Deno Custom Nodes by Deno2026 - Practical ComfyUI custom nodes focused on fast real-world workflow improvements including (Deno) Resize Box, Multi Image Loader, LTX Sequencer, LTX Model Loader, Easy Model Download Helper, LTX Multi LoRA Loader, and LTX Prompt Guide.
PromptRelay by kijai - Enables consistent multilingual lip-sync while maintaining voice consistency across languages. Distributes video latent frames across segments with smart prompt node supporting inline and block syntax styles.
WhatDreamsCost ComfyUI by WhatDreamsCost - A variety of custom ComfyUI nodes and workflows for creating AI-generated video content including Multi Image Loader, LTX Sequencer, LTX Keyframer, Speech Length Calculator, Load Video UI, and Load Audio UI.
ComfyUI-Sapiens2 by kijai - ComfyUI nodes for Sapiens2 computer vision models from Facebook Research. Supports pose estimation, body-part segmentation, surface normal estimation, and pointmap estimation with model variants from 400M to 5B parameters.

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ LoRA Training

For training LTX LoRAs, the community uses a variety of official scripts, community-developed forks, and cloud-based platforms.

Primary Local Training Tools

Official LTX-2 Trainer: This is the standard Python-based package for training LoRAs, full fine-tuning, and In-Context (IC) LoRAs. It is designed for Linux and requires CUDA and Triton.
- Link: Official LTX-2 Trainer Repository
Musubi-Tuner (AkaneTendo25 Fork): Widely considered the fastest and most efficient local trainer for LTX-2 and 2.3. It features significantly smaller cache sizes (up to 12x smaller than AI Toolkit) and better iteration speeds, reaching up to 2 iterations per second on an RTX 5090.
- Link: AkaneTendo25 Musubi-Tuner Fork
AI Toolkit (by Ostris): A popular third-party tool that supports LTX-2 character and image-to-video LoRAs. While beginner-friendly, some users reported issues with audio training on the main branch.
- Link: Official AI Toolkit Repository
AI Toolkit: BIG-DADDY-VERSION (ArtDesignAwesome Fork): This specific fork was created to fix broken audio and voice training in the original AI Toolkit. It is optimized for hardware like the RTX 5090.
- Link: ArtDesignAwesome AI Toolkit Fork
rs-nodes (richservo): A collection of nodes that includes a full LTX Lora trainer directly within ComfyUI. It is designed to be memory-efficient, allowing training on cards with as little as 11GB-12GB of VRAM by using ComfyUI's native weight loaders.
- Link: rs-nodes ComfyUI Trainer
SimpleTuner: A highly optimized trainer for Linux that supports LTX-2 and is noted for its ability to handle larger datasets on limited VRAM via block swapping.
- Link: SimpleTuner Repository

Cloud Training Platforms

Fal.ai: Provides a dedicated cloud trainer for custom styles and effects, though it is primarily limited to image-based training datasets.
- Link: Fal.ai LTX2 Video Trainer
RunComfy: A cloud service that offers a pre-configured AI Toolkit setup specifically for LTX-2 training.
- Link: RunComfy LTX-2 Training

Essential Dataset & Captioning Tools

Taz's Ultimate Captioning Tool: A Hugging Face space frequently used by the community to generate the long, detailed, cinematographic prompts (around 200 words) that LTX-2 requires for high-quality training.
- Link: LoRA Caption Assistant (Hugging Face)
AI Video Clipper & LoRA Captioner: A modular pipeline designed to automate local dataset creation using WhisperX and Qwen2-VL, including support for RTX 5090 Blackwell cards.
- Link: AI Video Clipper & LoRA Captioner

Training Requirements Summary

Dataset: Videos should typically be cut to 121 frames (exactly 4.84 seconds) to align with the model's architectural "8n+1" rule.
Hardware: While 16GB VRAM is possible with extreme offloading in tools like rs-nodes, 24GB is the practical minimum for quantized training. For best results and speed, 48GB to 80GB (H100 or RTX 6000) is preferred.
Precision: It is now officially recommended to train on the full BF16 model for LTX 2.3 rather than FP8 for superior quality.

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ Workflow & Technical Notes

❖ Lightricks

LTX-2.3:

LTX-2:

❖ vrgamedevgirl84

vrgamedevgirl84 LTX 2.3 Music Video Creator:

Music Video Creator Workflow
- Prompt Creator Workflow - Audio upload, beat detection, scene timing, lyrics analysis, style selection, prompt generation
- Text-to-Video Workflow - LoRA integration, advanced prompt controls, Remake Mode, video stitching
- Image-to-Video Workflow - Uses Z-Image Turbo and LTX 2.3
- Requirements: ComfyUI, LTX 2.3 models, Z-Image Turbo model, FFmpeg, vrgamedevgirl custom nodes

❖ ComfyUI

❖ RuneXX

RuneXX LTX-2.3 Workflows:

Movie-Maker:

Talking-Avatar-TTS:

Video-2-Video:

Music-Video-Creator:

Others:

I2V T2V Basic custom audio with Gemma-API example

Custom-Audio:

First-Last-Frame:

Long-Video-Experimental:

3-Pass-Experimental:

Control-reference:

Helper-Workflows:

Other-examples:

I2V T2V Basic Custom Audio with Gemma-API

RuneXX LTX-2 Workflows old pre_feb2026

Awesome LTX-2

Intro

▓ Apps & Tools

LTX2.3-Multifunctional

▓ Models

▣ Checkpoints

silveroxides Quantizations (mxfp8)

Distilled LoRA

▣ TenStrip Distilled LoRA Experiments

Spatial Upscaler

Temporal Upscaler

▣ Merges

▣ Finetunes

▣ GGUF Quantized Models

QuantStack LTX-2.3

QuantStack LTX-2

Unsloth LTX-2.3 GGUF

Unsloth LTX-2.3 GGUF - Distilled 1.1

Unsloth LTX-2 GGUF

Vantage AI GGUFs

Special Quantization: PolarQuant Q5

▓ Text Encoders

▣ Comfy-Org Optimized Encoders

Gemma-3-12b Abliterated

Why Choose Abliterated Encoders?

Safetensors

GGUF

▓ Separated Components

▣ Diffusion Models (Transformer Only)

▣ VAE (Video & Audio)

▣ Embedding Connectors & Text Projection

▓ LoRA

▣ Enchancer, special

▣ Styles

▣ Special

▣ ID-LoRA (Identity-Driven In-Context LoRA)

▓ ComfyUI Nodes

▣ Custom Node Collections

▓ LoRA Training

Primary Local Training Tools

Cloud Training Platforms

Essential Dataset & Captioning Tools

Training Requirements Summary

▓ Workflow & Technical Notes

❖ Lightricks

❖ vrgamedevgirl84

❖ ComfyUI

❖ RuneXX

关于 About

语言 Languages

提交活跃度 Commit Activity

核心贡献者 Contributors