CivArchive

    ALL QUESTIONS WILL BE ANSWERED IN THE COMMENT :)

    Cause there is a lot to say about how the model and workflows works...
    I will try to answer and help you if you get some problems :)


    (Upload Still in progress...)


    Upload In Progress....
    COMING SOON :

    • LTX+2 Q6_K.GGUF

    • LTX2 Audio VAE BF16

    • TEXT ENCODER GEMMA 3 12B

    • LTX-2-1B-Embeddings_connector_distill_BF16

    • Special GGUF EDITED WORFLOW IMG2IMG / TXT2IMG

    • LORA DISTILLED VERSION

    • SPATIAL UPSCALER

    • TEMPORAL UPSCALER

    • CAMERA CONTROL LORAS.

    • CONTROLNET AIO LTX2

    • Workflows I2V / V2V / T2V / VDETAILER.

    (Keep in mind i'm not the owner of this model)
    I made this upload from the official Public LTX Huggingface repo for Civitai.
    https://huggingface.co/Lightricks
    I just want to centralize the most useful models on this Civitai page,
    so everyone doesnโ€™t get lost among the many models available on the original repo.
    Iโ€™m not asking for any โ€œbuzzโ€ or anything like that, itโ€™s just to help the community have a solid LTX base on Civitai (for people not familiar with Hugging Face).
    I wonโ€™t upload big models, I leave that to people who actually want to use them, even if the benefits of doing so will be really minimal.
    But I will also centralize here my own releases and base-model variations of unofficial LTX.


    โšก LTX-2 FP8 โ€” Distilled (Fast & Lightweight)

    What is LTX-2 FP8 Distilled?

    The FP8 Distilled version is a compressed and accelerated variant of LTX-2, trained to replicate the behavior of the full model while being faster and lighter.

    Distillation reduces model complexity, making it more efficient โ€” at the cost of some fine-grained detail.

    โœ… Key Characteristics

    • Faster generation speed

    • Lower VRAM requirements

    • Quicker prompt response

    • Slightly reduced fine detail compared to full FP8

    • Excellent quality-to-performance ratio

    ๐ŸŽฏ Best Use Cases

    • Rapid iteration & testing

    • Prompt exploration

    • Draft videos and previews

    • Creators with limited hardware

    Recommended if:
    You want speed and accessibility, and are willing to trade a small amount of detail for faster results.


    ๐Ÿ”น LTX-2 FP8 โ€” Standard (Full Quality)

    What is LTX-2 FP8 (Standard)?

    The FP8 Standard version is a full-quality LTX-2 model quantized to FP8 precision.
    It preserves the complete architecture and capabilities of the original model while reducing memory usage.

    This is NOT a simplified model.
    Only the numerical precision is reduced โ€” the modelโ€™s intelligence, structure, and behavior remain intact.

    โœ… Key Characteristics

    • High visual fidelity and detail

    • Strong temporal consistency

    • Full audio-video synchronization

    • Lower VRAM usage than FP16

    • Stable and reliable for long generations

    ๐ŸŽฏ Best Use Cases

    • Cinematic video generation

    • Final renders and high-quality outputs

    • Creators who want maximum quality with lower hardware requirements

    Recommended if:
    You want the best possible quality in FP8, with no compromise on features or flexibility.


    ๐Ÿง  Which One Should You Choose?

    • ๐ŸŽฌ Go with FP8 Standard if quality and consistency matter most

    • โšก Go with FP8 Distilled if speed and efficiency are your priority

    Both versions are fully compatible with ComfyUI workflows and part of the same LTX-2 creative ecosystem.


    ๐Ÿ“Œ What is LTX-2?

    LTX-2 is a powerful multimodal AI model that transforms text prompts, images, or other media into fully synchronized audiovisual videos โ€” with motion, dialogue, music, and ambient sound generated in one unified pass. Itโ€™s built on a hybrid Diffusion-Transformer (DiT) architecture designed specifically for efficient spatiotemporal generation and audio-video alignment. LTX-2+1

    This approach lets creators go from idea to cinematic result without stitching separate audio tracks manually โ€” a major step beyond typical text-to-video systems. LTX-2


    โœจ Key Features & Capabilities

    ๐ŸŽฅ Cinematic Quality Output

    • Native 4K resolution support with playback up to 50 FPS, delivering smooth, high-detail video clips ideal for cinematic, commercial, or creative use. LTX-2

    ๐ŸŽต Unified Audio & Visual Generation

    • Generates synchronized audio โ€” including dialogue, ambience and music โ€” alongside the video in a single generation pass, removing the need for external audio sync tools. LTX-2

    ๐Ÿ”„ Flexible Input & Output Modes

    • Works with text prompts, image references, multi-keyframe conditioning, and more to animate concepts or stills into motion. LTX-2

    โš™๏ธ Performance Modes

    • Multiple performance configurations (Fast, Pro, Ultra) allow creators to balance speed and quality according to project needs โ€” from quick drafts to production-ready renders. LTX-2

    ๐Ÿง  Efficient & Accessible

    • Highly optimized for consumer-grade GPUs โ€” efficient enough to run on ~16 GB VRAM hardware with FP8/FP4 quantization options โ€” making AI video production more accessible. Reddit

    ๐Ÿ› ๏ธ Open & Extensible

    • Fully open weights, codebase, and workflows, enabling fine-tuning, custom LoRAs, and integration into tools like ComfyUI. Hugging Face


    ๐Ÿ“ˆ Improvements Over Earlier Versions

    Compared to the original LTX family and other open video models, LTX-2 raises the bar in several key areas:

    โœ… Audio Integration Built-In
    Instead of generating silent videos and requiring post-processing, LTX-2 outputs audio and visual streams together with temporal coherence. LTX-2

    โœ… Higher Resolution & Frame Rates
    Supports native 4K at up to 50 frames per second, reaching cinema-grade quality unlike many earlier community models that cap at lower resolutions or fps. LTX-2

    โœ… Longer Clips
    Offers extended duration generation (up to ~20 s clips) with continuous quality and audio coherence โ€” exceeding many alternatives. LTX-2+1

    โœ… Expanded Workflows
    Native support in ComfyUI plus custom workflows empowers users with text-to-video, image-to-video, multi-keyframe conditioning, and creative control nodes. comfyui.org+1


    ๐Ÿง  Typical Use Cases

    ๐Ÿ”น Cinematic storyboarding & concept visuals
    ๐Ÿ”น Social media & marketing video content
    ๐Ÿ”น Animated storytelling & motion design
    ๐Ÿ”น Game cutscenes & immersive narratives
    ๐Ÿ”น Product visualizations & dynamic ads

    Whether for rapid prototyping or production output, LTX-2 empowers creators with professional-grade generative video. LTX-2


    ๐Ÿงฉ Included Files & Variants

    Depending on the checkpoint uploaded, this collection may include:

    • Full Model Checkpoints (bf16 / fp8 / fp4) โ€” maximum quality with quantization options

    • Distilled Variants โ€” faster iteration with lighter compute cost

    • Spatial & Temporal Upscalers โ€” improve resolution or frame rate via multiscale pipelines

    • LoRA & Fine-Tuning Packs โ€” custom stylistic or control extension modules Hugging Face


    ๐Ÿ”ง ComfyUI Integration & Workflows

    Included workflow templates help you use LTX-2 in ComfyUI with nodes for:

    ๐Ÿ“Œ Text-to-Video โ€” generate animated clips from prompts
    ๐Ÿ“Œ Image-to-Video โ€” animate still images with camera motion and style
    ๐Ÿ“Œ Video Conditioning โ€” extend clips forward/backward or refine motions
    ๐Ÿ“Œ Keyframe Controls โ€” precise guidance over scene transitions

    These workflows are designed for ease-of-use and creative flexibility while demonstrating best practices for prompt structure and smooth temporal motion. LTX Documentation


    ๐Ÿง  Foundation Model Philosophy

    LTX-2 goes beyond a single task โ€” itโ€™s a foundation model for audiovisual creative AI. Open access to its weights, code, and tools encourages developers, artists, researchers, and hobbyists alike to customize, extend and innovate on a common platform. Hugging Face


    ๐Ÿ“Œ Summary

    LTX-2 is not just another video model โ€” it is a production-ready, synchronized audio-video foundation model that pushes the boundaries of what open discourse video generation can achieve. With cinematic output quality, flexible workflows, and a fully open ecosystem, LTX-2 stands as one of the most capable generative video tools available today. LTX-2

    Description

    ๐ŸŽ›๏ธ LTX-2 ControlNet โ€” CANNY

    LTX-2 ControlNet Canny

    The Canny ControlNet allows you to guide LTX-2 video generation using edge detection, giving you strong control over shapes, silhouettes, and composition.

    What it does:

    • Uses Canny edge maps as structural guidance

    • Preserves outlines and object boundaries

    • Keeps compositions consistent across frames

    Best for:

    • Precise scene layout control

    • Stylized or illustrated visuals

    • Maintaining architecture, props, or character shapes

    Why use it:
    Canny ControlNet is ideal when you want the model to respect a specific structure while still allowing creative freedom in textures, lighting, and style.

    Creator tip:
    Works especially well with clean input images or frames with strong contrast.

    Checkpoint
    LTXV

    Details

    Downloads
    46
    Platform
    SeaArt
    Platform Status
    Available
    Created
    1/24/2026
    Updated
    1/24/2026
    Deleted
    -

    Files

    Available On (1 platform)

    Same model published on other platforms. May have additional downloads or version variants.