CivArchive
    LTX 2.3 — Image-to-Video Workflow - v1.0
    NSFW
    Preview 136320345

    Привет, друзья! 👋

    Перед вами простой собранный по стандартной схеме (за некоторым исключением) workflow для генерации изображений и видео, который я специально заточил под максимальное качество. Данный workflow использует стандартные ноды и не требователен к настройке.

    Процесс не идеален (бывают ошибки при генерации, и он довольно требователен к промпту), поэтому в нём нет автопромптера — пишем всё руками. По этой причине очень рекомендую внимательно прочитать раздел «КАК ПИСАТЬ ПРОМПТЫ» — там подробная инструкция. Также я добавил системные команды, которые понимает модель. Они хорошо помогают «прибить» камеру, не дают ей летать куда попало и в целом делают результат более предсказуемым.

    В отличие от моих прошлых работ (можете посмотреть в профиле), здесь я реально уделил особое внимание качеству видео. Многие смотрят на красивые превью и не понимают, почему у них так не выходит. Всё просто: здесь двойной апскейл (второй можно отключить, если ресурсов мало — иначе комп начнёт сильно греться), точные настройки нод и максимально адекватная модель.

    Важный момент: этот процесс лучше всего работает на коротких роликах 5–10 секунд. Именно на такой длине модель выдаёт самое стабильное и красивое видео. Если делать длиннее — качество постепенно проседает, появляются глюки и артефакты.

    Главные фишки схемы 🔥

    1. Двойной апскейл — один внутри генератора, второй на выходе. Благодаря этому картинка и видео получаются заметно чётче и вкуснее.

    2. Подробная инструкция по промптам — нормальная человеческая инструкция + поддержка системных команд. Становится гораздо проще получать предсказуемый результат.

    3. Модульная структура — всё разложено по удобным блокам и субграфам.

    4. Контроль лица в динамичных сценах — можно использовать контроль геометрии лица в движении (при желании полностью отключается).

    Теперь о плохом (честно про железо) ⚠️

    • Видеокарта: минимум 12 ГБ. На 12 ГБ работать будет, но придётся жертвовать разрешением и временем генерации. Комфортнее всего на 16 ГБ и выше. С меньшим объёмом памяти хороший результат дать будет тяжело, особенно с агрессивным апскейлом.

    • Оперативная память: лучше сразу 32 ГБ и больше. На меньшем ComfyUI начинает тормозить, вылетать и активно жрать диск. 32+ ГБ — это когда всё работает спокойно и без нервов.

    • Диск: желателен быстрый NVMe SSD от 512 ГБ (а лучше от терабайта), особенно если у вас ещё куча моделей лежит.

    Благодарности ❤️

    Большое спасибо создателю базовой схемы — LTX2.3 Dasiwa TI2V (大丝袜图文生视频0418).

    Я взял его работу, выкинул лишнее, поменял модели, подкрутил настройки и сделал процесс удобнее, сохранив главное — желание делать качественно.

    # 🎬 High-Quality Image-to-Video Production Workflow (LTX 2.3) [Max Quality Focus]

    Hello friends! 👋

    This is a straightforward, standard-based (with a few custom tweaks) image and video generation workflow that I specifically optimized for maximum possible visual quality. It utilizes core nodes and requires minimal initial setup.

    > ⚠️ Please Note: This process isn't completely foolproof (occasional generation errors can occur, and it is quite sensitive to prompt structure). There is no AI auto-prompter here — everything is written manually. Because of this, I highly recommend carefully reading the "HOW TO WRITE PROMPTS" guide included inside. I have also added model-specific system commands. These help lock the camera in place, prevent unwanted camera drift, and make the final movement much more predictable.

    Unlike my previous lighter setups, I focused heavily on pure video fidelity here. Many users see beautiful previews on Civitai but wonder why their own outputs don't match. The secret is simple: this workflow implements a dual upscale system (the second pass can be disabled if your PC overheats or runs out of resources), precise node tuning, and a highly responsive model configuration.

    Optimal Video Length: This workflow performs best with short 5 to 10-second clips. Keeping generations within this timeframe ensures the most stable, artifact-free, and visually stunning videos. Pushing for longer clips will gradually degrade quality and introduce visual glitches.

    ---

    ### 🔥 Key Features & Highlights

    * Dual-Stage Upscaling: One upscale pass occurs inside the generator, and a second refinement pass is applied at the very end. This makes both static images and final videos noticeably sharper, cleaner, and more professional.

    * In-Depth Prompting Guide: Comes with a comprehensive, easy-to-understand manual prompting guide alongside native system command support. This gives you much better control over the final scene.

    * Modular Clean Layout: Everything on the canvas is neatly organized into clear, colored blocks and subgraphs for easy navigation.

    * Dynamic Face Geometry Control: Includes an optional facial geometry lock for high-motion scenes to prevent facial warping during fast movements (can be completely toggled off if not needed).

    ---

    ### 💻 Honest Hardware Requirements (The Reality Check) ⚠️

    Because this workflow is tuned for premium quality, it requires decent hardware to run smoothly:

    * **VRAM (Graphics Card):** 12 GB minimum. It will run on a 12 GB card, but you will have to compromise on base resolution and wait longer for generations. For a comfortable experience, 16 GB VRAM or higher is recommended. Achieving great results with lower memory is difficult, especially with the dual upscale enabled.

    * **RAM (System Memory):** 32 GB or more is highly recommended. With less RAM, ComfyUI might stutter, crash, or heavily rely on disk paging. 32+ GB ensures everything loads smoothly without stress.

    * Storage (SSD): A fast NVMe SSD with at least 512 GB (ideally 1 TB+) of free space is recommended, especially if you host multiple heavy models.

    ---

    ### 🤝 Credits & Acknowledgments ❤️

    Huge thanks to the original creator of the foundational layout: *LTX2.3 Dasiwa TI2V (大丝袜图文生视频0418)**.

    * I took their solid framework, stripped out the unnecessary bloat, swapped the models, fine-tuned the internal settings, and restructured the layout to make it user-friendly while keeping the main goal intact — delivering top-tier output quality.

    Description

    FAQ