This ComfyUI workflow generates a lip-synced video from a single static image and an audio track using the MiniMax-H3 Ref2VA model. Accelerated by Kijai's lightx2v 4-step Turbo LoRA and SageAttention patch. There is no need to include the lyrics in the prompt, the model is able to lipysnc to the provided reference audio track.
Description
FAQ
Looks like we don't have an active mirror for this file right now.
CivArchive is a community-maintained index — we catalog mirrors that volunteers upload to HuggingFace, torrents, and other public hosts. Looks like no one has uploaded a copy of this file yet.
Some files do get recovered over time through contributions. If you're looking for this one, feel free to ask in Discord, or help preserve it if you have a copy.
Details
Downloads
160
Platform
CivitAI
Platform Status
Deleted
Created
8/14/2026
Updated
8/16/2026
Deleted
8/14/2026