Yet Another Workflow : easy t2v + i2v
I've aimed at a user-friendly UI for ComfyUI. There's a balance between complexity and ease of use, and this workflow aims to give you useful controls with clear guidance on what you need to care about. I hope these will be helpful to anyone strugging with quality and the general UI-isms of ComfyUI. I've taken the time to color code and add lots of notes. Please read the notes, I've tried to make them useful!
This is the workflow I use, it's not aimed at a skill level. It's designed to be easy to use and adjust with some UI concessions and labeling to ensure you can pilot it with less experience in a way that is more sophisticated than the official example workflows, which can be easy to break.
The primary goal with this workflow is to give you a strong foundational place to generate either text to video (T2V) or image to video (I2V) outputs without having to fuss too much. Lightx2\ning is on by default. (It's an accelerator that trades variety for generation speed.)
The green controls are the stuff you generally want to mess with.
The secondary goal here is to provide a consistent interface to interact with different samplers.
What's new v0.50?
The main "new" things are the "low-res to high-res" option, and a new trick to make it easier to do low-noise only for lightx2/ning (combine that with a 5 step overide for high), and you basically get a short cut to using some of Wan 2.2's full power. (You can also do both!) It's not without compromise, but it's more performant that doing it with lightx2/ning disabled.
There's more going on under the hood. The hope is that everything is tuned a bit more tightly and that you have a few new options. The routing is completely different; the new SamplingPlanner nodes I addded are there to help simplify the way the model is being fed models and sampling parameters.
It brings the behavior closer to an MoE style sampler, while still exposing the controls and independent phases. I have some questions about the approach still, but it does allow the process to provide numbers closer to what what Wan is expecteding for T2V and I2V without needing to fuss with it too much and providing overrides where needed.
Versions
The "main" workflows (the one's without parenthetical version labels) support the basic ksampler node, but also includes a toggle to enable the ClownsharKSampler sampler and the TripleKSampler once you have some experience and want to mess around.
I generally recommend the main workflow. It's my daily driver. It offers the most control with the least fuss. Each version has its place tho!
I am in the process of updating the workflows to v0.50, but they will all work as-is. Below this line is still in the queue for updating.
If you are extremely new to Comfy and Wan? Consider using the MoE version. It removes a few nodes and options while providing mostly the same interface with slightly less visual complexity to help you get acclimated. Once you get comfy with this, step up up to the main version for more options.
Want better edge case prompt adherance? I've created a version of the workflow that supports the WanVideo nodes. I don't recommend using this one until you're more comfy with the standard version. It has increased visual complexity. These nodes work completely different to other systems, and I hope to make it more accessible by providing you with the same interface to engage with it. WanVideo tends to produce completely different results, so it can be another intersting thing to explore.
Want more fluid motion and jiggle? I've also created a Smooth Mix version to support the Smooth Mix checkpoint. What is it? Like with Stable Diffusion checkpoints, it merges many LoRA's into the base Wan 2.2 model to create a more opinionated model to create videos with. This version follows the recommendations in their official workflow, while offering you the improved YAW UI experience. I like this checkpoint for it's detail and motion, but it is also more prone to motion artifacts. It also has some built in support for anime styles. A self-forcing LoRA (Lightx\ning) is built-in, so the sampler options are kept simple for this one. Please note that, due to the additional 80gb of size, my RunPod template will only include this as an optional download. Also checkout the LoRA version, which I find much more useful as you can adjust the strength of the effect.
Expect an update of the RunPod template to include the new templates soon after.
As of version v0.38, I'm doing a revision of this article so patchnotes will be removed for clarity. The changes are noted in the file details section (and in the templates themselves).
Like it?
Give it a like! Tag it as a Resource when you use it! Support on Patreon or a tip on Ko-fi are also welcome. Yellow Buzz will go towards promoting awareness here on Civit.
Need help?
I like helping people get going with this stuff, so if you want help message me. If you want extended one-on-one help, there's an option on the Patreon. I'm happy to walk you through the details, answer your questions, and give you some extra tips and tricks, and scripts. I've done this for a few folks, I'll save you money and headaches.
I've also written an article here on getting it going with my Runpod template. The template will vastly expedite and simplify getting things up and running.
General Advice
Make lots of videos! Post your videos! Don't fuss with the tech! Be smart about how you spend your time with this stuff. It's easy to burn out if you spend more time trying to get things to work than making videos you like. That's really why I'm posting this.
Use RunPod. Use the RTX 5090 or the H100 SXM. Use my Wan 2.2 template. If you've not used RunPod before, sign up with my link; we'll both get some free credit. See the article for more.
If you use a service like RunPod, if you're doing I2V, it can be smart to have your images ready in advance to make sure the server stays busy while you are using it.
If you run this outside of Runpod, you'll need to install some custom nodes. To do that, click the "Manager" button at the top of the Comfy interface, and then click the "Install Missing Custom Nodes". Click "Install" on each one - I recommend in order; you'll need to wait till each has installed. Do not bother restarting ComfyUI until they are all installed. The RunPod template has them preinstalled. (There's a manual patch for the LTXFilmGrain node here.)
If the wires bother you, there's a button in the bottom right on the floating UI that will hide them.
What is Lightx2\ning? That's just my short hand for refering to Lightx2 and Lightning (which is just the Wan 2.2 version) self-forcing LoRA's .
I've made it easy to turn off Lightx2\ning as well, if you want to try without, but note that it's much slower! I really only recommend this with the H100 SXM. Do try though, especially with text-to-video! The full Wan 2.2 has some amazing capability.
This workflow is setup for .safetensors models, but you can use GGUF if you want to make the changes node changes.
If the having the Clownshark/TripleK sampler in the UI is distracting, you can delete the group with no negative consequence. (You could also delete the purple mute node for the sampler selection as well.)
Costs?
I'm updating the data here to reflect additional testing: In case you are curious, the example videos take around 4.5 / 3 minutes (720x1280). (I don't normally do that resolution when I'm just making stuff and experimenting.) I can generally make nice looking videos in 1-2 minutes. I'm generally running at either $1.04 or $3.04 per hour with the RTX 5090 or the faster but more expensive H100 SXM; in generally I think between 15 to 68 high quality videos per hour is what I tend to see, so about $0.02 - $0.13 per video, (rounding up). (With a session startup cost for loading the pod, probably adding a cent to so to that.) 1 to 2 minutes is probably my gen sweet spot for time, so it's either great or a bit over my ideal depending on resolution/scene complexity, but that's a cost consideration.
Troubleshooting
If a node is missing (bright thick red outline with a warning when you open the workflow), you can install them by going to Manager > Install Missing Custom Nodes, and pressing Install on any the nodes that show up there.
If you are getting any errors related to a custom node, it's possible something has changed recently in the software. It might be useful to change a version back to the last "stable" build in these situations.
For example, the nightly build of WanVideoWrapper might introduce an error that wasn't there last time. With a workflow open, you can go to Manager > Custom Nodes in Workflow. This will show you all of the custom nodes. If you click, Switch Ver, you can see all of the releases. Consider trying the first numbered on at the top of the list.
If that doesn't work, or there seem to be more significant problems and you are using RunPod, you may have forgotten to select CUDA 12.8. Try restarting the server. If that doesn't work, terminate the pod, and make a new one. This will fix a surprising number of possible issues.
Longer video generation support?
One day. Probably.
I'm always looking for a good solution to this. I've not found a good solution to this problem yet that isn't very complex. To talk through them a bit:
There are some specialized solutions like Wan Animate and Infinite Talk that achieve longer videos by utilizing other technology to specific ends (remapping motion/making a talking head video), and while VACE is promising, it's very complex to setup and use and requires multiple steps. There are also techniques that involve making keyframes for your scene and using first/last frame to fill in the actual animations, and you can use interpolation as a post processing step to blend those clips in a way that can hide seams. Most of this also requires color correction or ipadapter to keey faces consistent.
The SVI LoRA is a newer technique. It stabilizes consistency across videos, but lowers the base quality (everything gets less sharp), and the scenes become volitile to big changes while increasing the overal consistency across multiple videos. It's not perfect, and cannot go infinite, but if you're dead set on longer videos, this is a decent technique. It doesn't meet my quality bar. I find the overall drop in fidelity to be disappointing.
At the end of the day, it's either a ton of work to make a still-short video, or you've introduced a ton of compromise on what's already a compromise. That's not what I'm selling here.
I see this as the biggest problem in the AI video space, whether you do this as a hobby, like most of us, or you're a company trying to figure out how to seriously use this stuff commercially. These problems are also not unique to Wan, though they vary from company to company. There's a technology problem for how to extend video, so I suspect that there's a lot of economic pressure and research effort that will probably lead to better videos that aren't "more VRAM", as that doesn't scale well.
To be clear: You can do this now by using the last frame as the first frame. v0.38 adds that capability. You'll generally get 2 or 3 decent extensions, but you're taking a quality hit each time, but any camera movement or motion may not look consistent between clips. (Using the same seed, sadly, not not ensure consistency.)
Sound?
Once it get's much better. Sora 2 and the other private models can do amazing sound, but the available public models create audio that I really dislike. You can certainly add it yourself, if you like it, but I won't officially support until it improves. LTX-2 can do decent sound and lipsync, but has a lot of issues which I'll cover elsewhere.
Description
- Created a new image size custom node to reduce size adjustments for i2v
- Switching node isolate to improve install issues
FAQ
Comments (35)
Hello. Great job! It's incredible. Is it possible to add Loras via Civicomfy to the workflow?
As someone just getting back into video gen, this workflow is incredibly intuitive and helpful. I had a bunch of links open with all the things i want to try in wan2.2, but with no real guidance on what's important or actually works. The options and explanations here are so so good. Keep up the amazing work.
Thank you for the kind words! This is always what I hope to hear; and ultimately the goal I'm aiming at. It's definitely not for everyone, but I aim to leverage my experience with such things to help make this stuff more accessible. (And thanks for the tip!)
For the love of god stop making these dumb workflows where everything is dumped into one giant mess that no one but you can read. People usually want to actually learn and take something useful away from your workflows. Make them readable from left to right with enough space between everything. People who say this workflow is “incredibly intuitive” are insane. To understand how anything works and what’s connected to what, you have to spend like 10 minutes dissecting this pile of shit. And the author deliberately pinned every tiny little thing in this mess to prevent you from tweaking it.
How did you ever think it was a good idea to pile a mountain of crap from top to bottom and paint it in a million colors? My god I just can’t take it anymore
Sorry it didn't work for you. There are plenty of others to check out.
Don't be a whiny bitch. If this is too complex go to an easier workflow. Comfy ain't intuitive and there's a learning curve.
@Lifshift People on the internet get fussy and dramatic. I try to show some grace. I have strong opinions about this stuff too, but some people don't want to read. Percieved complexity can overwhelm people, and that's fine. My workflow is not for there. There's plenty of "simple" workflows out there.
I dunno man, our boy's workflow worked for me out-of-the-box. Other workflows haven't been able to do that.
It doesn't work for me either, it gives an error on Sampler! Currently the problem is WanMoeKSampler!!!!!! If not this one, some other Sempler
@tany6666372 If you're running local, it's possible you have an issue with the install for the node. Additionally, be careful with how you open the workflow. Dragging and dropping it from the Workflows side panel can break things. Just double click. I'd need to know a lot more to help further. Feel free to shoot me a message here post on the Discord.
Wow truly amazing! This workflow was perfect, worked great off the bat and seeing one working helped me figure out what I did wrong.
Oh those who say its too messy or too hard to figure out. Before we had AI...remember those days? Cant figure it out. Here's the solution....make google your friend! I get an error, I google it and presto! I got an answer! Google AI mode is very helpful too. Ask it the error and it will walk you through.
I look at this awesome worlflow as - i could say its hard and give up OR waste money on cloud services by some fly by night company selling you overpriced image or video generation. Well me I rather turn to community workflows than pay. Not to mention Comfy is very buggy. Im not even updating Comfy yet.
My situation is worse as I have the 5070ti and what I call the Blackwell nightmare. Or GPU driver hell. Lol
@brianr1973871 Thank you for your kind words. I'd welcome it as comment on the main post and/or a review.
I generally agree with you about most all in one, etc workflows on here but this one in particular is pretty straight forward
V0.39 is your best work yet Mr. Titmuffin !
Honestly out of all the WAN-type workflows, this thing loaded up on RunPod in less than 5 minutes from all models downloaded. I didn't have any LORAs or add any extra models but opened up the workflow. It was amazing. That was truly one of the best experiences I've had with these things. Anyways it was much better than most of them I tried.
Thanks again for the kind words. Yeah. The LTX one is especially fast to start because it doesn't have as much to download! Wan has bigger models, so it can be a bit slower, but even with all my extra downloads it's often not too bad. Just depends on the data center and the day.
Seems to only produce errors, I install what it needs and restart the UI to get an error about a node that I have no idea where to find in such a complex workflow. I then install the find node add in, restart again and run it to confirm which node ID was the issue and then I encounter 28 new errors - given up trying to use it, or make sense of it, how on earth people are getting any output at all from this?
As is, I'd need a lot more info to help. Since you're running locally, your personal ComfyUI configuration can have a lot of possible variables about what issues could come up. As an ecosystem, whatever you already have installed and which version you are running are all major factors in diagnosing something.
Hi! Firstly, thanks very much for creating the workflow. Very helpful for a returning user. I have managed to get both your runpod and local flows working, however I noticed a few 'bugs' with the local version (that I exported as a json from the runpod).
The runpod is super clean, but locally things are a mess. Wires all over the place and overlapping nodes. The rgthree i2v/t2v switch doesn't work, so I had to bypass it entirely. The video length setting is totally blank. You can click into it to set the length in the options, so it technically works, but visually it's just a blank green box. Also the KSampler won't generate high and low previews like with the runpod. Any thoughts?
Local is more difficult to diagnose, as there's a lot of things I cannot know about your install. You need to install all of the relevant custom nodes for it to behave as expected. You need to ensure Nodes 2.0 is turned off. Your previews is probably related. Install Video Helper Suite and ensure the latent2rgb preview type is selected.
You can hide wires if you want. The wires are just comfy being comfy. Theres a button for that in the bottom right (hide/show links). I disable it by default on the pod to make it a bit less intimidating. You don't generally need to see them unless you're doing modifications.
And you should install SageAttention if you haven't.
@boobkake22 Thanks for responding. Disabling nodes 2 did indeed fix most of the UI bugs for me :)
I have Video Helper Suite (according to the extensions manager, although my command prompt window keeps telling me it's not installed on the boot up). I wonder if it's just a RAM issue not displaying the animated High and Low Noise KSampler previews (instead of static images).
Anyway, I appreciate your help and again for the workflow!
@simonsmithrandom407 Yeah, no worries. You might want to check the "conflicts" in the manager too. I had a problem where the "Bleh" pack was breaking my generation previews at one point. Also check your log file, it will tell you during startup if someone cannot load, and more importantly, why.
挺棒的,一开始还有点手足无措,但是之后我发现这是我这种新手在网站上能找到的最棒的I2V工作流,谢谢你
"It's great! I was a little overwhelmed at first, but then I realized it's the best I2V workflow a beginner like me could find on a website. Thank you!"
I'm glad it's useful for you. That is what I am hoping to provide! I should work well for T2V or I2V.
Thank you. I’m the kind of new WAN user you mentioned. It’s quite tough learning to operate it on my own. I’ve tried generating videos using the Smooth Mix base model, yet the outputs I get are vastly inferior to yours. How come your footage is so smooth? Could you share which base model you’re using?
My first guess would be your resolution is too low. If you're using this workflow, interpolation should be on. Set it to 60 fps if you have not. Try those things first.
The base model is generally Wan 2.2 14b, or if you're using the Smooth Mix workflow, just the normal one. There's no magic LoRA here or anything. All my videos tag as many things as I can.
@boobkake22 Thank you, my friend! I finally got everything running correctly.
I have a question about the facial cumshot videos. I'm using your LoRA together with your prompts, but I can't seem to get any camera cuts. The entire video stays on the same shot from beginning to end.
Your videos, on the other hand, switch between multiple camera angles and shots, and they look amazing. Could you tell me how you achieve those shot transitions? Is there a specific workflow, setting, or prompt that makes the camera change during the video?
@sqdhl That's all related to sh00tz. The weights matter. sh00tz needs to be near full strength do do what it does. I'd generally keep strength at 1.0 if that's what you want.
That killed my confyui installation. Nothing works anymore. Missing notes all over the flow.
A workflow cannot kill your ComfyUI install. You had to have changed something else, and I'd need more info to guess.
@boobkake22 huge hot take buddy. If a node you used suddenly releases an update that also updates a vital component in the dependency chain, it can break a LOT of stuff. Your workflow also broke my ComfyUI python environment when installing node packs that you use. One of them forcibly updates numpy to a post 2.x version. You should run a review over your workflow, check nodes for updates, and try to identify the one causing this.
@haroldjimst It's not a hot take. I am not responsible for your installation. If you want to update nodes you can. My workflow doesn't force that to happen. Manager warns you about conflicts. You have to manage your install. There are a ton of versions of Comfy and a ton of add-ons, and an individual developer cannot account for all of them.
@boobkake22 It's not about, that your workflow will not work, it's a arning to others, that if you install all the components the possibility is high, that you break your installation. I have an backup, so that is not that dramatically, but very annoying. And because you layered various nodes that are missing from my installation, it’s difficult to track them down. 50% of the nodes used were unrecognized by my installation, and a couple of the ones you used no longer exist in the current ComfyUI. After two hours of trial and error, I just lost interest.
think you mean you updated your nodes bro :) its common for nodes to disconnect... just reconnect them,
@volkersauber550 if youre struggling to find the nodes that dont work go to settings and click on Modern Node Design (Nodes 2.0) to turn it on... before you even run a workflow itll highlight the node in red. so its easier. if you dont have that on then you have to understand the log details which give node codes... easier to just turn the 2.0 on then off once done.
@Hiitpoint
Thats what i normally do, But when—like in this flow—three different nodes are stacked on top of each other, and the one at the very bottom is the one you need to edit, it’s just incredibly tedious and no fun at all—especially since that node appears 30 times in the flow, and every single time you have to move all the ones above it out of the way first (and they’re pinned, too). Absolutely no fun!