This workflow contains system prompts of LLM prompt enhancer for all T2VA, I2VA, L2VA, FL2VA and Ref2VA.
It supports local gguf LLM models. e.g. Gemma 4, Qwen 3.6
No API needed.
Purpose: To make user prompt conform with MiniMax official prompting guide.
Choose a big MoE LLM model (i.e. something with -A?B) instead of a small non-MoE model.
Set "cpu_moe": true to speed up LLM with limited vram. (Q8 Gemma 4 26B-A4B model needs only 6GB vram.)
To set image_min_tokens for Gemma 4, follow n_ubatch > image_max_tokens > image_min_tokens. (e.g. 2240, 2240, 560)
Qwen 3.6 (e.g. Q6 35B-A3B) can also be used instead of Gemma 4. Just set image_min_tokens to 1024 and n_ctx to a larger value (e.g. 16384).
Custom node used:
ComfyUI_Simple_Qwen3-VL-gguf
ComfyUI_Simple_Qwen3-VL-gguf requires installation of llama-cpp-python wheel and Nvidia CUDA Toolkit. (Comfyui's built-in CUDA might not work.)
e.g.
cmd
cd /d C:\StabilityMatrix\Data\Packages\ComfyUI\venv\Scripts\
python.exe -m pip install ???.whl
ComfyUI-SolAttn_triton
System prompts modified from:
H3_LLM_Instructions
rzgar
Thanks for their work. Cheers.
Description
FAQ
Comments (3)
I tried referencing a video in this workflow, but it stops here with an error.
https://files.catbox.moe/35rq2y.png
If I disconnect "Get_input_vid1" and try to reconnect it, the connection fails. It seems it doesn't support image-type inputs; is there any workaround for this?
I personally haven't experienced this error before. The custom node has always been accepting image-type videos. Maybe you could ask the author. He/she's really nice and has helped me alot with the usage of the node in the issue page. Cheers.
@cloudreadypc
I came across a post in the Issues section that appeared to be yours. I tried something based on the ideas there, and it worked.
The issue was due to the version. It seems that version 3.9 is installed automatically, but it doesn't accept input; switching to version 3.9-nightly resolved the problem. Thank you.


