Wan 2.2 Animate in ComfyUI: Workflow, Files and Custom Nodes
Wan 2.2 Animate in ComfyUI: the native template, every file it loads with exact sizes, the three custom nodes it needs, GGUF builds and the Kijai wrapper route.
Quick answer
The official Wan 2.2 Animate ComfyUI workflow is a built-in template, video_. You give it one character image and one video of a person moving. It either animates your character with that person's motion, or puts your character into the video in that person's place.
What it takes to run the official template:
- Six model files, 28.83 GB in total. The largest is the 18.40 GB FP8 Animate model from Kijai's repository, not from Comfy-Org.
- Three custom node packs. The template's own note names
comfyui_,controlnet_ aux ComfyUI-andKJNodes ComfyUI-. ComfyUI's tutorial lists only the first two.segment- anything- 2 - Three preprocessing models, 0.68 GB, which those node packs download on the first run: two DWPose files and one SAM2 file.
Two things the official tutorial does not tell you:
- The template loads a second LoRA,
WanAnimate_(1.44 GB). The tutorial's download list leaves it out.relight_ lora_ fp16 .safetensors - Kijai has replaced the FP8 file the template asks for. His note on the repository says the first upload gives grid-pattern noise in native ComfyUI, and a
_file fixes it. The template and tutorial still point at the first one.v2
We read the template, ComfyUI's tutorial, Wan's README and paper and the node packs' source on 2026-10-09. Sizes are from the Hugging Face API, in decimal gigabytes. We have not run Wan 2.2 Animate ourselves.
What Wan 2.2 Animate does
Wan2.2-Animate-14B is a separate 14B model from the Wan team, released on 2025-09-19. Its model card lists Wan2.2-I2V-A14B as the base model, and its paper is titled Wan-Animate: Unified Character Animation and Replacement with Holistic Replication. Wan's README and the paper name two modes:
| Wan's name | ComfyUI template's name | What you get |
|---|---|---|
| Wan's nameAnimation mode | ComfyUI template's nameMove | What you getYour character image, moving the way the person in the video moves. The video's background is not used. |
| Wan's nameReplacement mode | ComfyUI template's nameMix | What you getThe original video with the person swapped for your character, relit to match the scene's lighting and color. |
Per the paper, body motion is carried by a skeleton extracted from the video and the expression by features taken from face crops. Replacement mode also needs a mask of the person, and it uses the relighting LoRA to blend your character into the scene.
That is why Animate needs more preprocessing than other Wan 2.2 templates. Before the model runs, the video has to become a pose video, a face video and, for replacement, a mask and a background video with the person blacked out.
Wan-AI has since published a separate model, Wan2.2-, released on 2026-08-07 by its README. This page covers the original Animate-14B, which is what the ComfyUI template loads.
Every file the native template loads
From the models metadata inside video_, checked against each repository's file list.
Folder under ComfyUI/ | File | Repository | Bytes | Size |
|---|---|---|---|---|
Folder under ComfyUI/models/diffusion_ | FileWan2_ | RepositoryKijai/ | Bytes18,401,760,586 | Size18.40 GB |
Folder under ComfyUI/models/loras/ | Filelightx2v_ | RepositoryKijai/ | Bytes738,005,744 | Size0.74 GB |
Folder under ComfyUI/models/loras/ | FileWanAnimate_ | RepositoryKijai/ | Bytes1,436,672,440 | Size1.44 GB |
Folder under ComfyUI/models/clip_ | Fileclip_ | RepositoryComfy- | Bytes1,264,219,396 | Size1.26 GB |
Folder under ComfyUI/models/text_ | Fileumt5_ | RepositoryComfy- | Bytes6,735,906,897 | Size6.74 GB |
Folder under ComfyUI/models/vae/ | Filewan_ | RepositoryComfy- | Bytes253,815,318 | Size0.25 GB |
Total: 28,830,380,381 bytes, 28.83 GB. The tutorial's folder tree says clip_; the real folder is clip_, as our Wan 2.2 models folder page explains. The note inside the template gives sizes such as "17.14 GB" for the model; those are binary gibibytes of the same file.
Other versions of the Animate model you may see linked:
| File | Repository | Bytes | Size |
|---|---|---|---|
FileWan2_ | RepositoryKijai/ | Bytes17,317,143,060 | Size17.32 GB |
FileWan2_ | RepositoryKijai/ | Bytes18,401,760,586 | Size18.40 GB |
FileWan2_ | RepositoryKijai/ | Bytes17,317,143,060 | Size17.32 GB |
Filewan2.2_ | RepositoryComfy- | Bytes34,549,787,368 | Size34.55 GB |
Kijai's note on the v2 files: the first uploads kept the face encoder layers in bf16. ComfyUI's native loader then casts them down to plain FP8, which he says puts noise in a grid pattern into the output. The v2 files store those layers as scaled FP8 too. For his own wrapper, he says the first files are still fine.
So in the native template, download the _ e4m3fn file and pick it in the Load Diffusion Model node. The e5m2 files are the same models in the other FP8 format.
The custom nodes and what they download
The template's "Make sure these custom nodes are installed" note lists three packs. ComfyUI-Manager's Install missing nodes button finds them from the red nodes.
| Pack | Nodes the template uses | What it does here |
|---|---|---|
PackFannovel16/ | Nodes the template usesDWPreprocessor ×2, PixelPerfectResolution | What it does hereTurns the video into pose frames. One DWPose node is set to body and hands only, the other to face only |
Packkijai/ | Nodes the template usesDownloadAndLoadSAM2Model, Sam2Segmentation | What it does hereCuts the person out of the video as a mask, from the points you place |
Packkijai/ | Nodes the template usesPointsEditor, BlockifyMask, DrawMaskOnImage | What it does herePointsEditor is where you click the person; the other two shape the mask and black out the person in the background video |
The sampler node, WanAnimateToVideo, is built into ComfyUI. If it shows red, update ComfyUI itself; a custom node will not fix it.
On the first run, two of those packs download their own models:
| File | Downloaded from | Saved to | Bytes | Size |
|---|---|---|---|---|
Fileyolox_ | Downloaded fromyzd- | Saved tocustom_ | Bytes216,746,733 | Size0.22 GB |
Filedw- | Downloaded fromhr16/ | Saved tocustom_ | Bytes135,059,124 | Size0.14 GB |
Filesam2_ | Downloaded fromKijai/ | Saved tomodels/ | Bytes323,407,992 | Size0.32 GB |
The ckpts path is controlnet_aux's default and can be changed in its config. On a machine that cannot reach Hugging Face while ComfyUI runs, put these files in place by hand first.
How the template runs
The defaults, read from the template's widgets:
- Size 640 × 640. Width and height must be multiples of 16, a limit of
WanAnimateToVideo. The template's note says the small default is there to avoid running out of VRAM. - 77 frames per segment at 16 fps, about 4.8 seconds. Each copy of the "Video Extend" subgraph adds another 77 frames. For a longer clip, copy it again and link
batch_andimages video_from the previous one.frame_ offset - 6 steps at CFG 1, with both LoRAs at strength 1. The lightx2v LoRA is what makes so few steps work.
- An Image Scale node resizes the input video to the same width and height before preprocessing. The template's note warns that a large video takes a very long time to preprocess.
The template opens in Mix mode. For Move, disconnect the background_ and character_ inputs from the sampling subgraph. The template's note warns that bypassing those nodes is not enough, because a bypassed node still passes the video through.
The PointsEditor canvas is empty until it has the video's first frame. Run the workflow once, or load the frame yourself, then Shift-click to place points: left for the person, right for areas to exclude.
GGUF builds of Wan 2.2 Animate
QuantStack/ has the Animate model in 11 quantisations. Its card describes it as a direct conversion of the official model.
| Quant | Bytes | Size |
|---|---|---|
| QuantQ2_K | Bytes6,457,431,872 | Size6.46 GB |
| QuantQ3_K_S | Bytes7,969,675,072 | Size7.97 GB |
| QuantQ3_K_M | Bytes8,630,769,472 | Size8.63 GB |
| QuantQ4_0 | Bytes10,402,699,072 | Size10.40 GB |
| QuantQ4_K_S | Bytes10,592,753,472 | Size10.59 GB |
| QuantQ4_K_M | Bytes11,496,331,072 | Size11.50 GB |
| QuantQ5_K_S | Bytes12,349,118,272 | Size12.35 GB |
| QuantQ5_0 | Bytes12,526,065,472 | Size12.53 GB |
| QuantQ5_K_M | Bytes13,003,659,072 | Size13.00 GB |
| QuantQ6_K | Bytes14,605,195,072 | Size14.61 GB |
| QuantQ8_0 | Bytes18,719,217,472 | Size18.72 GB |
The card says to put the file in ComfyUI/ and load it with city96's ComfyUI-GGUF node pack. In the native template, that means replacing the Load Diffusion Model node with Unet Loader (GGUF). Everything else in the template stays: the LoRAs, CLIP Vision, text encoder, VAE and the preprocessing. The repository also carries a community example workflow, credited to a Discord user. Note that Q8_0 is larger than Kijai's FP8 files.
The Kijai WanVideoWrapper route
The alternative to the native template is Kijai's ComfyUI-, which uses its own loader and sampler nodes. It ships two Animate examples in example_:
wanvideo_does the same preprocessing as the native template: DWPose, SAM2 and the KJNodes points editor.WanAnimate_ example_ 01 .json wanvideo_uses Kijai'sWanAnimate_ preprocess_ example_ 02 .json ComfyUI-pack instead. It runs ViTPose and a YOLO detector, the same kind of models as Wan's own preprocessing script.WanAnimatePreprocess
Both load the same Animate FP8 file and the same two LoRAs. They differ from the native template in the text encoder and VAE:
| File | Repository | Bytes | Size |
|---|---|---|---|
Fileumt5- | RepositoryKijai/ | Bytes11,361,845,464 | Size11.36 GB |
FileWan2_ | RepositoryKijai/ | Bytes253,806,278 | Size0.25 GB |
Fileyolov10m (preprocess example only) | RepositoryWan- | Bytes61,659,339 | Size0.06 GB |
Filevitpose_ + _ (preprocess example, Huge) | RepositoryKijai/ | Bytes2,549,378,992 | Size2.55 GB |
The WanAnimatePreprocess README says its models go in ComfyUI/ for now, and that a Large ViTPose model from JunkyByte/ also works. The Huge ViTPose is two files that must sit in the same folder.
The first example ships with block swap set to 25 blocks, which moves part of the model to system RAM, and with sageattn as the attention mode. Kijai has said the only black Wan output he had seen was SageAttention-related; if you get black frames, see our black video page.
How much VRAM Wan 2.2 Animate needs
Nobody official has published a figure. Wan's README gives an 80 GB statement for several of its other 14B scripts, but none for Animate. ComfyUI's tutorial says only to start at a small size in case you do not have enough VRAM.
What the file sizes say, as arithmetic and not as a measurement:
- The template's FP8 file is 18.40 GB and the v2 file 17.32 GB. Neither fits on a 16 GB card by size. They can still run through ComfyUI's weight streaming or Kijai's block swap, which move part of the model to system RAM.
- The Q4_K_M GGUF is 11.50 GB. It fits a 16 GB card with room to spare, and a 12 GB card with very little.
- Animate is one model, not two experts. Unlike the 14B text- and image-to-video templates, there is no second file to swap in. Our Wan 2.2 VRAM page covers how those compare.
- Preprocessing adds its own load. The template loads SAM2 on
cudain fp16, and it runs before the video model does.
For a sense of how far streaming goes on a 12 GB card, our MiniMax H3 run on an RTX 3060 12GB peaked at 11,649 MiB of VRAM and 43,587 MiB of system RAM; that is a different model, recorded in our RTX 3060 test card.
What nobody has published yet
- An official minimum VRAM figure for Wan 2.2 Animate in ComfyUI or in Wan's own script.
- A side-by-side of the first and v2 FP8 files in the native template, beyond Kijai's note.
- A measured peak VRAM and system RAM figure for the native template at its 640 × 640 default.
Licence and downloads
Wan2.2-Animate-14B is released under the Apache 2.0 licence. Wan's README adds use restrictions on illegal and harmful content; read it before building on the model. The other files and nodes, as each source states it:
- Apache 2.0 tag:
Comfy-,Org/ Wan_ 2.2_ ComfyUI_ Repackaged Comfy-,Org/ Wan_ 2.1_ ComfyUI_ repackaged Kijai/,WanVideo_ comfy_ fp8_ scaled QuantStack/,Wan2.2- Animate- 14B- GGUF Kijai/,sam2- safetensors Kijai/,vitpose_ comfy yzd-andv/ DWPose hr16/. QuantStack's card adds that the original model's terms still apply.DWPose- TorchScript- BatchSize5 - No licence tag:
Kijai/, the repository both LoRAs and the wrapper's text encoder come from, andWanVideo_ comfy JunkyByte/. Check their pages before using those files commercially.easy_ ViTPose - Custom nodes:
comfyui_,controlnet_ aux ComfyUI-,segment- anything- 2 ComfyUI-,WanVideoWrapper ComfyUI-andWanAnimatePreprocess ComfyUI-are Apache 2.0.GGUF ComfyUI-is GPL-3.0.KJNodes - Wan's own script can optionally use FLUX.1-Kontext-dev for pose retargeting. That model carries the FLUX.1 [dev] non-commercial licence. The ComfyUI workflows here do not use it.
We do not host any of these files. Download them from the repositories named above. For other templates and how to read a workflow before you queue it, see our workflows page.
To check your GPU and system RAM against a local video model, use the system requirements checker. Wan 2.2 Animate is not one of its presets yet; only MiniMax H3 is.
GenVidKit is an independent guide. It is not affiliated with Alibaba, the Wan team, Comfy Org, Kijai, QuantStack, Hugging Face or the authors of the custom nodes named here.
Sources
All read on 2026-10-09.
- Wan2.2 Animate, ComfyUI documentation — the template, its download list, the two custom nodes it names, Mix and Move, size and extend instructions.
- Comfy-Org/workflow_templates:
video_— the files, nodes, node packs and default settings the template actually uses, and its notes.wan2_ 2_ 14B_ animate .json - Wan-Video/Wan2.2 on GitHub and its preprocessing guide — the two modes, preprocessing outputs, FLUX use, licence.
- Wan-Animate paper, arXiv 2509.14055 — how motion, expression and relighting work.
- Wan-AI/Wan2.2-Animate-14B and Wan-AI/Wan2.2-Animate-2-14B — licence tags, the YOLO file, the later model's release date.
- Kijai/WanVideo_comfy_fp8_scaled, Wan22Animate — FP8 files, byte counts and the v2 note.
- Kijai/WanVideo_comfy, Comfy-Org/Wan_2.2_ComfyUI_Repackaged and Comfy-Org/Wan_2.1_ComfyUI_repackaged — LoRA, encoder and VAE byte counts.
- QuantStack/Wan2.2-Animate-14B-GGUF — GGUF byte counts, folder and loader.
- kijai/ComfyUI-WanVideoWrapper and kijai/ComfyUI-WanAnimatePreprocess — the wrapper's two Animate examples and the preprocessing models.
- Fannovel16/comfyui_controlnet_aux and kijai/ComfyUI-segment-anything-2 — where the DWPose and SAM2 models download from and to.
- yzd-v/DWPose, hr16/DWPose-TorchScript-BatchSize5, Kijai/sam2-safetensors and Kijai/vitpose_comfy — preprocessing model byte counts and licence tags.
- kijai/ComfyUI-KJNodes and city96/ComfyUI-GGUF — node pack licences and the GGUF loader.