CogVideoX-Fun Model Setup Guide
a. Model Links and Storage Locations
Required Files:
V1.5:
| Name | Storage Space | Hugging Face | Model Scope | Description |
|---|---|---|---|---|
| CogVideoX-Fun-V1.5-5b-InP | 20.0 GB | 🤗Link | 😄Link | Our official graph-generated video model is capable of predicting videos at multiple resolutions (512, 768, 1024) and has been trained on 85 frames at a rate of 8 frames per second. |
| CogVideoX-Fun-V1.5-Reward-LoRAs | - | 🤗Link | 😄Link | The official reward backpropagation technology model optimizes the videos generated by CogVideoX-Fun-V1.5 to better match human preferences. | |
V1.1:
| Name | Storage Space | Hugging Face | Model Scope | Description |
|---|---|---|---|---|
| CogVideoX-Fun-V1.1-2b-InP | 13.0 GB | 🤗Link | 😄Link | Our official graph-generated video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. |
| CogVideoX-Fun-V1.1-5b-InP | 20.0 GB | 🤗Link | 😄Link | Our official graph-generated video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. Noise has been added to the reference image, and the amplitude of motion is greater compared to V1.0. |
| CogVideoX-Fun-V1.1-2b-Pose | 13.0 GB | 🤗Link | 😄Link | Our official pose-control video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. |
| CogVideoX-Fun-V1.1-2b-Control | 13.0 GB | 🤗Link | 😄Link | Our official control video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. Supporting various control conditions such as Canny, Depth, Pose, MLSD, etc. |
| CogVideoX-Fun-V1.1-5b-Pose | 20.0 GB | 🤗Link | 😄Link | Our official pose-control video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. |
| CogVideoX-Fun-V1.1-5b-Control | 20.0 GB | 🤗Link | 😄Link | Our official control video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. Supporting various control conditions such as Canny, Depth, Pose, MLSD, etc. |
| CogVideoX-Fun-V1.1-Reward-LoRAs | - | 🤗Link | 😄Link | The official reward backpropagation technology model optimizes the videos generated by CogVideoX-Fun-V1.1 to better match human preferences. | |
V1.0:
| Name | Storage Space | Hugging Face | Model Scope | Description |
|---|---|---|---|---|
| CogVideoX-Fun-2b-InP | 13.0 GB | 🤗Link | 😄Link | Our official graph-generated video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. |
| CogVideoX-Fun-5b-InP | 20.0 GB | 🤗Link | 😄Link | Our official graph-generated video model is capable of predicting videos at multiple resolutions (512, 768, 1024, 1280) and has been trained on 49 frames at a rate of 8 frames per second. |
Storage Location:
📂 ComfyUI/
├── 📂 models/
│ └── 📂 Fun_Models/
| ├── 📂 CogVideoX-Fun-V1.1-2b-InP/
| ├── 📂 CogVideoX-Fun-V1.1-5b-InP/
| ├── 📂 CogVideoX-Fun-V1.5-5b-InP/
│ └── 📂 CogVideoX-Fun-V1.1-5b-Control/
b. Node types
- LoadCogVideoXFunModel
- Loads the CogVideoX-Fun model
- FunTextBox
- Write the prompt for CogVideoX-Fun model
- CogVideoXFunInpaintSampler
- CogVideoX-Fun Sampler for Image to Video
- CogVideoXFunT2VSampler
- CogVideoX-Fun Sampler for Text to Video
- CogVideoXFunV2VSampler
- CogVideoX-Fun Sampler for Video to Video
c. ComfyUI Json Workflows
i. Video to video generation
Download link for v1.5.
Download link for v1.1.
You can run the demo using following video: demo video
ii. Image to video generation
Download link for v1.5.
Download link for v1.1.
You can run the demo using following photo:

iii. Text to video generation
Download link for v1.5.
Download link for v1.1.
iv. Control video generation
Download link for v1.1.
You can run the demo using following video: demo video
v. Lora usage.
Download link for v1.1.




