main
ComfyUI-MiniCPM
English | 简体中文
This is a custom node for ComfyUI that utilizes the multimodal capabilities of MiniCPM-o.
The node's functionality is still being expanded, with the goal of implementing real-time audio and video capabilities in ComfyUI to create interesting and practical applications.
Currently supported model version: MiniCPM-o 2.6 (Released January 2024)
Features
Single Image i2t Prompt Inference
You can choose from preset prompts or input your own prompts.
Multi-Image i2t Prompt Inference
Outputs combined prompts from multiple images.
Installation
Method 1: Using ComfyUI Manager (Recommended)
- Install ComfyUI Manager in ComfyUI
- Open ComfyUI and click on the "Manager" tab in the top right
- Search for "MiniCPM-o" in the search box
- Click the install button to complete installation
Method 2: Manual Installation
- Clone this repository into your ComfyUI custom_nodes folder:
cd ComfyUI/custom_nodes
git clone https://github.com/CY-CHENYUE/ComfyUI-MiniCPM-o.git
- Install dependencies using ComfyUI's Python:
..\..\..\python_embeded\python.exe -m pip install -r requirements.txt
Installation Guide
-
Download Model Files
- Download MiniCPM-o 2.6 model files from Hugging Face Repository
-
Place Model Files
- Put the downloaded model files in ComfyUI's model directory:
ComfyUI └── models └── MiniCPM └── MiniCPM-o-2_6 ├── image_processing_minicpmv.py ├── configuration_minicpm.py ├── modeling_minicpmo.py └── other model files... -
Model File Structure
- Ensure all necessary files are in the model directory
- Do not modify the file structure or filenames
-
Remember to install dependencies using ComfyUI's Python
Contact Me
- X (Twitter): @cychenyue
- TikTok: @cychenyue
- YouTube: @CY-CHENYUE
- BiliBili: @CY-CHENYUE
- Xiaohongshu: @CY-CHENYUE
Languages
Python
100%


