fe9b975866f7a38f7f3b8a0e2b2f1bfa3749b9ac
ComfyUI ExLlama Nodes
A simple prompt generator for ComfyUI utilizing ExLlama.
Installation
Clone the repository to custom_nodes in your ComfyUI directory and install the dependencies:
git clone https://github.com/Zuellni/ComfyUI-ExLlama-Nodes
python -m pip install -r requirements.txt
If you see any errors related to ExLlama while loading the nodes, you should manually install the wheel matching your system from here.
For example, on Windows with Python 3.11 and PyTorch CUDA 12.1, you would use:
python -m pip install https://github.com/jllllll/exllama/releases/download/0.0.17/exllama-0.0.17+cu121-cp311-cp311-win_amd64.whl
Nodes
| Name | Description |
|---|---|
| Loader | Loads 4-bit GPTQ Llama/2 models. You can find a lot of them on Hugging Face. Clone the model repository or download all the files in it and place them in an empty directory, then specify the path in model_dir. The model.safetensors file won't work on its own.ExLlama allocates memory based on max_seq_len. Lowering it is a good way to save on VRAM. It's currently not possible to offload the model to RAM. |
| Generator | Returns a string based on the given prompt for use with other nodes. Default values correspond to the simple-1 preset from text-generation-webui. ExLlama isn't deterministic, so the outputs may differ slightly even with the same seed.To load a LoRA specify the path to its directory in lora_dir. It should contain adapter_model.bin and adapter_config.json. |
| Previewer | Displays generated outputs in the UI. |
Workflow
The workflow below can be loaded directly in ComfyUI. Model used: MythoLogic-Mini-7B.
Languages
Python
89.9%
JavaScript
10.1%