Update README.md
This commit is contained in:
@@ -1,6 +1,8 @@
|
||||
# ComfyUI DiaTest TTS Node
|
||||
|
||||
This node pack integrates the [Nari Labs Dia](https://github.com/nari-labs/dia) text-to-speech model into ComfyUI using a single node for loading (onto GPU) and generation.
|
||||
Warning LLM Code there are probably better options
|
||||
|
||||
This node pack partially integrates the [Nari Labs Dia](https://github.com/nari-labs/dia) text-to-speech model into ComfyUI using a single node for loading (onto GPU) and generation.
|
||||
|
||||
Dia allows generating dialogue with speaker tags (`[S1]`, `[S2]`) and non-verbal sounds (`(laughs)`, etc.). This node loads the model from Hugging Face Hub and generates audio directly using float32 precision. It requires a CUDA-enabled GPU.
|
||||
|
||||
@@ -57,4 +59,4 @@ Loads the specified Dia model from Hugging Face Hub onto the GPU (if not already
|
||||
* The first time you run the node for a specific `repo_id`, it will download the model files from Hugging Face Hub, which may take some time. Subsequent runs will use the cached model.
|
||||
* Changing the `repo_id` will trigger a model reload.
|
||||
* The model uses float32 precision internally.
|
||||
* The Descript Audio Codec (DAC) dependency (`descript-audio-codec`) must be installed via `requirements.txt`.
|
||||
* The Descript Audio Codec (DAC) dependency (`descript-audio-codec`) must be installed via `requirements.txt`.
|
||||
|
||||
Reference in New Issue
Block a user