docs: clarify Qwen3 model type emotion support and UI priority

- Split Qwen3-TTS into CustomVoice (supports instruct) and
  Base (strips tags, doesn't support instruct) in compat table
- Add note: UI dropdown emotion takes priority over inline tags
This commit is contained in:
Hawk Lee
2026-02-16 10:35:47 +08:00
parent e01e7c3eff
commit 3ef9fd4601
+4 -1
View File
@@ -1167,13 +1167,16 @@ Script Parser → Emotion Annotator → Dialogue TTS / Qwen Dialogue TTS
| TTS 引擎 | 消费方式 | 效果 |
|----------|---------|------|
| **CosyVoice** | `[Happy] 文本...` 格式注入 | ✅ 精准情感控制 |
| **Qwen3-TTS** | 合并到 `instruct` 指令(按情感自动拆批) | ✅ 自动按情感分批生成 |
| **Qwen3-TTS (CustomVoice)** | 提取标签 → `instruct` 指令(按情感自动拆批) | ✅ 自动按情感分批生成 |
| **Qwen3-TTS (Base)** | 自动清洗标签(Base 模型不支持 instruct) | ✅ 标签被清洗,不影响克隆 |
| **VibeVoice** | 自动清洗 `[Emotion]` 标签后生成 | ✅ 无影响,不会朗读标签 |
> **Qwen3 智能分批**:当使用 Qwen3 Dialogue TTS 时,连续相同情感的句子会自动合并为一个批次,情感变化时自动拆分为新批次,确保每个批次的 `instruct` 只包含单一情感,语义准确。
> **VibeVoice 安全保障**:VibeVoice Standard / Realtime 节点在文本预处理阶段会自动清洗所有 24 种情感标签(如 `[Happy]`、`[Calm]`),因此即使通过 Podcast Splitter 传递了嵌标签的文本,VibeVoice 也不会将其作为普通文字朗读。
> **优先级规则**:如果用户在 Qwen3 节点的 UI 下拉框中手动选择了情感,该选择将**优先于**文本中的内嵌标签。标签仍会被清洗,但不会覆盖 UI 设置。
##### 典型工作流
**基础流程**(适合大多数场景):