Commit Graph
6 Commits
Author SHA1 Message Date
dseditorandClaude a90f71066c Add Qwen GPU Inference node with intelligent memory management
- Add QwenGPUInference node for AI photo prompt optimization
- Implement smart GPU memory management with automatic detection
- Support CPU offload when GPU memory is insufficient
- Auto-download model config files from HuggingFace
- Remove <think> tags from model output
- Add bilingual (Chinese/English) support
- Remove deprecated GGUF inference node and related files
- Update README with comprehensive documentation

Features:
- Automatic model detection (qwen_3_4b.safetensors)
- Three loading strategies: Full GPU / CPU Offload / CPU-only
- Memory conflict prevention with ComfyUI models
- Professional photography prompt generation
- Default max_tokens: 2048 for detailed prompts
- Custom system prompt for photo optimization

Performance:
- Full GPU: ~26-30 tokens/second
- CPU Offload: ~1-2 tokens/second (reliable fallback)
- First load: 7-130 seconds depending on hardware
- Subsequent loads: Near-instant (model cached)

🤖 Generated with [Claude Code](https://claude.com/claude-code)

Co-Authored-By: Claude <noreply@anthropic.com>
2025-12-04 11:15:12 +08:00
dseditor 434cc205c8 add llm nodes 2025-09-02 22:50:25 +08:00
dseditor 25eb1ad213 Adding Split Fade 2025-07-06 01:15:44 +08:00
dseditor 112c39b277 Readme 2025-06-21 16:58:54 +08:00
dseditor af9bec2e00 Readme 2025-06-21 16:55:13 +08:00
dseditor 026ffa3b47 Initial commit for ComfyUI-ListHelper 2025-06-21 16:47:32 +08:00