- PromptRefiner unload_model toggle now actually controls eviction:
keep_alive is "0s" when on, None (server default) when off, so the
UI no longer unloads against the user's choice.
- Cleanup always runs; unload fires only when unload_model is on AND the
subprocess fallback left a model loaded, instead of skipping cleanup
entirely when the toggle is off.
- release_vram() clears empty_cache() on every CUDA device, not just
device 0, for multi-GPU rigs.
- Add docstrings to INPUT_TYPES methods and language ids to plan doc
code fences.