ComfyUI-SelVA

Files

T

Ethanfel e56ece9c1c feat: add SelVA Textual Inversion Trainer and Loader nodes

Learns K CLIP token embeddings ([K, 1024]) with all model weights frozen,
keeping generated latents on the decoder's natural manifold — avoids the
quality degradation that affects LoRA on BJ's audio dataset.

- selva_textual_inversion_trainer.py: trains learned_tokens via AdamW,
  injects into last K positions of 77-token CLIP embedding, checkpoints
  with eval audio + spectral metrics
- selva_textual_inversion_loader.py: loads .pt bundle, returns
  TEXTUAL_INVERSION dict for sampler
- selva_sampler.py: optional textual_inversion input; injects into both
  text_clip and neg_text_clip before preprocess_conditions
- __init__.py: registers both new nodes

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>

2026-04-08 23:01:44 +02:00

__init__.py

feat: add SelVA Textual Inversion Trainer and Loader nodes

2026-04-08 23:01:44 +02:00

selva_audio_preprocessors.py

feat: add SelVA HF Smoother and Spectral Matcher preprocessing nodes

2026-04-08 20:28:16 +02:00

selva_dataset_browser.py

feat: thorough overnight sweep + dataset browser updates

2026-04-08 00:38:19 +02:00

selva_feature_extractor.py

fix: set OUTPUT_NODE=True on SelVA Feature Extractor so it runs without connected outputs