ComfyUI-Omnivoice

Author	SHA1	Message	Date
Ethanfel	772f6654d4	feat: add language selector for voice_design + Chinese instruct support - Generate: language dropdown (auto/English/Chinese), passed only in voice_design and auto_voice modes where it selects the instruct vocab - VoiceDesign: Chinese mode with dialect/age/pitch/gender dropdowns using the model's validated Chinese instruct vocabulary (全角逗号) Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 20:22:25 +02:00
Ethanfel	e26bac3684	remove: language parameter from Generate (model auto-detects from text) Language is inferred from the text content — the parameter had no effect. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 20:21:27 +02:00
Ethanfel	194e0b0e09	fix: trim accent list to model-validated values only The model's _resolve_instruct() validates against a fixed vocabulary. Only 10 accents are supported — removed all unsupported additions. Updated tooltip to reflect actual constraints. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 20:17:59 +02:00
Ethanfel	d4bf7c825e	feat: pass instruct in voice_cloning mode for accent/style influence If instruct is set alongside ref_audio, it is now forwarded to model.generate() — allowing accent/style transfer on top of the cloned voice identity. Model may or may not honour both simultaneously. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 20:07:18 +02:00
Ethanfel	d2cb5c4249	feat: expand language and accent lists to full coverage Language: ~170 world languages with type-to-filter dropdown Accent: 50+ regional varieties grouped by area Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 20:04:12 +02:00
Ethanfel	c1558efad9	feat: add Voice Design node + language and guidance_scale to Generate OmniVoiceVoiceDesign: structured dropdowns for gender/age/pitch/accent that compose into an instruct string — wire to Generate's instruct input. OmniVoiceGenerate: new optional language dropdown (auto + 11 languages) and guidance_scale (CFG, default 2.0) parameters. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 20:02:06 +02:00
Ethanfel	8805665a22	Add seed parameter to OmniVoice Generate for consistent voice across chunks OmniVoice chunks long text internally; each chunk is a separate diffusion pass with different random noise, causing voice drift between paragraphs. Setting the same seed before each generate() call anchors the RNG state and keeps the voice consistent. seed=0 means random (default behaviour). Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 18:53:58 +02:00
Ethanfel	8d77dd6cd5	Remove torchcodec workaround; recommend Whisper node for ref_text Users should connect a ComfyUI Whisper node to ref_text instead of relying on omnivoice's internal ASR. Removes the error-catch workaround and updates the tooltip accordingly. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 17:49:25 +02:00
Ethanfel	30f46fc3ef	Revert transformers cap; catch torchcodec ASR failure with clear message install.py: restore transformers>=5.0.0 (capping it would break other nodes). generator.py: catch the torchcodec RuntimeError that fires when ref_text is blank and transformers 5.x auto-transcription requires missing FFmpeg libs. Raises a human-readable error telling the user to fill in ref_text manually. Also updates the ref_text tooltip to recommend providing it explicitly. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 17:35:54 +02:00
Ethanfel	5dfaa0b300	Replace torchaudio.save with soundfile.write; add EPUB loader node - nodes/generator.py: swap torchaudio.save for soundfile.write to avoid torchcodec/FFmpeg dependency crash in environments without FFmpeg shared libs - nodes/epub_loader.py: new OmniVoiceEpubLoader node for loading EPUB chapters - tests/test_epub_loader.py: 8 tests for the EPUB loader - install.py: add beautifulsoup4 to runtime deps - __init__.py, nodes/__init__.py: register OmniVoiceEpubLoader Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 17:24:18 +02:00
Ethanfel	5366f6992e	feat: add tooltips with inline tag reference to generator node inputs	2026-04-05 15:35:21 +02:00
Ethanfel	b273f8f2d7	refactor: remove redundant condition and rename shadowed waveform variable	2026-04-05 10:41:08 +02:00
Ethanfel	0ffd624471	fix: protect os.unlink in finally block from masking original exceptions	2026-04-05 10:38:38 +02:00
Ethanfel	a2c542a2bc	fix: move output waveform to CPU and cast sample_rate to int	2026-04-05 10:34:53 +02:00
Ethanfel	18fe6359cf	fix: add input validation and cpu() guard in OmniVoiceGenerate Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 09:09:52 +02:00
Ethanfel	95712e5504	feat: add OmniVoiceGenerate node with voice cloning, design, and auto modes Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-04-05 09:07:20 +02:00
Ethanfel	0ed43a83ca	feat: scaffold ComfyUI-Omnivoice node package	2026-04-05 08:43:17 +02:00

17 Commits