Etherll/Timbre
Extract a target speaker’s clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.
Commercial license
✓ Commercial OKApache-2.0
可商用,通常只需保留著作權聲明/授權條款
Topics
audio-processingljspeechnemopythonpython3speaker-diarizationspeaker-verificationspeech-separationtts-datasetvoice-cloningwhisper
Ad