thomcle/chatterbox-tts
Convert text to speech with optional zero-shot voice cloning from a short reference audio clip. Accepts text and an opti...
Found 104 models (showing 21-40)
Convert text to speech with optional zero-shot voice cloning from a short reference audio clip. Accepts text and an opti...
Convert spoken audio to a target speakerβs voice using a reference sample. Provide a source audio file (content) and a r...
Clone a voice from a short reference clip and generate speech from text. Accepts text and a reference audio sample; outp...
Performs zero-shot speech editing and text-to-speech synthesis using audio input and text transcripts. Supports four mai...
Synthesize speech from text in a cloned voice using a reference audio sample. Provide a text prompt and speaker referenc...
Generate expressive multilingual speech from text. Accept a text prompt and a language selection (ar, da, de, el, en, es...
Generate English speech from text with zero-shot voice cloning from a 5β10s reference clip. Provide text, a short refere...
Generate speech from text with optional voice cloning from a reference voice sample. Accepts a text prompt and optional...
Convert text to expressive speech, with optional speaker style cloning from a short reference audio. Accepts text input...
Create song covers by cloning voices from audio files using RVC v2 technology. Takes an input audio file and transforms...
Generate Spanish speech from text by cloning the voice from a reference audio. Provide Spanish text, a reference audio s...
Synthesizes speech from text using a reference audio file and its transcript to clone the speaker's voice. Takes text in...
Generate speech from text and convert voices. Use zero_shot voice cloning to synthesize speech in the style of a prompt_...
Clone voices from audio samples with ultra-low latency streaming synthesis. Supports zero-shot voice cloning, cross-ling...
Generate multilingual speech from text with zero-shot voice cloning. Provide a short reference audio clip and its transc...
Converts text to speech using a speaker reference audio file to clone the voice characteristics and speaking style of th...
Generate spoken audio from text. Clone a target voice by providing a prompt audio sample (voice_cloning mode), or synthe...
Generates speech from text input using a reference speaker audio and corresponding text. Takes text to synthesize, a ref...
Clone a voice from a short reference audio and synthesize speech from text. Provide at least 6 seconds of speaker audio...
Convert speech to a target voice using RVC v2 voice models. Takes an input speech audio clip and outputs converted audio...