x-lance/f5-tts
Synthesize speech from text in a cloned voice using a reference audio sample. Provide a text prompt and speaker referenc...
Found 19 models (showing 1-19)
Synthesize speech from text in a cloned voice using a reference audio sample. Provide a text prompt and speaker referenc...
Clone a target voice and generate speech audio from text. Provide a short speaker reference audio and its transcript (te...
Clone a voice from a short reference clip and generate speech from text. Accepts text and a reference audio sample; outp...
Converts text into speech using a reference speaker audio file and reference text. Takes a text input to be synthesized,...
Converts text to speech using a speaker reference audio file to clone the voice characteristics and speaking style of th...
Generate multilingual speech from text with zero-shot voice cloning. Provide a short reference audio clip and its transc...
Generate expressive speech from text with zero-shot voice cloning using a reference speaker audio input. Control emotion...
Converts text into spoken audio using the Parler TTS tiny 1.0 model. Takes text input and optional speaker reference aud...
Generate speech audio from text while cloning a target voice from a reference audio sample. Provide the text to speak, a...
Generate Hololive VTuber-style speech from text or convert a reference audio clip into those voices. Takes text input or...
Clone voices from audio samples with ultra-low latency streaming synthesis. Supports zero-shot voice cloning, cross-ling...
Converts text input into spoken audio output using StyleTTS2 technology. Accepts an optional speaker reference audio fil...
Convert speech to a target voice using RVC v2 voice models. Takes an input speech audio clip and outputs converted audio...
Generate spoken audio from text, optionally cloning a target voice from a short speaker reference audio. Accepts text as...
Convert text to expressive speech, with optional speaker style cloning from a short reference audio. Accepts text input...
Converts text to speech with instant voice cloning capabilities using a reference audio sample. Takes a text input and a...
Generate speech audio from text, with optional voice cloning from a reference speaker clip. Accepts text as the primary...
Generate speech from text with zero-shot voice cloning using a reference voice sample. Accepts text, a speaker reference...
Converts speech from one voice to another while preserving the original content using soft speech units. Takes an audio...