Skip to main content
KYM

Detected Skills:

Text-to-Speech (100% match)

Found 0 registries and 29 entities for "Text-to-Speech"

All Agents & Models

@cf/myshell-ai/melotts

model

MeloTTS is a high-quality multi-lingual text-to-speech library by MyShell.ai.

Speech Recognition Text-to-Speech
@cf Score: 0

@cf/deepgram/aura-2-es

model

Aura-2 is a context-aware text-to-speech (TTS) model that applies natural pacing, expressiveness, and fillers based on the context of the provided text. The quality of your text input directly impacts the naturalness of the audio output.

Speech Recognition
@cf Score: 0

@cf/deepgram/aura-1

model

Aura is a context-aware text-to-speech (TTS) model that applies natural pacing, expressiveness, and fillers based on the context of the provided text. The quality of your text input directly impacts the naturalness of the audio output.

Speech Recognition
@cf Score: 0

@cf/deepgram/aura-2-en

model

Aura-2 is a context-aware text-to-speech (TTS) model that applies natural pacing, expressiveness, and fillers based on the context of the provided text. The quality of your text input directly impacts the naturalness of the audio output.

Speech Recognition
@cf Score: 0

qwen/qwen3-tts

model

A unified Text-to-Speech demo featuring three powerful modes: Voice, Clone and Design

Text Generation Speech Recognition Text-to-Speech
qwen Score: 0

minimax/speech-2.6-turbo

model

Low‑latency MiniMax Speech 2.6 Turbo brings multilingual, emotional text-to-speech to Replicate with 300+ voices and real-time friendly pricing

Text Generation Speech Recognition Text-to-Speech
minimax Score: 0

elevenlabs/v2-multilingual

model

Generate multilingual text-to-speech audio in over 30 languages

Text Generation Speech Recognition Text-to-Speech
elevenlabs Score: 0

vladpolbennikov/kokoro-82m-all-voices

model

Kokoro v1.0 2025 Jan 27 - text-to-speech (82M params, based on StyleTTS2)

Text Generation Speech Recognition Text-to-Speech
vladpolbennikov Score: 0

lucataco/indextts-2

model

Emotionally Expressive and Duration-Controlled Text-to-Speech

Text Generation Speech Recognition Text-to-Speech
lucataco Score: 0

microsoft/vibevoice

model

Microsoft's VibeVoice text-to-speech model that can generate long-form speech from text with sample voices.

Text Generation Speech Recognition Text-to-Speech
microsoft Score: 0

lucataco/higgs-audio-v2

model

Higgs Audio v2, a powerful text-to-speech audio foundation model that excels in expressive audio generation

Text Generation Speech Recognition Text-to-Speech
lucataco Score: 0

cjwbw/voicecraft

model

Zero-Shot Speech Editing and Text-to-Speech in the Wild

Text Generation Speech Recognition Text-to-Speech
cjwbw Score: 0

lucataco/step-audio-tts-3b

model

Step-Audio-TTS-3B represents the industry's first Text-to-Speech (TTS) model trained on a large-scale synthetic dataset utilizing the LLM-Chat paradigm

Text Generation Speech Recognition Text-to-Speech
lucataco Score: 0

cuuupid/zonos

model

Zonos-v0.1 beta, a SOTA text-to-speech Transformer model with extraordinary expressive range, built by Zyphra.

Text Generation Speech Recognition Text-to-Speech
cuuupid Score: 0

alphanumericuser/kokoro-82m

model

Kokoro v1.0 - text-to-speech (82M params, based on StyleTTS2)

Text Generation Speech Recognition Text-to-Speech
alphanumericuser Score: 0

jaaari/kokoro-82m

model

Kokoro v1.0 - text-to-speech (82M params, based on StyleTTS2)

Text Generation Speech Recognition Text-to-Speech
jaaari Score: 0

e1100x/chattts

model

ChatTTS is a text-to-speech model designed specifically for dialogue scenarios such as LLM assistant.

Text Generation Speech Recognition Text-to-Speech
e1100x Score: 0

zsxkib/hololive-style-bert-vits2

model

🎙️Hololive text-to-speech and voice-to-voice (Japanese🇯🇵 + English🇬🇧)

Text Generation Speech Recognition Text-to-Speech
zsxkib Score: 0

lee101/guided-text-to-speech

model

Guided Text to Speech Generator

Text Generation Speech Recognition Text-to-Speech
lee101 Score: 0

cjwbw/parler-tts

model

lightweight text-to-speech (TTS) model, trained on 10.5K hours of audio data

Text Generation Speech Recognition Text-to-Speech
cjwbw Score: 0