NEWVectors or files. Pick a path.Start →

    Text To Speech Models

    Browse AI models for multimodal decomposition and recomposition pipelines: plug any model into your extractors.

    333 models available

    Showing 1-24 of 333 models

    Text To Speech

    hexgrad/Kokoro-82M

    11.6M
    6,665
    Text To Speech

    coqui/XTTS-v2

    8.3M
    3,723
    coqui
    Text To Speech

    ResembleAI/chatterbox

    2.2M
    1,733
    chatterbox
    Text To Speech

    Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice

    2.1M
    1,867
    Text To Speech

    Qwen/Qwen3-TTS-12Hz-0.6B-CustomVoice

    1.5M
    174
    Text To Speech

    onnx-community/Kokoro-82M-v1.0-ONNX

    1.3M
    243
    transformers.js
    Text To Speech

    k2-fsa/OmniVoice

    765K
    1,248
    omnivoice
    Text To Speech

    SWivid/F5-TTS

    742K
    1,191
    f5-tts
    Text To Speech

    openbmb/VoxCPM2

    688K
    1,522
    voxcpm
    Text To Speech

    microsoft/VibeVoice-Realtime-0.5B

    600K
    1,267
    transformers
    Text To Speech

    Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign

    475K
    386
    qwen-tts
    Text To Speech

    Qwen/Qwen3-TTS-12Hz-0.6B-Base

    431K
    275
    Text To Speech

    fishaudio/s2-pro

    414K
    1,218
    Text To Speech

    bosonai/higgs-tts-2-3b-base

    395K
    694
    transformers
    Text To Speech

    pnnbao-ump/VieNeu-TTS-v3-Turbo

    320K
    47
    Text To Speech

    bosonai/higgs-tts-3-4b

    314K
    707
    transformers
    Text To Speech

    speechbrain/tts-hifigan-libritts-22050Hz

    270K
    6
    speechbrain
    Text To Speech

    wasmdashai/lahja-sa-ahmad-v1

    267K
    5
    transformers
    Text To Speech

    ai4bharat/indic-parler-tts

    252K
    283
    transformers
    Text To Speech

    kenpath/svara-tts-v1

    252K
    50
    transformers
    Text To Speech

    wasmdashai/lahja-sa-huba-v1

    221K
    3
    transformers
    Text To Speech

    Serveurperso/Qwen3-TTS-GGUF

    201K
    30
    gguf
    Text To Speech

    bosonai/higgs-audio-v2-generation-3B-base

    199K
    684
    transformers
    Text To Speech

    sesame/csm-1b

    185K
    2,424
    transformers
    1 / 14