nowfound

Alternatives

Products that do what IndexTTS2 does

Precise duration & emotional zero-shot tts

  1. 1

    Turn text into natural speech in seconds.

    Oct 2025

  2. 2

    Voice AI that feels as good as it sounds

    May 2026 · inworld.ai

  3. 3
    MARS5 TTS489

    Open-source, insanely prosodic text-to-speech model

    2024

  4. 4KT

    Kitten TTS is an open-source series of tiny and expressive text-to-speech models for on-device applications. We are excited to launch a preview of our smallest model, which is less than 25 MB. This model has 15M parameters. This release supports English text-to-speech applications in eight voices: four male and four female. The model is quantized to int8 + fp16, and it uses onnx for runtime. The model is designed to run literally anywhere eg. raspberry pi, low-end smartphones, wearables, browsers etc. No GPU required! We're releasing this to give early users a sense of the latency and voices…

    2025 · github.com

  5. 5

    Describe any AI voice and prompt its emotional delivery

    2025

  6. 6

    Voice AI that’s 5% of the cost. 100% of the quality.

    2025

  7. 7

    The voice for your real-time AI applications

    2025

  8. 8

    Convert text to speech with natural voices free online

    2024

  9. 9
    Qwen3-TTS155

    Voice design, cloning & 97ms streaming

    Jan 2026 · qwen.ai

  10. 10

    Generate endless looping sound effects with a single prompt

    2025

  11. 11TA

    2012 · tts-api.com

  12. 12

    Highly expressive TTS model with high fidelity voice cloning

    2025

  13. 13

    Expressive Voice Cloning and Text-to-Speech

    Oct 2025

  14. 14TN

    Kitten TTS (https:&#x2F;&#x2F;github.com&#x2F;KittenML&#x2F;KittenTTS) is an open-source series of tiny and expressive text-to-speech models for on-device applications. We had a thread last year here: https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=44807868. Today we're releasing three new models with 80M, 40M and 14M parameters. The largest model (80M) has the highest quality. The 14M variant reaches new SOTA in expressivity among similar sized models, despite being <25MB in size. This release is a major upgrade from the previous one and supports English text-to-speech applications in…

    Mar 2026 · github.com

  15. 15

    Open-source TTS with emotion & voice cloning

    2025

  16. 16

    Real-time text-to-speech model you can self-host

    May 2026 · kugelaudio.com

  17. 17

    First TTS model to support all 22 Indic languages + English

    2024

  18. 18

    Expressive TTS with voice cloning in 15 languages

    Jun 2026 · microsoft.ai

  19. 19

    Text-to-speech API with natural language voice direction

    Apr 2026 · blog.google

  20. 20

    Multilingual TTS model with realistic and expressive speech

    Mar 2026 · mistral.ai

  21. 21

    Fast, expressive, open source TTS with native watermarking

    Dec 2025 · resemble.ai

  22. 22

    Make anything an audiobook or podcast

    2025

  23. 23

    Turn text into speech with zero-shot voice cloning.

    2025

  24. 24

    Open-Source Multilingual Text-to-Speech with Voice Cloning

    2024

Ranked by how close each launch is in meaning, then by votes. Refine with a description →