nowfound

Alternatives

Products that do what Audio AI had a wild day – 5 major open-source / real-time TTS drops does

The audio&#x2F;TTS space just moved fast. In the last week alone: NVIDIA – PersonaPlex-7B Open-source, full-duplex conversational speech model. Inworld AI – TTS-1.5 Realtime TTS (<250ms), $0.005&#x2F;min, currently #1 on Artificial Analysis. Flash Labs – Chroma 1.0 First open-source, end-to-end, real-time speech-to-speech model. Alibaba Qwen – Qwen3-TTS Fully open-sourced TTS family: Base, CustomVoice, VoiceDesign. Kyutai Labs – Pocket TTS Runs locally on a laptop. No GPU required. Feels like TTS is hitting the same acceleration moment LLMs had last year. Realtime, open-source, and local is…

  1. 1KT

    Kitten TTS is an open-source series of tiny and expressive text-to-speech models for on-device applications. We are excited to launch a preview of our smallest model, which is less than 25 MB. This model has 15M parameters. This release supports English text-to-speech applications in eight voices: four male and four female. The model is quantized to int8 + fp16, and it uses onnx for runtime. The model is designed to run literally anywhere eg. raspberry pi, low-end smartphones, wearables, browsers etc. No GPU required! We're releasing this to give early users a sense of the latency and voices…

    2025 · github.com

  2. 2

    Voice AI that’s 5% of the cost. 100% of the quality.

    2025

  3. 3

    The voice for your real-time AI applications

    2025

  4. 4

    Voice AI that feels as good as it sounds

    May 2026 · inworld.ai

  5. 5TN

    Kitten TTS (https:&#x2F;&#x2F;github.com&#x2F;KittenML&#x2F;KittenTTS) is an open-source series of tiny and expressive text-to-speech models for on-device applications. We had a thread last year here: https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=44807868. Today we're releasing three new models with 80M, 40M and 14M parameters. The largest model (80M) has the highest quality. The 14M variant reaches new SOTA in expressivity among similar sized models, despite being <25MB in size. This release is a major upgrade from the previous one and supports English text-to-speech applications in…

    Mar 2026 · github.com

  6. 6
    MARS5 TTS489

    Open-source, insanely prosodic text-to-speech model

    2024

  7. 7

    Expressive TTS with voice cloning in 15 languages

    Jun 2026 · microsoft.ai

  8. 8

    Multilingual TTS model with realistic and expressive speech

    Mar 2026

  9. 9RT
  10. 10

    First TTS model to support all 22 Indic languages + English

    2024

  11. 11

    Open-source TTS with emotion & voice cloning

    2025

  12. 12MO

    I wanted to share our new speech to text model, and the library to use them effectively. We're a small startup (six people, sub-$100k monthly GPU budget) so I'm proud of the work the team has done to create streaming STT models with lower word-error rates than OpenAI's largest Whisper model. Admittedly Large v3 is a couple of years old, but we're near the top the HF OpenASR leaderboard, even up against Nvidia's Parakeet family. Anyway, I'd love to get feedback on the models and software, and hear about what people might build with it.

    Feb 2026 · github.com

  13. 13

    Build Powerful Voice Agents

    2025

  14. 14
    OpenWispr190

    100% local open source AI speech-to-text model

    2025

  15. 15

    Generate endless looping sound effects with a single prompt

    2025

  16. 16IO

    Hi HN! Last year the project I launched here got a lot of good feedback on creating speech to speech AI on the ESP32. Recently I revamped the whole stack, iterated on that feedback and made our project fully open-source—all of the client, hardware, firmware code. This Github repo turns an ESP32-S3 into a realtime AI speech companion using the OpenAI Realtime API, Arduino WebSockets, Deno Edge Functions, and a full-stack web interface. You can talk to your own custom AI character, and it responds instantly. I couldn't find a resource that helped set up a reliable, secure websocket (WSS) AI…

    2025 · github.com

  17. 17

    Fast & Affordable Text-to-Speech API

    2025

  18. 18

    Premium AI voice quality without the premium price tag.

    2025

  19. 19

    Convert text to speech with natural voices free online

    2024

  20. 20

    Real-time text-to-speech model you can self-host

    May 2026 · kugelaudio.com

  21. 21
    Qwen3-TTS155

    Voice design, cloning & 97ms streaming

    Jan 2026

  22. 22TA

    2012 · tts-api.com

  23. 23
    VoxCPM2110

    Open-source 48kHz TTS with voice design and cloning

    Apr 2026

  24. 24

    Ultra-realistic AI voices & cloning

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →