nowfound

Alternatives

Products that do what TTSLab does

Test TTS & STT models in your browser. No server required.

  1. 1KT

    Kitten TTS is an open-source series of tiny and expressive text-to-speech models for on-device applications. We are excited to launch a preview of our smallest model, which is less than 25 MB. This model has 15M parameters. This release supports English text-to-speech applications in eight voices: four male and four female. The model is quantized to int8 + fp16, and it uses onnx for runtime. The model is designed to run literally anywhere eg. raspberry pi, low-end smartphones, wearables, browsers etc. No GPU required! We're releasing this to give early users a sense of the latency and voices…

    2025 · github.com

  2. 2TA

    I built TTSLab — a free, open-source tool for running text-to-speech and speech-to-text models directly in the browser using WebGPU and WASM. No API keys, no backend, no data leaves your machine. When you open the site, you'll hear it immediately — the landing page auto-generates speech from three different sentences right in your browser, no setup required. You can then try any model yourself: type text, hit generate, hear it instantly. Models download once and get cached locally. The most experimental feature: a fully in-browser Voice Agent. It chains speech-to-text → LLM → text-to-speech,…

    Feb 2026 · ttslab.dev

  3. 3TN

    Kitten TTS (https:&#x2F;&#x2F;github.com&#x2F;KittenML&#x2F;KittenTTS) is an open-source series of tiny and expressive text-to-speech models for on-device applications. We had a thread last year here: https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=44807868. Today we're releasing three new models with 80M, 40M and 14M parameters. The largest model (80M) has the highest quality. The 14M variant reaches new SOTA in expressivity among similar sized models, despite being <25MB in size. This release is a major upgrade from the previous one and supports English text-to-speech applications in…

    Mar 2026 · github.com

  4. 4MO

    I wanted to share our new speech to text model, and the library to use them effectively. We're a small startup (six people, sub-$100k monthly GPU budget) so I'm proud of the work the team has done to create streaming STT models with lower word-error rates than OpenAI's largest Whisper model. Admittedly Large v3 is a couple of years old, but we're near the top the HF OpenASR leaderboard, even up against Nvidia's Parakeet family. Anyway, I'd love to get feedback on the models and software, and hear about what people might build with it.

    Feb 2026 · github.com

  5. 5
    MARS5 TTS489

    Open-source, insanely prosodic text-to-speech model

    2024

  6. 6

    Build Powerful Voice Agents

    2025

  7. 7

    Convert text to speech with natural voices free online

    2024

  8. 8PS

    Welcome to Project S.A.T.U.R.D.A.Y. This is a project that allows anyone to easily build their own self-hosted J.A.R.V.I.S-like voice assistant. In my mind vocal computing is the future of human-computer interaction and by open sourcing this code I hope to expedite us on that path. I have had a blast working on this so far and I'm excited to continue to build with it. It uses whisper.cpp [1], Coqui TTS [2] and OpenAI [3] to do speech-to-text, text-to-text and text-to-speech inference all 100% locally (except for text-to-text). In the future I plan to swap out OpenAI for llama.cpp [4]. It is…

    2023 · github.com

  9. 9

    First TTS model to support all 22 Indic languages + English

    2024

  10. 10IM

    A few years ago, right after high school, I decided to try to make a simultaneous translation app for Android as a side project, it took longer than expected (about 2 years) and I had to make a lot of compromises (I had to use Google's API and therefore make users use a developer key because at the time there were no free solutions for speech recognition and translation that had good quality). At the end of university, I decided to pick it up again and finally, using OpenAi's Whisper for speech recognition and Meta's NLLB for translation (with both running locally on the phone), I managed to…

    2024 · github.com

  11. 11

    Open-source TTS with emotion & voice cloning

    2025

  12. 12

    Highly expressive TTS model with high fidelity voice cloning

    2025

  13. 13

    The voice for your real-time AI applications

    2025

  14. 14

    Generate endless looping sound effects with a single prompt

    2025

  15. 15
    OpenWispr190

    100% local open source AI speech-to-text model

    2025

  16. 16
    Mimic 377

    Privacy-focused neural text-to-speech (TTS) engine

    2022

  17. 17TA

    2012 · tts-api.com

  18. 18LV
  19. 19
    VoxCPM2110

    Open-source 48kHz TTS with voice design and cloning

    Apr 2026 · github.com

  20. 20

    Voice AI that feels as good as it sounds

    May 2026 · inworld.ai

  21. 21

    Real-time text-to-speech model you can self-host

    May 2026 · kugelaudio.com

  22. 22
    ChattyUI149

    Run open-source LLMs locally in the browser using WebGPU

    2024

  23. 23

    Multilingual TTS model with realistic and expressive speech

    Mar 2026 · mistral.ai

  24. 24
    Kokori101

    Transform text to speech with a powerful macOS app

    Jan 2026

Ranked by how close each launch is in meaning, then by votes. Refine with a description →