nowfound

Alternatives

Products that do what TTSLab – A voice AI agent and TTS lab running in the browser via WebGPU does

I built TTSLab — a free, open-source tool for running text-to-speech and speech-to-text models directly in the browser using WebGPU and WASM. No API keys, no backend, no data leaves your machine. When you open the site, you'll hear it immediately — the landing page auto-generates speech from three different sentences right in your browser, no setup required. You can then try any model yourself: type text, hit generate, hear it instantly. Models download once and get cached locally. The most experimental feature: a fully in-browser Voice Agent. It chains speech-to-text → LLM → text-to-speech,…

  1. 1
    TTSLab5

    Test TTS & STT models in your browser. No server required.

    Mar 2026

  2. 2KT

    Kitten TTS is an open-source series of tiny and expressive text-to-speech models for on-device applications. We are excited to launch a preview of our smallest model, which is less than 25 MB. This model has 15M parameters. This release supports English text-to-speech applications in eight voices: four male and four female. The model is quantized to int8 + fp16, and it uses onnx for runtime. The model is designed to run literally anywhere eg. raspberry pi, low-end smartphones, wearables, browsers etc. No GPU required! We're releasing this to give early users a sense of the latency and voices…

    2025 · github.com

  3. 3
    MARS5 TTS489

    Open-source, insanely prosodic text-to-speech model

    2024

  4. 4

    Build Powerful Voice Agents

    2025

  5. 5

    Generate endless looping sound effects with a single prompt

    2025

  6. 6

    Describe any AI voice and prompt its emotional delivery

    2025

  7. 7

    Convert text to speech with natural voices free online

    2024

  8. 8

    First TTS model to support all 22 Indic languages + English

    2024

  9. 9

    Voice AI that’s 5% of the cost. 100% of the quality.

    2025

  10. 10TN

    Kitten TTS (https:&#x2F;&#x2F;github.com&#x2F;KittenML&#x2F;KittenTTS) is an open-source series of tiny and expressive text-to-speech models for on-device applications. We had a thread last year here: https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=44807868. Today we're releasing three new models with 80M, 40M and 14M parameters. The largest model (80M) has the highest quality. The 14M variant reaches new SOTA in expressivity among similar sized models, despite being <25MB in size. This release is a major upgrade from the previous one and supports English text-to-speech applications in…

    Mar 2026 · github.com

  11. 11PS

    Welcome to Project S.A.T.U.R.D.A.Y. This is a project that allows anyone to easily build their own self-hosted J.A.R.V.I.S-like voice assistant. In my mind vocal computing is the future of human-computer interaction and by open sourcing this code I hope to expedite us on that path. I have had a blast working on this so far and I'm excited to continue to build with it. It uses whisper.cpp [1], Coqui TTS [2] and OpenAI [3] to do speech-to-text, text-to-text and text-to-speech inference all 100% locally (except for text-to-text). In the future I plan to swap out OpenAI for llama.cpp [4]. It is…

    2023 · github.com

  12. 12TA

    2012 · tts-api.com

  13. 13

    Fast, accurate STT for production-grade voice agents

    May 2026 · ringg.ai

  14. 14
    OpenWispr190

    100% local open source AI speech-to-text model

    2025

  15. 15

    The fastest generative AI Text-to-Speech API

    2023

  16. 16IM

    A few years ago, right after high school, I decided to try to make a simultaneous translation app for Android as a side project, it took longer than expected (about 2 years) and I had to make a lot of compromises (I had to use Google's API and therefore make users use a developer key because at the time there were no free solutions for speech recognition and translation that had good quality). At the end of university, I decided to pick it up again and finally, using OpenAi's Whisper for speech recognition and Meta's NLLB for translation (with both running locally on the phone), I managed to…

    2024 · github.com

  17. 17IO

    Hi HN! Last year the project I launched here got a lot of good feedback on creating speech to speech AI on the ESP32. Recently I revamped the whole stack, iterated on that feedback and made our project fully open-source—all of the client, hardware, firmware code. This Github repo turns an ESP32-S3 into a realtime AI speech companion using the OpenAI Realtime API, Arduino WebSockets, Deno Edge Functions, and a full-stack web interface. You can talk to your own custom AI character, and it responds instantly. I couldn't find a resource that helped set up a reliable, secure websocket (WSS) AI…

    2025 · github.com

  18. 18

    Open-source TTS with emotion & voice cloning

    2025

  19. 19LV
  20. 20

    The voice for your real-time AI applications

    2025

  21. 21

    Text-to-speech API with natural language voice direction

    Apr 2026

  22. 22OS

    Our goal with this project is to build a completely open source, state of the art turn detection model that can be used in any voice AI application. I've been experimenting with LLM voice conversations since GPT-4 was first released. (There's a previous front page Show HN about Pipecat, the open source voice AI orchestration framework I work on. [1]) It's been almost two years, and for most of that time, I've been expecting that someone would "solve" turn detection. We all built initial, pretty good 80&#x2F;20 versions of turn detection on top of VAD (voice activity detection) models. And…

    2025 · github.com

  23. 23

    Open-Source Multilingual Text-to-Speech with Voice Cloning

    2024

  24. 24MP

    I built a pipeline that turns tweets about ML papers into a podcast. Code's up here. Happy hacking. https:&#x2F;&#x2F;github.com&#x2F;yacineMTB&#x2F;scribepod

    2023 · scribepod.substack.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →