nowfound

Alternatives

Products that do what VoiceSaaS Pro does

Next.js 15 Starter Kit for Twilio + OpenAI Voice Agents

  1. 1

    Build Powerful Voice Agents

    2025

  2. 2OS

    Hey HN, we've been working with OpenAI for the past few months on the new Realtime API. The goal is to give everyone access to the same stack that underpins Advanced Voice in the ChatGPT app. Under the hood it works like this: - A user's speech is captured by a LiveKit client SDK in the ChatGPT app - Their speech is streamed using WebRTC to OpenAI’s voice agent - The agent relays the speech prompt over websocket to GPT-4o - GPT-4o runs inference and streams speech packets (over websocket) back to the agent - The agent relays generated speech using WebRTC back to the user’s device The…

    2024 · github.com

  3. 3VP

    Imagine creating a podcast where Mark Zuckerberg interviews Elon Musk – using their actual voices? What sounds like science fiction is now reality. Voice-Pro is an open-source Gradio WebUI that breaks the boundaries of audio manipulation. Powered by cutting-edge Whisper engines, this tool turns voice replication into child's play. Key Features: - Zero-shot Voice Cloning - Voice Changer with 50+ Celebrity Voices - YouTube Audio Downloading - Vocal Isolation - Multi-Language Text-to-Speech (Edge-TTS, F5-TTS) - Multi-Language Translation - Powered by Whisper Engines (Whisper, Faster-Whisper,…

    2024 · github.com

  4. 4

    Premium AI voice quality without the premium price tag.

    2025

  5. 5AO

    I've been obsessed for the past ~year with the possibilities of talking to LLMs. I built a bunch of one-off prototypes, shared code on X, started a Meetup group in SF, and co-hosted a big hackathon. It turns out that there are a few low-level problems that everybody building conversational/real-time AI needs to solve on the way to building/shipping something that works well: low-latency media transport, echo cancellation, voice activity detection, phrase endpointing, pipelining data between models/services, handling voice interruptions, swapping out different…

    2024 · github.com

  6. 6

    Real Expressive AI Voices

    Mar 2026 · fish.audio

  7. 7

    Create realistic AI Voiceovers within seconds

    2022

  8. 8GA

    Hi! It's a complete product with integrations to Auth0, OpenAI, Google Cloud and Stripe, which consists of Next.js Web App, Node.js + Express Web API and Python + FastAPI AI API I've built this software, because I wanted to make money by selling tokens to enable users talking with the chatbot. But I think Google / Apple will include such AI-powered assistant in their products soon, so nobody will pay me for using it So I open source the product today and share it as a GNU GPL-2 licensed software I'm happy to assist in case if something is unclear or requires additional docs and answer…

    2023 · github.com

  9. 9PS

    Welcome to Project S.A.T.U.R.D.A.Y. This is a project that allows anyone to easily build their own self-hosted J.A.R.V.I.S-like voice assistant. In my mind vocal computing is the future of human-computer interaction and by open sourcing this code I hope to expedite us on that path. I have had a blast working on this so far and I'm excited to continue to build with it. It uses whisper.cpp [1], Coqui TTS [2] and OpenAI [3] to do speech-to-text, text-to-text and text-to-speech inference all 100% locally (except for text-to-text). In the future I plan to swap out OpenAI for llama.cpp [4]. It is…

    2023 · github.com

  10. 10

    Voice agents powered by Simba 3.2 the world's #1 voice model

    Jul 2026 · speechify.ai

  11. 11LV
  12. 12

    The most accurate streaming speech model for voice agents.

    Mar 2026 · assemblyai.com

  13. 13

    24/7 custom AI livechat + realtime voice chatbot

    2024

  14. 14IO

    Hi HN! Last year the project I launched here got a lot of good feedback on creating speech to speech AI on the ESP32. Recently I revamped the whole stack, iterated on that feedback and made our project fully open-source—all of the client, hardware, firmware code. This Github repo turns an ESP32-S3 into a realtime AI speech companion using the OpenAI Realtime API, Arduino WebSockets, Deno Edge Functions, and a full-stack web interface. You can talk to your own custom AI character, and it responds instantly. I couldn't find a resource that helped set up a reliable, secure websocket (WSS) AI…

    2025 · github.com

  15. 15
    Qwen3-TTS155

    Voice design, cloning & 97ms streaming

    Jan 2026 · qwen.ai

  16. 16

    Voice AI agents, zero setup, fast

    2025

  17. 17

    One API to build production-ready voice agents

    Apr 2026 · assemblyai.com

  18. 18

    Open-source TTS with emotion & voice cloning

    2025

  19. 19VA

    Voxos is an open-source desktop voice assistant that aims to put Clippy to shame while supporting new desktop workflows powered by LLMs. Tired of copy and pasting ChatGPT responses between your web browser and IDE? Does your copilot not quite do what you need it to do? I invite you to give Voxos a try and maybe even become a contributor!

    2024 · gitlab.com

  20. 20

    Real-time text-to-speech model you can self-host

    May 2026 · kugelaudio.com

  21. 21

    Make an AI voice chat app in 21 lines of JavaScript

    2024

  22. 22

    Drag & drop voice generator tool

    Dec 2025 · apps.apple.com

  23. 23AF

    2024 · swift-ai.vercel.app

  24. 24

    Stop typing, Start talking to your apps and agents

    Jun 2026 · voicetypr.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →