nowfound

Alternatives

Products that do what Houndify – A sophisticated voice recognition platform (join us in beta) does

https://www.houndify.com/ Hey all, I'm a software engineer at SoundHound where we are very actively developing our sophisticated new voice recognition API. You can see it at work in these vids of the Hound assistant: https://www.youtube.com/watch?v=M1ONXea0mXg https://www.youtube.com/watch?v=9KVlKRK-wkQ https://www.youtube.com/watch?v=ckXei1d-zr0 We're in the process of giving out verification codes but ping me (rob at soundhound dot com) if your interested in using it right away or if your in Toronto or the Bar area and want help…

  1. 1
    Hound230

    Voice search and assistant app. A better Siri.

    2016

  2. 2
    Houndify148

    Add a voice enabled conversational interface to anything

    2015

  3. 3AO

    I've been obsessed for the past ~year with the possibilities of talking to LLMs. I built a bunch of one-off prototypes, shared code on X, started a Meetup group in SF, and co-hosted a big hackathon. It turns out that there are a few low-level problems that everybody building conversational/real-time AI needs to solve on the way to building/shipping something that works well: low-latency media transport, echo cancellation, voice activity detection, phrase endpointing, pipelining data between models/services, handling voice interruptions, swapping out different…

    2024 · github.com

  4. 4VP

    Imagine creating a podcast where Mark Zuckerberg interviews Elon Musk – using their actual voices? What sounds like science fiction is now reality. Voice-Pro is an open-source Gradio WebUI that breaks the boundaries of audio manipulation. Powered by cutting-edge Whisper engines, this tool turns voice replication into child's play. Key Features: - Zero-shot Voice Cloning - Voice Changer with 50+ Celebrity Voices - YouTube Audio Downloading - Vocal Isolation - Multi-Language Text-to-Speech (Edge-TTS, F5-TTS) - Multi-Language Translation - Powered by Whisper Engines (Whisper, Faster-Whisper,…

    2024 · github.com

  5. 5

    An audio-based web accessibility tool for pronouncing names.

    2016

  6. 6

    Music discovery, identification, & voice-controlled player

    2017

  7. 7

    Your voice. Your personality. Twitter for audio.

    2017

  8. 8PS

    Welcome to Project S.A.T.U.R.D.A.Y. This is a project that allows anyone to easily build their own self-hosted J.A.R.V.I.S-like voice assistant. In my mind vocal computing is the future of human-computer interaction and by open sourcing this code I hope to expedite us on that path. I have had a blast working on this so far and I'm excited to continue to build with it. It uses whisper.cpp [1], Coqui TTS [2] and OpenAI [3] to do speech-to-text, text-to-text and text-to-speech inference all 100% locally (except for text-to-text). In the future I plan to swap out OpenAI for llama.cpp [4]. It is…

    2023 · github.com

  9. 9VA

    Voxos is an open-source desktop voice assistant that aims to put Clippy to shame while supporting new desktop workflows powered by LLMs. Tired of copy and pasting ChatGPT responses between your web browser and IDE? Does your copilot not quite do what you need it to do? I invite you to give Voxos a try and maybe even become a contributor!

    2024 · gitlab.com

  10. 10
    VoiceGist102

    Code sharing with integrated voice-over recording

    2021

  11. 11EA
  12. 12OF

    I wanted a voice-to-text app but didn't trust any of the proprietary ones with my privacy. So I decided to see if I could vibe code it with 0 macOS app & Swift experience. It uses a local binary of whisper.cpp (a fast implementation of OpenAI's Whisper voice-to-text model in C++). Github: https://github.com/richardwu/openwhisper I also decided to take this as an opportunity to compare 3 agentic coding harnesses: Cursor w/ Opus 4.6: - Best one-shot UI by far - Didn't get permissioning correct - Had issues making the "Cancel recording" hotkey being turned on all the…

    Feb 2026 · github.com

  13. 13IB

    I originally added this to my site to speed up my video editing process. Last year I started a youtube channel and for some of my longer videos it's annoying to rely on youtube or capcut to transcribe when Whisper is open source. Capcut also recently updated their T&Cs to say they own your content if you use their app, so I cancelled my subscription. Another use-case I have is recording my claude prompts as audio, transcribing them, and then pasting them into my terminal. I mostly work on the CLI (claude, ffmpeg, whisper), but I wanted to make a browser version. Not reinventing the wheel…

    2025 · meetcosmos.com

  14. 14SA

    Hi HN, We're a new Bay Area startup aiming to connect all of your things visually - supporting multiple triggers (AND and OR), multiple actions, cascading flows, etc. We're already integrating with dozens of services, both physical and digital - SmartThings, LIFX, Hue, Jawbone, Withings, Misfit, Dropbox, Google Drive, and the list goes on. We're really hoping to get some additional beta testers. We're on iOS only for now - so if you have an iPhone and want to give it a shot, get on our invite list @ https://www.stringify.com and we'll send you a TestFlight invite in the next few…

    2015

  15. 1501

    Hey HN! I've been working on a side project to create an audio transcription API based on the OpenAI whisper model. Sign up link: https://whisperapi.com I tried to make the API really easy to use and get setup with. Also, because the Whisper model is so good, turns out I can offer the service for about 75% cheaper than what seems like the industry average. I'm always looking to make improvements, so would appreciate any feedback anyone has!

    2022 · whisperapi.com

  16. 16V4

    Hey guys, Just finished up hacking on my submission for the Wufoo API contest. VoiceForms 4000 takes any existing Wufoo form and turns it into a phone survey. The app is written in Python, runs on App Engine, and uses Twilio for all the phone magic. https://voiceforms4000.appspot.com/ Go ahead and fill one of the sample forms and tell me what you think. Enjoy.

    2010

  17. 17AF

    I've developed a simple AI-powered British accent generator. Enter or paste your text, select the voice that best fits your project's tone, and generate speech for free. It supports up to 500 characters and offers 8 distinct, lifelike voices. Everything runs entirely within your browser. I'm primarily seeking feedback on output quality, user experience, and any technical improvements worth exploring.

    Feb 2026 · audioconvert.ai

  18. 18M

    URL: http://mixcastr.com/ The end of the weekend is here, and my project (roughly 15 hours) is done. It's called Mixcastr, and it uses the Sound Cloud API to serve songs from their platform, based on user queries. So yeah, it's really just a glorified search interface - but I had fun making it. Hopefully next weekend I'll expand on its functionality/utility. Breakdown: Sinatra, Twitter Bootstrap, SoundCloud JavaScript SDK, and Heroku, Polymorphs, CloudMade

    2011

  19. 19

    Free TTS API with 1,200+ voices in 75+ languages

    Feb 2026

  20. 20FT

    I built a simple text-to-speech converter at texttospeech.site Free tier: 10 generations/day, standard voices, no account needed. Pro tier: Neural2 voices, 2000 chars, downloadable MP3s. Stack: Next.js, Google Cloud TTS API, Vercel. The $2 domain was an SEO experiment after my speechtotext.xyz satellite drove 22% of traffic to my main product. Curious if exact-match keyword domains still work for TTS searches. Feedback welcome — especially on voice quality and UX.

    Jan 2026 · texttospeech.site

  21. 21CC

    Hey there HN! I believe the future of AI communication will be more voice and less text. Low-latency realistic voice interactions are finally becoming feasible. I've built a few voice-first apps on Retell AI using Elevenlabs voices. This one uses Claude Haiku for responses and Mixtral to switch between posts and comments. The AI knows about the top 30 posts and their comments on Hacker News right now. After a Google sign-in you can try it free for 10 minutes. I'd love to hear your thoughts!

    2024 · callhackernews.com

  22. 22

    Your voice is your moat. Posts in your voice. Every day.

    Jun 2026 · voicemoat.com

  23. 23ST

    I'm currently working on a "home for voice AI developers" called Vocalized. Here is a link to the playground where you can split-test different voice AI providers (I'm still working on making it mobile-responsive so try it on desktop!). I want the playground to be an complete documentation of every offering that exists in the space (whether relevant or not) so each can be compared. The intent for the site is to be a lightweight site you can bookmark & come back to to poke around if interested. The site also has a company directory that I'm working on. I wanted to offload my browser bookmarks…

    2024 · vocalized.dev

  24. 24VM

    Long time lurker, first time poster here. I have been teaching myself web and app development and I just built my first web app in my spare time. It's called Voxxcast, and it allows you to leave voice messages on Facebook via your phone. I'm definitely interested in any feedback you can give me. I plan on creating a lot more apps and developing this one even further. http://www.voxxcast.com

    2011

Ranked by how close each launch is in meaning, then by votes. Refine with a description →