nowfound

Alternatives

Products that do what I hacked my son’s Duplo train to go faster using my voice does

  1. 1IB

    I built a voice agent from scratch that averages ~400ms end-to-end latency (phone stop → first syllable). That’s with full STT → LLM → TTS in the loop, clean barge-ins, and no precomputed responses. What moved the needle: Voice is a turn-taking problem, not a transcription problem. VAD alone fails; you need semantic end-of-turn detection. The system reduces to one loop: speaking vs listening. The two transitions - cancel instantly on barge-in, respond instantly on end-of-turn - define the experience. STT → LLM → TTS must stream. Sequential pipelines are dead on arrival for natural…

    Mar 2026 · ntik.me

  2. 2

    Change your voice to anyone in realtime with AI, for free

    2023 · dubbingai.io

  3. 3

    Next gen audio/video editor with personalized voice cloning

    2020

  4. 4VP

    Imagine creating a podcast where Mark Zuckerberg interviews Elon Musk – using their actual voices? What sounds like science fiction is now reality. Voice-Pro is an open-source Gradio WebUI that breaks the boundaries of audio manipulation. Powered by cutting-edge Whisper engines, this tool turns voice replication into child's play. Key Features: - Zero-shot Voice Cloning - Voice Changer with 50+ Celebrity Voices - YouTube Audio Downloading - Vocal Isolation - Multi-Language Text-to-Speech (Edge-TTS, F5-TTS) - Multi-Language Translation - Powered by Whisper Engines (Whisper, Faster-Whisper,…

    2024 · github.com

  5. 5

    Point at anything to learn its name in any language

    Jun 2026 · apps.apple.com

  6. 6LV
  7. 7
    Voxiyo112

    Record, organise and transform with AI

    2025

  8. 8OS

    Our goal with this project is to build a completely open source, state of the art turn detection model that can be used in any voice AI application. I've been experimenting with LLM voice conversations since GPT-4 was first released. (There's a previous front page Show HN about Pipecat, the open source voice AI orchestration framework I work on. [1]) It's been almost two years, and for most of that time, I've been expecting that someone would "solve" turn detection. We all built initial, pretty good 80/20 versions of turn detection on top of VAD (voice activity detection) models. And…

    2025 · github.com

  9. 9

    Transcribe and export to VTT and SRT for free

    2024

  10. 10

    Explore the most advanced voice cloning software ever

    2023

  11. 11

    The fastest generative AI Text-to-Speech API

    2023

  12. 12

    Live voice changer & face filters for video selfies.

    2017

  13. 13

    Create a persona and run AI powered prank calls for fun

    2024

  14. 14

    Your voice narrating their bedtime stories, every night

    Jun 2026 · benostudio.com

  15. 15IU

    Hi Hacker News, This is definitely out of my comfort zone. I just wanted to show you guys because I'm super proud of it. It's a 100% faithful recreation based off of the schematics, patents, and ROMs that were found online. So please watch the video and tell me what you think https://youtu.be/auOlZXI1VxA The reason why I think this is relevant is because I've been a programmer for 25 years and AI scares the shit out of me. I'm not a programmer anymore. I'm something else now. I don't know what it is but it's multi-disciplinary, and it doesn't involve writing code myself--for…

    Jan 2026

  16. 16

    Generate endless looping sound effects with a single prompt

    2025

  17. 17
    Schedodo156

    Instantly transform lectures and memos into organized notes

    2025

  18. 18SS

    Supertone's Shift offers real-time voice changing technology. It lets users immediately switch to any selected voice. Just pick a voice and begin speaking. Shift is suited for VTubers, content creators, and gamers, as well as anyone who wishes to accurately express their chosen persona's voice. Try out Supertone Shift now. >> https://product.supertone.ai/shift

    2024 · product.supertone.ai

  19. 19

    Build voice features remarkably fast

    2022

  20. 20

    Free ultra realistic text to speech for voiceover projects

    2020

  21. 21AP

    TLDR: We created a personalised Andrej Karpathy tutor that can response to questions about his Youtube videos in sub 1 second responses (voice-to-voice). We do this using a voice enabled RAG agent. See later in the post for demo link, Github Repo and blog write up. A few weeks ago we released the worlds fastest voice bot, achieving 500ms voice-to-voice response times, including a 200ms delay waiting for a user to stop speaking. After reaching the front page of HN, we thought about how we could take this a step further based on feedback we were getting from the community. Many companies were…

    2024 · educationbot.cerebrium.ai

  22. 22

    Children’s Speech Recognition Technology (Offline)

    2018

  23. 23

    A hackable robot that responds only in GIFs & videos

    2019

  24. 24
    Yoto80

    A clever speaker for kids

    2017

Ranked by how close each launch is in meaning, then by votes. Refine with a description →