Alternatives
Products that do what Outspeed – Platform for realtime voice and video AI does
Hey HN! Janak here from Outspeed (https://outspeed.com). We’re excited to show you Outspeed : a purpose-built platform for realtime voice & video AI applications. Here’s a demo of some cool apps you can create using Outspeed: https://www.youtube.com/watch?v=a11LQIlXelM Outspeed emerged from our frustration of needing to stitch together multiple tools such as livekit, vocode, langflow, silero etc. just to make a simple voice bot. Even after all that hard work, it still wasn’t production-ready. So we decided to work on a complete framework that could stand production…
- 1

- 2

- 3

- 4

- 5

- 6

- 7

- 8

- 9

- 10

- 11

- 12

- 13OS
Hey HN, it’s Russ - cofounder of LiveKit. An open source stack for building realtime AI applications. We’re sharing our first homegrown AI model for turn detection. Here’s a live demo: https://cerebras.vercel.app/ Voice AI has come a long way in the last year. We now have end-to-end systems that can generate a response to user input in 300-500ms — human level speeds! As latency reduces, a common problem that surfaces is the LLM responds too quickly. Any time there’s a short pause in a user’s speech, it ends up interrupting them. This is largely due to how voice AI applications…
2024
- 14WO
We kept hitting the same wall building voice AI systems. Pipecat and LiveKit are great projects, genuinely. But getting it to production took us weeks of plumbing - wiring things together, handling barge-ins, setting up telephony, Knowledge base, tool calls, handling barge in etc. And every time we needed to tweak agent behavior, you were back in the code and redeploying. We just wanted to change a prompt and test it in 30 seconds. Thats why Vapi retell etc exist. So we wrote the entire code and open sourced it as a Visual drag-and-drop for voice agents ( same as vapi or n8n for voice).…
Mar 2026 · github.com
- 15CC
Hey there HN! I believe the future of AI communication will be more voice and less text. Low-latency realistic voice interactions are finally becoming feasible. I've built a few voice-first apps on Retell AI using Elevenlabs voices. This one uses Claude Haiku for responses and Mixtral to switch between posts and comments. The AI knows about the top 30 posts and their comments on Hacker News right now. After a Google sign-in you can try it free for 10 minutes. I'd love to hear your thoughts!
2024 · callhackernews.com
- 16OS
Hi, I’m Sagar. We just open-sourced a framework to build real-time AI-powered video avatars you can drop into any app or website. You can use it to create sales assistants, customer success agents, mock interviewers, language coaches, or even historical characters. It’s modular (choose your STT, LLM, and TTS provider), production-ready, and optimized for ultra-low latency video generation. Features: - Real-time speech-to-video avatars (<300ms) - Native turn detection, VAD, and noise suppression - Modular pipelines for STT, LLM, TTS, and avatars with real-time model switching - Built-in RAG +…
2025 · github.com
- 17IM
SpeakFast helps you prep by actually talking—not just reading tips or recording yourself. The AI interviews you, challenges you in real time, and gives coaching based on how you're doing. It can even build a custom roadmap around your weak spots. There are 200K+ real roles from top tech companies to practice with. You can also paste a job description and get a tailored mock interview instantly. It’s composable, flexible, and built to feel real
2025 · speakfast.ai
- 18BA
cerebrium.ai/blog/how-to-build-a-real-time-ai-avatar-for-training-and-coaching At Cerebrium, we have recently built a few demos showing voice AI capabilities (worlds fastest voice agent & realtime RAG agent) but we wanted to push the boundary and see if we could create realistic, human-like situations to train and onboard teams to perform better - recreating real life scenarios! An example of this is a sales coach for your sales team, an investor pitch or even prep for a notoriously stressful YC interview . To achieve this there were a few difficult problems to solve, namely: - How…
2024 · coaching.cerebrium.ai
- 19AE
Hey HN! Have been working on Audentic, a platform that simplifies adding voice AI to your website. Think of it as a copy-paste voice assistant that you can embed directly into your site. Setting up voice AI has traditionally been a bit of a hassle, often requiring juggling multiple components like speech recognition, text processing, and text-to-speech systems. With Audentic, we've streamlined this into a more straightforward, end-to-end voice model. While OpenAI's Realtime API has made real-time, multimodal AI interactions more feasible, integrating these capabilities into a website can…
2025 · audentic.io
- 20IB
Hey HN, I'm Damiano Rodriguez, and I've been working solo on a project called Caisual Games. It's a free platform where anyone can create web games just by describing them with their voice. Here's how it works: You simply describe your game idea ("Make a cookie clicker game with 5 upgrades"), and the platform generates a playable web game in seconds using HTML and JavaScript. There's no coding required and no downloads needed—you can create and play directly in your browser. What's more, anyone can improve a published game, fostering a collaborative environment for game creation. The…
2024 · caisual.com
- 21AA
Hello Hacker News, I've built a platform that lets anyone create their own voice-based AI tutor. To showcase what the platform can do, I've created a voice-based AI tutor for Stanford's famous CS229 Machine Learning course. I took the course myself and know how challenging it can be, so I built this to make it easier to follow and stick with. It's free for now (as long as budget lasts) -> https://vivaverbalis.com/ml The goal of my app is simple: Whether you're a student, teacher, or just curious, you can instantly spin up a conversational AI tutor, and share it, with whoever…
2025 · vivaverbalis.com
- 22ET
For a while I've wanted to try out the new AI voices for long-form narration, but everything I found required a subscription that didn't justify my limited usage. I came across the open Kokoro model [0] and the voices are very good -- good enough to listen to for hours without the fatigue I got from legacy, robotic TTS voices. The model is 82m parameters and designed to run fast, but I still struggled to get reasonable times from CPU inference on my 12-core laptop. I thought a cloud-based GPU service would let me generate audiobooks fast enough to feed my own self-hosted library, and that…
Jun 2026 · ebookaloud.com
- 23

- 24ST
I'm currently working on a "home for voice AI developers" called Vocalized. Here is a link to the playground where you can split-test different voice AI providers (I'm still working on making it mobile-responsive so try it on desktop!). I want the playground to be an complete documentation of every offering that exists in the space (whether relevant or not) so each can be compared. The intent for the site is to be a lightweight site you can bookmark & come back to to poke around if interested. The site also has a company directory that I'm working on. I wanted to offload my browser bookmarks…
2024 · vocalized.dev
Ranked by how close each launch is in meaning, then by votes. Refine with a description →