Alternatives
Products that do what Yuyin - AI Chinese Pronunciation Coach does
Speak Chinese with confidence — AI that actually listens
- 1

- 2

- 3IT
Built this because tones are killing my spoken Mandarin and I can't reliably hear my own mistakes. It's a 9M Conformer-CTC model trained on ~300h (AISHELL + Primewords), quantized to INT8 (11 MB), runs 100% in-browser via ONNX Runtime Web. Grades per-syllable pronunciation + tones with Viterbi forced alignment. Try it here: https://simedw.com/projects/ear/
Jan 2026 · simedw.com
- 4
- 5

- 6

- 7

- 8

- 9
- 10IM
2022 · freesubtitles.ai
- 11

- 12

- 13

- 14

- 15

- 16

- 17IB
Hello Hacker News, When learning foreign languages, I made the most progress by speaking them throughout the day, every day. So I made a site where you can *speak* to an AI language teacher to practice both listening and speaking. # The product *What I have now:* * Multilingual speech recognition: You can ask a question in English and get an answer in your target language. * Feedback on your grammar. * Suggestions: See examples of what to say next to keep the conversation flowing. * Speed: Choose a lower speed for beginners or a faster one for advanced levels. * Translations: Click to see a…
2023 · gliglish.com
- 18IB
Hi, I built TalkBits because most language apps focus on vocabulary or exercises, but not actual conversation. The hard part of learning a language is speaking naturally under pressure. TalkBits lets you have real-time spoken conversations with an AI that acts like a native speaker. You can choose different scenarios (travel, daily life, work, etc.), speak naturally, and the AI responds with natural speech back. The goal is to make it feel like talking to a real person rather than doing lessons. Techwise, it uses realtime speech input, transcription, LLM responses, and tts streaming to keep…
Jan 2026 · apps.apple.com
- 19

- 20

- 21

- 22

- 23RT
Hi HN -- voice chat with AI is very popular these days, especially with YC startups (https://twitter.com/k7agar/status/1769078697661804795). The current approaches all do a cascaded approach, with audio -> transcription -> language model -> text synthesis. This approach is easy to get started with, but requires lots of complexity and has a few glaring limitations. Most notably, transcription is slow, is lossy and any error propagates to the rest of the system, cannot capture emotional affect, is often not robust to code-switching/accents, and more. Instead, what…
2024 · demo.tincans.ai
- 24

Ranked by how close each launch is in meaning, then by votes. Refine with a description →