nowfound

Alternatives

Products that do what OpenWhistle does

Audio in. Structured intelligence out.

  1. 1

    Free and open-source AI to turn voice into structured notes

    2024

  2. 2AJ

    Hey HN, we’re building an open specification that lets agents discover and invoke APIs with natural language, built on the OpenAPI standard. agents.json clearly defines the contract between LLMs and API as a standard that's open, observable, and replicable. Here’s a walkthrough of how it works: https://youtu.be/kby2Wdt2Dtk?si=59xGCDy48Zzwr7ND. There’s 2 parts to this: 1. An agents.json file describes how to link API calls together into outcome-based tools for LLMs. This file sits alongside an OpenAPI file. 2. The agents.json SDK loads agents.json files as tools for an LLM that…

    2025 · github.com

  3. 3
    OpenClaw841

    The AI that actually does things

    Jan 2026 · openclaw.ai

  4. 4

    Turn any static web page into a dynamic exchange

    2018

  5. 5

    Your Voice-Powered Sales Agent. One URL.

    2025

  6. 6

    Fast voice-to-text on 92 languages

    2023

  7. 7

    Open-source components for AI audio & voice agents

    Oct 2025

  8. 8

    Control your computer with natural language

    2023

  9. 9

    Extremely accurate, AI powered voice-to-text

    2024

  10. 10

    Build Powerful Voice Agents

    2025

  11. 11
    OpenWispr190

    100% local open source AI speech-to-text model

    2025

  12. 12
    GoWhisper150

    Cross-platform, privacy-first audio transcription app

    2023

  13. 13
    Openbase216

    Manage your team of AI agents by voice, from anywhere

    Jul 2026 · openbase.cloud

  14. 14OS

    Hey HN, we've been working with OpenAI for the past few months on the new Realtime API. The goal is to give everyone access to the same stack that underpins Advanced Voice in the ChatGPT app. Under the hood it works like this: - A user's speech is captured by a LiveKit client SDK in the ChatGPT app - Their speech is streamed using WebRTC to OpenAI’s voice agent - The agent relays the speech prompt over websocket to GPT-4o - GPT-4o runs inference and streams speech packets (over websocket) back to the agent - The agent relays generated speech using WebRTC back to the user’s device The…

    2024 · github.com

  15. 15FA
  16. 16

    A catalogue of artificial intelligence APIs by Microsoft

    2015

  17. 17

    Extremely accurate, AI powered voice-to-text for macOS

    2023

  18. 18IO

    Hi HN! Last year the project I launched here got a lot of good feedback on creating speech to speech AI on the ESP32. Recently I revamped the whole stack, iterated on that feedback and made our project fully open-source—all of the client, hardware, firmware code. This Github repo turns an ESP32-S3 into a realtime AI speech companion using the OpenAI Realtime API, Arduino WebSockets, Deno Edge Functions, and a full-stack web interface. You can talk to your own custom AI character, and it responds instantly. I couldn't find a resource that helped set up a reliable, secure websocket (WSS) AI…

    2025 · github.com

  19. 19OO

    Hello everyone. This is Yujong from the Hyprnote team (https://github.com/fastrepl/hyprnote). We built OWhisper for 2 reasons: (Also outlined in https://docs.hyprnote.com/owhisper/what-is-this) (1). While working with on-device, realtime speech-to-text, we found there isn't tooling that exists to download / run the model in a practical way. (2). Also, we got frequent requests to provide a way to plug in custom STT endpoints to the Hyprnote desktop app, just like doing it with OpenAI-compatible LLM endpoints. The (2) part is still kind of WIP, but…

    2025 · docs.hyprnote.com

  20. 20

    Turn your work activity into structured AI context.

    Feb 2026

  21. 21
    ZooData272

    The data layer for AI agents

    Jul 2026 · zoodata.ai

  22. 22

    The fastest generative AI Text-to-Speech API

    2023

  23. 23PS

    Welcome to Project S.A.T.U.R.D.A.Y. This is a project that allows anyone to easily build their own self-hosted J.A.R.V.I.S-like voice assistant. In my mind vocal computing is the future of human-computer interaction and by open sourcing this code I hope to expedite us on that path. I have had a blast working on this so far and I'm excited to continue to build with it. It uses whisper.cpp [1], Coqui TTS [2] and OpenAI [3] to do speech-to-text, text-to-text and text-to-speech inference all 100% locally (except for text-to-text). In the future I plan to swap out OpenAI for llama.cpp [4]. It is…

    2023 · github.com

  24. 24NO

    Hello HN! The day has finally come to stop adding features and start sharing what I've been building the last 5-6 months. It's a bit of CrewAI, OpenDevon, LangFuse/Cloud all in one, providing devs who prefer TypeScript an integrated framework thats provides a lot out of the box to start experimenting and building agents with. It started after peeking at the LangChain docs a few times and never liking the example code. I began experimenting with automating a simple Jira request from the engineering team to add an index to one of our Google Spanner databases (for context I'm the…

    2024 · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →