nowfound

Alternatives

Products that do what Australian Acoustic Observatory Search does

The Australian Acoustic Observatory (https://acousticobservatory.org/) has 360 microphones across the continent, and over 2 million hours of audio. However, none of it is labeled: We want to make this enormous repository useful to researchers. We have found that researchers are often looking for 'hard' signals - specific call-types, birds with very little available training data, and so on. So we built an acoustic-similarity search tool, allowing researchers to provide an example of what they're looking for, which we then match against embeddings from the A2O dataset. Here's…

  1. 1IT

    Hey HN! I just shipped a project I’ve been working on called Maroofy: https://maroofy.com You can search for any song, and it’ll use the song’s audio to find other similar-sounding music. Demo: https://twitter.com/subby_tech/status/1621293770779287554 How does it work? I’ve indexed ~120M+ songs from the iTunes catalog with a custom AI audio model that I built for understanding music. My model analyzes raw music audio as input and produces embedding vectors as output. I then store the embedding vectors for all songs into a vector database, and use semantic…

    2023 · maroofy.com

  2. 2MA

    I am excited to announce a new tool for music producers and audio enthusiasts - a music audio search engine. With just a simple description of the groove you're looking for, our semantic search engine will output the most similar audio in seconds. I used the Freesound.org API to upload over 3,000 grooves to MongoDB, and combined all the relevant data such as tags, title, description, BPM, etc. into OpenAI's Text-Davinci to generate a unique description of each sound. I then embedded these descriptions using the Ada Embeddings Model and inserted them into Pinecone DB vector database, making…

    2023 · muzic-sage.vercel.app

  3. 3IV
  4. 4
    MARS5 TTS489

    Open-source, insanely prosodic text-to-speech model

    2024

  5. 5

    A neural net for speech recognition

    2022

  6. 6

    Find anything inside audio, video, images & documents

    2022

  7. 7
    Cekura431

    Observe and analyze your voice and chat AI agents

    Mar 2026

  8. 8

    Fast, accurate STT for production-grade voice agents

    May 2026 · ringg.ai

  9. 9

    Expressive Voice Cloning and Text-to-Speech

    Oct 2025

  10. 10

    Multilingual speech AI model trained on 12.5M hours of data

    2024

  11. 11

    Advancing automatic speech recognition for 1,600+ languages

    Nov 2025

  12. 12CA

    In 2020, Magenta released DDSP [1], a machine learning algorithm / python library which made it possible to generate good sounding instrument synthesizers from about 6-10 minutes of data. While working with DDSP for a project, we realised how it was actually quite hard to find 6-10 minute of clean recordings of monophonic instruments. In this project, we have combined the DDSP architecture with a domain adaptation technique from speech synthesis [2]. This domain adaptation technique works by pre-training our model on many different recordings from the Solos dataset [3] first and then…

    2022 · erlj.notion.site

  13. 13

    Real Expressive AI Voices

    Mar 2026

  14. 14SS

    I built a podcast search engine that performs a semantic search within thousands of podcasts transcriptions. Technologies used are: OpenAI Whisper for transcriptions, BERT embeddings and FAISS as a vector database

    2022 · voilib.com

  15. 15WD
  16. 16OS

    Our goal with this project is to build a completely open source, state of the art turn detection model that can be used in any voice AI application. I've been experimenting with LLM voice conversations since GPT-4 was first released. (There's a previous front page Show HN about Pipecat, the open source voice AI orchestration framework I work on. [1]) It's been almost two years, and for most of that time, I've been expecting that someone would "solve" turn detection. We all built initial, pretty good 80/20 versions of turn detection on top of VAD (voice activity detection) models. And…

    2025 · github.com

  17. 17CO

    Hey HN! Mike & Warren here from HyperDX (now part of ClickHouse)! We’ve been building ClickStack, an open source observability stack that helps you collect, centralize, search/viz/alert on your telemetry (logs, metrics, traces) in just a few minutes - all powered by ClickHouse (Apache2) for storage, HyperDX (MIT) for visualization and OpenTelemetry (Apache2) for ingestion. You can check out the quick start for spinning things up in the repo here: https://github.com/hyperdxio/hyperdx ClickStack makes it really easy to instrument your application so you can go…

    2025 · github.com

  18. 18
    Orate258

    The AI toolkit for speech

    2025

  19. 19

    Listen to your Google Search results

    2025

  20. 20IV
  21. 21
    Songbird220

    Shazam for bird songs

    2017

  22. 22
    Gumply70

    Whatever search, get in 5-minute audio

    2019

  23. 23
    Qwen3-TTS155

    Voice design, cloning & 97ms streaming

    Jan 2026

  24. 24
    SAM Audio146

    Segment any sound with text, visual, or time prompts

    Dec 2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →