nowfound

Alternatives

Products that do what Cloning a musical instrument from 16 seconds of audio does

In 2020, Magenta released DDSP [1], a machine learning algorithm / python library which made it possible to generate good sounding instrument synthesizers from about 6-10 minutes of data. While working with DDSP for a project, we realised how it was actually quite hard to find 6-10 minute of clean recordings of monophonic instruments. In this project, we have combined the DDSP architecture with a domain adaptation technique from speech synthesis [2]. This domain adaptation technique works by pre-training our model on many different recordings from the Solos dataset [3] first and then…

  1. 1

    High-quality voice clones with just 60 seconds of audio

    2023

  2. 2AA
  3. 3FA

    Hi all, I've spent some time working on music demixing or music source separation algorithms, which take in a mixed song and output estimates of isolated components (e.g. vocals, drums, bass, other). I took a popular PyTorch model with good performance (Open-Unmix, UMX-L weights), reimplemented the inference steps in C++, and compiled it to WebAssembly for a free client-side music demixer.

    2023 · sevag.xyz

  4. 4BM

    If you already know what bytebeat is, you don't need an explanation. If not, check my project :) Here is how it looks: SELECT mono(output( arraySum(x -> 1 / 6 * running_envelope(30 * (1 + x / 6), time, 0.05 * x, 0.005, lfo(0, 0.25, sine_wave, time / 8), 0.1) * sine_wave(time * 80 * exp2(x / 3)), range(12)))) FROM table; To check how it sounds, find the examples in the repository https://github.com/ClickHouse/NoiSQL

    2023 · github.com

  5. 5T3
  6. 6T1
  7. 7

    Generative AI for audio made simple

    2023

  8. 8MA

    I am excited to announce a new tool for music producers and audio enthusiasts - a music audio search engine. With just a simple description of the groove you're looking for, our semantic search engine will output the most similar audio in seconds. I used the Freesound.org API to upload over 3,000 grooves to MongoDB, and combined all the relevant data such as tags, title, description, BPM, etc. into OpenAI's Text-Davinci to generate a unique description of each sound. I then embedded these descriptions using the Ada Embeddings Model and inserted them into Pinecone DB vector database, making…

    2023 · muzic-sage.vercel.app

  9. 9
    Udio251

    Generative song creator & mobile studio. Make your music.

    2025

  10. 10IM

    I recently got Ableton Move and was disappointed by how difficult it was to convert samples into Move-compatible drum racks, so I made a simple CLI tool for that. I hope others find it useful, and at some point, Ableton sees the need and provides us with a built-in solution.

    2024 · github.com

  11. 11IU

    Hi Hacker News, This is definitely out of my comfort zone. I just wanted to show you guys because I'm super proud of it. It's a 100% faithful recreation based off of the schematics, patents, and ROMs that were found online. So please watch the video and tell me what you think https://youtu.be/auOlZXI1VxA The reason why I think this is relevant is because I've been a programmer for 25 years and AI scares the shit out of me. I'm not a programmer anymore. I'm something else now. I don't know what it is but it's multi-disciplinary, and it doesn't involve writing code myself--for…

    Jan 2026

  12. 12WA
  13. 13KS
  14. 14WA

    Hi, I'm the author of this little Web Audio toy which does physical modeling synthesis using a simple spring-mass system. My current area of research is in sparse, event-based encodings of musical audio (https://blog.cochlea.xyz/sparse-interpretable-audio-codec-pa...). I'm very interested in decomposing audio signals into a description of the "system" (e.g., room, instrument, vocal tract, etc.) and a sparse "control signal" which describes how and when energy is injected into that system. This toy was a great way to start learning about physical modeling synthesis, which seems…

    2025 · blog.cochlea.xyz

  15. 15
    VoxCPM2110

    Open-source 48kHz TTS with voice design and cloning

    Apr 2026 · github.com

  16. 16IF

    Hi HN, Last time I showed free-music-demixer, which people seemed to enjoy. It was a static website with a Javascript + WASM module to perform music demixing (or music source separation) using an AI model UMX-L (Open-Unmix) running client-side in the browser. Since then, I have overhauled the project and made several improvements: - The demixing/separation quality is higher now, since I implemented the missing post-processing step - Memory usage is lower now by performing a custom segmented inference with a streaming LSTM, which should allow larger tracks (or, dare I say,…

    2023 · freemusicdemixer.com

  17. 17WI

    I built k-synth as an experiment to see if a minimalist, K-inspired array language could make sketching waveforms faster and more intuitive than traditional code. I’ve put together a web-based toolkit so you can try the syntax directly in the browser without having to touch a compiler: Live Toolkit: https://octetta.github.io/k-synth/ If you visit the page, here is a quick path to an audio payoff: - Click "patches" and choose dm-bell.ks. - Click "run"—the notebook area will update. Click the waveform to hear the result. - Click the "->0" button below the waveform to copy…

    Mar 2026 · octetta.github.io

  18. 18BW
  19. 19FO

    2020 · signal.vercel.app

  20. 20

    Just tell your guitar pedal how to sound

    May 2026 · jamtime.ai

  21. 21TA
  22. 22

    Create music using AI

    2024

  23. 23

    Make music & art using machine learning

    2018

  24. 24WD

    2018 · dsp.audio

Ranked by how close each launch is in meaning, then by votes. Refine with a description →