Alternatives
Products that do what MiMo-Audio does
Audio language models are few-shot learners
- 1

- 2

- 3

- 4

- 5

- 6

- 7

- 8

- 9

- 10

- 11

- 12

- 13PO
Hi HN, OpenAI recently released a model for automatic speech recognition called Whisper [0]. I decided to reimplement the inference of the model from scratch using C/C++. To achieve this I implemented a minimalistic tensor library in C and ported the high-level architecture of the model in C++. The entire code is less than 8000 lines of code and is contained in just 2 source files without any third-party dependencies. The Github project is here: https://github.com/ggerganov/whisper.cpp With this implementation I can very easily build and run the model - “make…
2022 · github.com
- 14

- 15
- 16

- 17

- 18

- 19

- 20

First TTS model to support all 22 Indic languages + English
2024
- 21

- 22

- 23

- 24

Listen to your favorite podcasters in your native language
2023
Ranked by how close each launch is in meaning, then by votes. Refine with a description →