Alternatives
Products that do what Mellum by JetBrains does
Fast LLMs for low-latency and high-performance workflows
- 1IB
Built a ~9M param LLM from scratch to understand how they actually work. Vanilla transformer, 60K synthetic conversations, ~130 lines of PyTorch. Trains in 5 min on a free Colab T4. The fish thinks the meaning of life is food. Fork it and swap the personality for your own character.
Apr 2026 · github.com
- 2TV
May 2026 · github.com
- 3SU
Here's a project I've been working on for the last few months. It's a new (I think) algorithm, that allows to adjust smoothly - and in real time - how many calculations you'd like to do during inference of an LLM model. It seems that it's possible to do just 20-25% of weight multiplications instead of all of them, and still get good inference results. I implemented it to run on M1/M2/M3 GPU. The mmul approximation itself can be pushed to run 2x fast before the quality of output collapses. The inference speed is just a bit faster than Llama.cpp's, because the rest of implementation…
2024 · asciinema.org
- 4

- 5GG
A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me. But then I thought, "I wonder how it would work on a normal computer like mine," and above all, "I wonder if it would work without going into OOM on a computer like mine." So I started working with the help of agents to test this possibility. I started converting the model to int4, understanding MTP usage, and if possible implementing DSA for long context.…
Jul 2026 · github.com
- 6TL
2023 · tinyllms.vercel.app
- 7

- 8

- 9

- 10PT
2024 · adamgrant.info
- 11

First TTS model to support all 22 Indic languages + English
2024
- 12TO
Mar 2026 · github.com
- 13FT
May 2026 · github.com
- 14

- 15CA
Hi HN! We’re been working hard on this low-code tool for rapid prompt discovery, robustness testing and LLM evaluation. We’ve just released documentation to help new users learn how to use it and what it can already do. Let us know what you think! :)
2023 · chainforge.ai
- 16MA
2023 · github.com
- 17

- 18

- 19

Fast, accurate STT for production-grade voice agents
May 2026 · ringg.ai
- 20

- 21

Working on Mac, Linux, and Windows now. I include a simple GUI to find new models and get things built and set up. It is working quite well across a few models for me. The GitHub README and DESIGN.md files go into detail of the how/why and it's working remarkably well so far. https://github.com/notactuallytreyanastasio/shoehorn
19d ago · notactuallytreyanastasio.github.io
- 22

- 23

Multilingual TTS model with realistic and expressive speech
Mar 2026 · mistral.ai
- 24EL
2023 · github.com
Ranked by how close each launch is in meaning, then by votes. Refine with a description →