AI · alternatives · 2026

24 alternatives to RLLama
Empowering LLMs with Memory-Augmented Reinforcement Learning
Below are 24 products that do a similar job, ranked by how close each is in meaning and then by launch-day votes.
- 1LF
2024 · github.com · its alternatives →
- 2WT
After working with LLMs for long enough, I found myself wanting a lightweight utility for doing various small tasks to prepare inputs, locate information and create evaluators. This library is two things: a very simple model and utilities that inference it (eg. fuzzy deduplication). The target platform is CPU, and it’s intended to be light, fast and pip installable — a library that lowers the barrier to working with strings semantically. You don’t need to install pytorch to use it, or any deep learning runtimes. How can this be accomplished? The model is simply token embeddings that are…
2024 · github.com · its alternatives →
- 3IB
Built a ~9M param LLM from scratch to understand how they actually work. Vanilla transformer, 60K synthetic conversations, ~130 lines of PyTorch. Trains in 5 min on a free Colab T4. The fish thinks the meaning of life is food. Fork it and swap the personality for your own character.
Apr 2026 · github.com · its alternatives →
- 4LL
2025 · github.com · its alternatives →
- 5

LLM reinforcement fine-tuning platform to improve LLM output
2025 · its alternatives →
- 6RA
Hey everyone! Along with my team, I've developed a reinforcement learning system that automatically optimizes LLM prompts, complete with a visualization feature to track both prompt structure and learning progress over time. Take a look here: https://nomadic-ml.github.io/nomadic/cookbooks/Nomadic_Promp... Check out our website too:https://www.nomadicml.com/ In terms of how this visualization works: The RL Prompt Optimizer employs a reinforcement learning framework to iteratively improve prompts used for language model evaluations. At each episode, the…
2024 · nomadic-ml.github.io · its alternatives →
- 7AA
Hey HN, I wanted to share a new project we've been working on for the last couple of months called ART (https://github.com/OpenPipe/ART). ART is a new open-source framework for training agents using reinforcement learning (RL). RL allows you to train an agent to perform better at any task whose outcome can be measured and quantified. There are many excellent projects focused on training LLMs with RL, such as GRPOTrainer (https://huggingface.co/docs/trl/main/en/grpo_trainer) and verl…
2025 · github.com · its alternatives →
- 8

RL-training an AI agent to RL-train AI agents. Contribute to Danau5tin/ai-trains-ai development by creating an account on GitHub.
Jul 2026 · github.com · its alternatives →
- 9

Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM - jmaczan/tiny-vllm
May 2026 · github.com · its alternatives →
- 10SP
I built a system that lets LLMs automatically learn and improve problem-solving strategies over time, inspired by Andrej Karpathy's idea of a "third paradigm" for LLM learning. The basic idea: instead of using static system prompts, the LLM builds up a database of strategies that actually work for different problem types. When you give it a new problem, it selects the most relevant strategies, applies them, then evaluates how well they worked and refines them. For example, after seeing enough word problems, it learned this strategy: 1) Read carefully and identify unknowns, 2) Define…
2025 · its alternatives →
- 11

Fine-tuning, RL, and inference in one CLI
Dec 2025 · github.com · its alternatives →
- 12CL
I have a proposal that addresses long-term memory problems for LLMs when new data arrives continuously (cheaply!). The program involves no code, but two Markdown files. For retrieval, there is a semantic filesystem that makes it easy for LLMs to search using shell commands. It is currently a scrappy v1, but it works better than anything I have tried. Curious for any feedback!
Apr 2026 · github.com · its alternatives →
- 13

- 14

Unlock your knowledge with 2000 LLM prompts
2023 · its alternatives →
- 15
CogniMemo▲120AI that remembers, learns, and evolves
Dec 2025 · app.cognimemo.com · its alternatives →
- 16

- 17LA
Hey Hacker News! I've been working on an open-source project called LLM Alignment Template, a comprehensive toolkit designed to help researchers, developers, and data scientists align large language models (LLMs) with human values using Reinforcement Learning from Human Feedback (RLHF). What the project does: Interactive Web Interface: Easily train models, visualize alignment metrics, and manage alignment with an accessible UI. Training with RLHF: Align models effectively to human preferences using feedback loops. Explainability: Built-in dashboards to help understand model behavior using…
2024 · github.com · its alternatives →
- 18
- 19L8
I've been tinkering with getting Llama-8B to bootstrap its own research skills through self-play. The model generates questions about documents, searches for answers, and then learns from its own successes/failures through RL (hacked up Unsloth's GRPO code). Started with just 23% accuracy on Apollo 13 mission report questions and hit 53% after less than an hour of training. Everything runs locally using open-source models. It's cool to see the model go from completely botching search queries to iteratively researching to get the right answer.
2025 · github.com · its alternatives →
- 20

- 21LA
A proof-of-concept language learning app that uses LLMs to generate definitions of unknown words using only previously mastered vocabulary.
Dec 2025 · simedw.com · its alternatives →
- 22IB
2025 · github.com · its alternatives →
- 23

The context manager and skills library for marketing teams
Apr 2026 · promptr.ai · its alternatives →
- 24NG
2024 · github.com · its alternatives →
Also compare
- LlamaGym – fine-tune LLM agents with online reinforcement learning alternatives
- Wordllama – Things you can do with the token embeddings of an LLM alternatives
- I built a tiny LLM to demystify how language models work alternatives
- Learn LLMs LeetCode Style alternatives
- Predibase Reinforcement Fine-Tuning alternatives
- RL Agent that can auto-optimize your LLM prompts alternatives
Ranked by how close each launch is in meaning, then by votes. Prices were read from each product’s own site when checked and can change. Refine with your own description →