nowfound

AI · alternatives · 2026

24 alternatives to RLLama

Empowering LLMs with Memory-Augmented Reinforcement Learning

Below are 24 products that do a similar job, ranked by how close each is in meaning and then by launch-day votes.

  1. 1LF
  2. 2WT

    After working with LLMs for long enough, I found myself wanting a lightweight utility for doing various small tasks to prepare inputs, locate information and create evaluators. This library is two things: a very simple model and utilities that inference it (eg. fuzzy deduplication). The target platform is CPU, and it’s intended to be light, fast and pip installable — a library that lowers the barrier to working with strings semantically. You don’t need to install pytorch to use it, or any deep learning runtimes. How can this be accomplished? The model is simply token embeddings that are…

    2024 · github.com · its alternatives →

  3. 3IB

    Built a ~9M param LLM from scratch to understand how they actually work. Vanilla transformer, 60K synthetic conversations, ~130 lines of PyTorch. Trains in 5 min on a free Colab T4. The fish thinks the meaning of life is food. Fork it and swap the personality for your own character.

    Apr 2026 · github.com · its alternatives →

  4. 4LL
  5. 5

    LLM reinforcement fine-tuning platform to improve LLM output

    2025 · its alternatives →

  6. 6RA

    Hey everyone! Along with my team, I've developed a reinforcement learning system that automatically optimizes LLM prompts, complete with a visualization feature to track both prompt structure and learning progress over time. Take a look here: https://nomadic-ml.github.io/nomadic/cookbooks/Nomadic_Promp... Check out our website too:https://www.nomadicml.com/ In terms of how this visualization works: The RL Prompt Optimizer employs a reinforcement learning framework to iteratively improve prompts used for language model evaluations. At each episode, the…

    2024 · nomadic-ml.github.io · its alternatives →

  7. 7AA

    Hey HN, I wanted to share a new project we've been working on for the last couple of months called ART (https://github.com/OpenPipe/ART). ART is a new open-source framework for training agents using reinforcement learning (RL). RL allows you to train an agent to perform better at any task whose outcome can be measured and quantified. There are many excellent projects focused on training LLMs with RL, such as GRPOTrainer (https://huggingface.co/docs/trl/main/en/grpo_trainer) and verl…

    2025 · github.com · its alternatives →

  8. 8

    RL-training an AI agent to RL-train AI agents. Contribute to Danau5tin/ai-trains-ai development by creating an account on GitHub.

    Jul 2026 · github.com · its alternatives →

  9. 9

    Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM - jmaczan/tiny-vllm

    May 2026 · github.com · its alternatives →

  10. 10SP

    I built a system that lets LLMs automatically learn and improve problem-solving strategies over time, inspired by Andrej Karpathy's idea of a "third paradigm" for LLM learning. The basic idea: instead of using static system prompts, the LLM builds up a database of strategies that actually work for different problem types. When you give it a new problem, it selects the most relevant strategies, applies them, then evaluates how well they worked and refines them. For example, after seeing enough word problems, it learned this strategy: 1) Read carefully and identify unknowns, 2) Define…

    2025 · its alternatives →

  11. 11

    Fine-tuning, RL, and inference in one CLI

    Dec 2025 · github.com · its alternatives →

  12. 12CL

    I have a proposal that addresses long-term memory problems for LLMs when new data arrives continuously (cheaply!). The program involves no code, but two Markdown files. For retrieval, there is a semantic filesystem that makes it easy for LLMs to search using shell commands. It is currently a scrappy v1, but it works better than anything I have tried. Curious for any feedback!

    Apr 2026 · github.com · its alternatives →

  13. 13
    Taylor AI▲118

    Fine-tune open source LLMs in minutes

    2023 · its alternatives →

  14. 14

    Unlock your knowledge with 2000 LLM prompts

    2023 · its alternatives →

  15. 15
    CogniMemo▲120

    AI that remembers, learns, and evolves

    Dec 2025 · app.cognimemo.com · its alternatives →

  16. 16
    LearnTime▲105

    LMS platform

    2024 · its alternatives →

  17. 17LA

    Hey Hacker News! I've been working on an open-source project called LLM Alignment Template, a comprehensive toolkit designed to help researchers, developers, and data scientists align large language models (LLMs) with human values using Reinforcement Learning from Human Feedback (RLHF). What the project does: Interactive Web Interface: Easily train models, visualize alignment metrics, and manage alignment with an accessible UI. Training with RLHF: Align models effectively to human preferences using feedback loops. Explainability: Built-in dashboards to help understand model behavior using…

    2024 · github.com · its alternatives →

  18. 18

    The missing memory layer for generative AI

    2024 · its alternatives →

  19. 19L8

    I've been tinkering with getting Llama-8B to bootstrap its own research skills through self-play. The model generates questions about documents, searches for answers, and then learns from its own successes/failures through RL (hacked up Unsloth's GRPO code). Started with just 23% accuracy on Apollo 13 mission report questions and hit 53% after less than an hour of training. Everything runs locally using open-source models. It's cool to see the model go from completely botching search queries to iteratively researching to get the right answer.

    2025 · github.com · its alternatives →

  20. 20

    A single memory for all your LLMs

    Nov 2025 · hiperyon.com · its alternatives →

  21. 21LA

    A proof-of-concept language learning app that uses LLMs to generate definitions of unknown words using only previously mastered vocabulary.

    Dec 2025 · simedw.com · its alternatives →

  22. 22IB
  23. 23

    The context manager and skills library for marketing teams

    Apr 2026 · promptr.ai · its alternatives →

  24. 24NG

Also compare

Ranked by how close each launch is in meaning, then by votes. Prices were read from each product’s own site when checked and can change. Refine with your own description →