nowfound

Alternatives

Products that do what W&B Training by Weights & Biases does

The fast and easy way to train AI agents with serverless RL

  1. 1

    RL-training an AI agent to RL-train AI agents. Contribute to Danau5tin/ai-trains-ai development by creating an account on GitHub.

    Jul 2026 · github.com

  2. 2TB

    After training calculator agent via RL, I really wanted to go bigger! So I built RL infrastructure for training long-horizon terminal/coding agents that scales from 2x A100s to 32x H100s (~$1M worth of compute!) Without any training, my 32B agent hit #19 on Terminal-Bench leaderboard, beating Stanford's Terminus-Qwen3-235B-A22! With training... well, too expensive, but I bet the results would be good! *What I did*: - Created a Claude Code-inspired agent (system msg + tools) - Built Docker-isolated GRPO training where each rollout gets its own container - Developed a multi-agent…

    2025 · github.com

  3. 3AA

    Hey HN, I wanted to share a new project we've been working on for the last couple of months called ART (https://github.com/OpenPipe/ART). ART is a new open-source framework for training agents using reinforcement learning (RL). RL allows you to train an agent to perform better at any task whose outcome can be measured and quantified. There are many excellent projects focused on training LLMs with RL, such as GRPOTrainer (https://huggingface.co/docs/trl/main/en/grpo_trainer) and verl…

    2025 · github.com

  4. 4

    Learn any language with AI in a fun, easy, personalized way

    2024

  5. 5

    Build AI agents that respond with UI instead of text

    Feb 2026 · thesys.dev

  6. 6

    Run 100s of coding agents on any machine from anywhere

    May 2026 · superset.sh

  7. 7

    Generate SEO-optimized content 10X faster with AI

    2023

  8. 8FD

    I worked on this applied Deep Reinforcement Learning course for the better part of 2021. I made a Datacamp course [0] before, and this served as my inspiration to make an applied Deep RL series. Normally, Deep RL courses teach a lot of mathematically involved theory. You get the practical applications near the end (if at all). I have tried to turn that on its head. In the top-down approach, you learn practical skills first, then go deeper later. This is much more fun. This course (the first in a planned multi-part series) shows how to use the Deep Reinforcement Learning framework RLlib to…

    2022 · courses.dibya.online

  9. 9

    Train AI chatbots in 3 clicks and help customers instantly

    2024

  10. 10

    Fine-tuning, RL, and inference in one CLI

    Dec 2025 · github.com

  11. 11ET
  12. 12

    No-code AI Lab: Train models, access datasets, run inference

    Feb 2026 · neuro-block.com

  13. 13

    Teach AI to learn from your workflows

    2024

  14. 14NG
  15. 15
    Synapis126

    Build & deploy AI models online – no code or hardware needed

    2025

  16. 16

    Build custom AI tools without code

    2023

  17. 17

    Vibe Training AI models

    Mar 2026

  18. 18
    Trainer93

    Train AI agents by recording your screen

    May 2026 · myagentrainer.com

  19. 19L8

    I've been tinkering with getting Llama-8B to bootstrap its own research skills through self-play. The model generates questions about documents, searches for answers, and then learns from its own successes/failures through RL (hacked up Unsloth's GRPO code). Started with just 23% accuracy on Apollo 13 mission report questions and hit 53% after less than an hour of training. Everything runs locally using open-source models. It's cool to see the model go from completely botching search queries to iteratively researching to get the right answer.

    2025 · github.com

  20. 20RA

    Hey everyone! Along with my team, I've developed a reinforcement learning system that automatically optimizes LLM prompts, complete with a visualization feature to track both prompt structure and learning progress over time. Take a look here: https://nomadic-ml.github.io/nomadic/cookbooks/Nomadic_Promp... Check out our website too:https://www.nomadicml.com/ In terms of how this visualization works: The RL Prompt Optimizer employs a reinforcement learning framework to iteratively improve prompts used for language model evaluations. At each episode, the…

    2024 · nomadic-ml.github.io

  21. 21

    Train and deploy AI agents in minutes to your website

    2025

  22. 22IB

    This integration allows for scalable evals and training of browser agents with hosted Prime Intellect eval + training pipelines and headless browser infrastructure on Browserbase to RL train browser agents with LoRA.

    Mar 2026 · github.com

  23. 23

    AI Journey Starts Here

    2025

  24. 24SR

Ranked by how close each launch is in meaning, then by votes. Refine with a description →