nowfound

Alternatives

Products that do what Reame – a CPU inference server that gets faster as it runs does

  1. 1
    Reame8

    Self-hosted LLM inference on the hardware you already have

    Jul 2026 · github.com

  2. 2
    Banana235

    Serverless GPUs for Machine Learning inference

    2022

  3. 3
    Rerun308

    The easiest way to build AI agents for all your tasks

    Jul 2026 · rerun.build

  4. 4

    Boost your social media confidence like a native speaker

    2023

  5. 5
    Re:scam342

    I’m an AI chatbot created to send scammers a message.

    2017

  6. 6
    Reword71

    Rewrite messages without leaving your workflow

    Jan 2026

  7. 7
    Groq®237

    Hyperfast LLM running on custom built GPUs

    2024

  8. 8LA
  9. 9
    MindsDB235

    In-database machine learning

    2022

  10. 10
    Rerun151

    Visualize computer vision

    2023

  11. 11
    Rehance204

    The AI copilot for hardcore SaaS

    2024

  12. 12

    Fast multimodal-native inference at scale

    Dec 2025

  13. 13
    Rembo79

    A chatbot that reminds you of your tasks by calling you

    2019

  14. 14
    Retrace101

    Debug AI agents by replaying and forking runs

    Jul 2026 · retraceai.tech

  15. 15
    RELE.AI96

    Virtual Assistant

    2020

  16. 16
    anon110

    Machine learning, automated

    2021

  17. 17MI

    2022 · max.io

  18. 18RE
  19. 19
    req71

    A lightweight, minimal yet powerful API testing tool

    2022

  20. 20

    An OKF-backed Model Context Protocol (MCP) server delivering persistent long-term memory and SQLite FTS5 search for AI agents. - fellowgeek/mcp-memory

    24d ago · github.com

  21. 21BO

    Read the full blogpost at https://rach.codes/blog/Introducing-Bhumi (click on reader to see the technical breakdown!) AI inference should be fast, but in practice it’s painfully slow. Inference bottlenecks slow down LLM-powered chatbots and AI workflows everywhere. I built Bhumi to fix that. Bhumi is a Python library designed for developers, yet its performance-critical core is implemented in Rust (via PyO3) for near-native speed. This hybrid approach delivers up to 2.5x faster response times across providers like OpenAI, Anthropic, and Gemini—without changing the…

    2025 · bhumi.trilok.ai

  22. 22GA

    2021 · inferrd.com

  23. 23RS

    What relai-sdk is an open-source toolkit for making AI agents reliable via a complete learning loop: simulate → evaluate → optimize. Why Agent runs are stochastic; tool-calls fail; hard to reproduce, measure, and fix at scale. It’s also hard to align behavior with goals across output quality/format, cost, and latency. We need a loop that integrates user feedback and LLM evaluators directly into the agent code (prompts, configs, models, graphs) without overfitting. How - Simulation: LLM personas, mocked MCP servers/tools, synthetic data; can condition on real traces - Evaluation:…

    Oct 2025 · github.com

  24. 24FS

    Hi everyone! I've been loving building with AI, and over the past few years I've been leaning more and more into Typescript (and bun). My team at inference.net is constantly trying to get more leverage out of AI and find ways to setup our codebase to be able to increase the level of correctness that our AI is able to write code at. This starter repo is a very opinionated way to lay out a repo to lean into AI heavily. It leverages Cloudflare Workers as a deployment target for the API (my goal is to never have to deploy an API on a AWS/Azure/GCP server ever again unless I get to a…

    2025 · abeahmed.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →