nowfound

Alternatives

Products that do what Self-hosted gateway to access LLMs, Ollama, ComfyUI and FFmpeg servers does

  1. 1

    Use any AI model with just one API

    2025

  2. 2

    Connect, observe & control LLMs, MCPs, Guardrails & Prompts

    Dec 2025 · truefoundry.com

  3. 3
    ZenMux382

    An enterprise-grade LLM gateway with automatic compensation

    Feb 2026 · zenmux.ai

  4. 4

    One private gateway for every AI model

    Aug 2026 · ngrok.ai

  5. 5

    One AI gateway with built-in observability and evals

    Jun 2026 · respan.ai

  6. 6

    Automate self-hosting of commercial open source products

    2023

  7. 7

    Use AI models without managing keys or billing

    Dec 2025 · docs.netlify.com

  8. 8
    SUB/WAVE102

    Self-hosted radio with an AI DJ and one shared stream

    Jul 2026 · getsubwave.com

  9. 9SH
  10. 10OR

    Hi HN A few folks and I have been working on this project for a couple weeks now. After previously working on the Docker project for a number of years (both on the container runtime and image registry side), the recent rise in open source language models made us think something similar needed to exist for large language models too. While not exactly the same as running linux containers, running LLMs shares quite a few of the same challenges. There are "base layers" (e.g. models like Llama 2), specific configuration to run correctly (parameters, temperature, context window sizes etc). There's…

    2023 · github.com

  11. 11
    AskCodi230

    Custom LLMs, without training. Use via openai compatible api

    Nov 2025 · askcodi.com

  12. 12

    Avoid OpenAI downtimes - one API for 30+ LLMs

    2023

  13. 13

    Hi HN, we built an open source model gateway. It's a single place to manage our own self hosted, frontier, and open source models in one place. It’s is rust native, built for concurrency, and implements all the config quirks across models and providers (streaming formats, tool calls, model parameters, rate limits, and different error behavior). The gateway adds under 1 ms for BYOK requests and under 2 ms when Experiential supplies the provider key. It has every major inference provider, and 1000+ models refreshed daily via a codex agent that opens a PR. Compared to other similar projects…

    10d ago · github.com

  14. 14AL

    We built any-llm because we needed a lightweight router for LLM providers with minimal overhead. Switching between models is just a string change : update "openai/gpt-4" to "anthropic/claude-3" and you're done. It uses official provider SDKs when available, which helps since providers handle their own compatibility updates. No proxy or gateway service needed either, so getting started is pretty straightforward - just pip install and import. Currently supports 20+ providers including OpenAI, Anthropic, Google, Mistral, and AWS Bedrock. Would love to hear what you think!

    2025 · github.com

  15. 15WM

    Try it out! https://glhf.chat/ Hey HN! We’ve been working for the past few months on a website to let you easily run (almost) any open-source LLM on autoscaling GPU clusters. It’s free for now while we figure out how to price it, but we expect to be cheaper than most GPU offerings since we can run the models multi-tenant. Unlike Together AI, Fireworks, etc, we’ll run any model that the open-source vLLM project supports: we don’t have a hardcoded list. If you want a specific model or finetune, you don’t have to ask us for it: you can just paste the Hugging Face link in and…

    2024 · glhf.chat

  16. 16
    ChattyUI149

    Run open-source LLMs locally in the browser using WebGPU

    2024

  17. 17

    One balance. Every model. Chat, image, video & audio.

    Jun 2026 · lounge.llmgateway.io

  18. 18

    Self-hosted LLM router. Cost effective, deterministic, and fast. Secure and private by default. - Northwood-Systems/millwright

    Jul 2026 · github.com

  19. 19

    Open-source, OpenAI-compatible LLM gateway you run yourself. One endpoint for 40+ providers, with virtual keys, budgets, and usage tracking. - mozilla-ai/otari

    Jul 2026 · github.com

  20. 20

    One AI API for production - streaming, failover, logs

    Jan 2026 · modelriver.com

  21. 21AR

    Hi HN — we're the team behind Arch (https://github.com/katanemo/archgw), an open-source proxy for LLMs written in Rust. Today we're releasing Arch-Router (https://huggingface.co/katanemo/Arch-Router-1.5B), a 1.5B router model for preference-based routing, now integrated into the proxy. As teams integrate multiple LLMs - each with different strengths, styles, or cost/latency profiles — routing the right prompt to the right model becomes a critical part of the application design. But it's still an open problem. Most routing systems fall into two…

    2025

  22. 22

    Hey HNers - Riz here. I got together with a few guys and we built an LLM gateway. It's designed for small teams working on early-stage products, and can be deployed to AWS using a single command (i.e. `mantis deploy`). It's self-hosted, and is designed to belong to you.

    Jun 2026 · github.com

  23. 23AO

    Hi HN, I've been developing Portkey Gateway, an open-source AI gateway that's now processing billions of tokens daily across 200+ LLMs. Today, we're launching a significant update: integrated Guardrails at the gateway level. Key technical features: 1. Guardrails as middleware: We've implemented a hooks architecture that allows guardrails to act as middleware in the request/response flow. This enables real-time LLM output evaluation and transformation. 2. Flexible orchestration: The gateway can now route requests based on guardrail verdicts. This allows for complex logic like fallbacks…

    2024 · github.com

  24. 24

    The intelligent, self-hosted LLM gateway

    Apr 2026 · provara.xyz

Ranked by how close each launch is in meaning, then by votes. Refine with a description →