nowfound

Alternatives

Products that do what ModelPointer does

Open-source AI Gateway for enterprise LLM infrastructure

  1. 1

    Use any AI model with just one API

    2025

  2. 2

    Hi HN, we built an open source model gateway. It's a single place to manage our own self hosted, frontier, and open source models in one place. It’s is rust native, built for concurrency, and implements all the config quirks across models and providers (streaming formats, tool calls, model parameters, rate limits, and different error behavior). The gateway adds under 1 ms for BYOK requests and under 2 ms when Experiential supplies the provider key. It has every major inference provider, and 1000+ models refreshed daily via a codex agent that opens a PR. Compared to other similar projects…

    10d ago · github.com

  3. 3

    One AI API for production - streaming, failover, logs

    Jan 2026

  4. 4

    Connect, observe & control LLMs, MCPs, Guardrails & Prompts

    Dec 2025

  5. 5GA

    Hi, I’m Jakub, a solo founder based in Warsaw. I’ve been building GoModel since December with a couple of contributors. It's an open-source AI gateway that sits between your app and model providers like OpenAI, Anthropic or others. I built it for my startup to solve a few problems: - track AI usage and cost per client or team - switch models without changing app code - debug request flows more easily - reduce AI spendings with exact and semantic caching How is it different? - ~17MB docker image - LiteLLM's image is more than 44x bigger ("docker.litellm.ai/berriai/litellm:latest" ~…

    Apr 2026 · github.com

  6. 6
    LLMWare358

    Dev tool to make AI apps to deploy privately or locally

    2024

  7. 7

    One private gateway for every AI model

    Aug 2026 · ngrok.ai

  8. 8

    Use AI models without managing keys or billing

    Dec 2025

  9. 9
    Sudo AI265

    One API for any LLM— routing, context, and monetization

    Sep 2025

  10. 10SM

    We built a model router that plugs into coding agents (e.g. Claude Code, Codex, Cursor, etc.) and intelligently sends requests to the best model to serve them. Here's a quick demo of running it locally: https://www.youtube.com/watch?v=isKhAyivtfM. At Weave, we write most of our code with AI, and it's been getting more expensive. This came to a head when Opus 4.7 was released and, thanks to its tokenizer changes, our costs shot up. We knew we didn't need Opus for everything but we didn't want to lose out on the intelligence for the cases where you really need it. So we decided…

    Jun 2026 · github.com

  11. 11
    Mindware281

    API gateway for AI agents

    2024

  12. 12

    Open source AI gateway turning traffic into a better model

    3d ago · experientiallabs.ai

  13. 13

    Avoid OpenAI downtimes - one API for 30+ LLMs

    2023

  14. 14AR

    Hi HN — we're the team behind Arch (https://github.com/katanemo/archgw), an open-source proxy for LLMs written in Rust. Today we're releasing Arch-Router (https://huggingface.co/katanemo/Arch-Router-1.5B), a 1.5B router model for preference-based routing, now integrated into the proxy. As teams integrate multiple LLMs - each with different strengths, styles, or cost/latency profiles — routing the right prompt to the right model becomes a critical part of the application design. But it's still an open problem. Most routing systems fall into two…

    2025

  15. 15

    The open-source AI gateway for AI-native startups

    Nov 2025

  16. 16

    LLM Provider arbitrage to get the best performance for the $

    2025

  17. 17

    Serve Any AI Model, Faster & Cheaper

    Mar 2026

  18. 18

    Optimize Performance, Cost, Speed & Carbon for each prompt

    Nov 2025

  19. 19

    Global APIs as MCP powered by AI Gateway

    2025

  20. 20RL

    Hello Hacker News! We're Yangqing, Xiang and JJ from lepton.ai. We are building a platform to run any AI models as easy as writing local code, and to get your favorite models in minutes. It's like container for AI, but without the hassle of actually building a docker image. We built and contributed to some of the world's most popular AI software - PyTorch 1.0, ONNX, Caffe, etcd, Kubernetes, etc. We also managed hundreds of thousands of computers in our previous jobs. And we found that the AI software stack is usually unnecessarily complex - and we want to change that. Imagine if you are a…

    2023 · lepton.ai

  21. 21AL

    We are Rohit & Ayush, we created Portkey this year March to help tackle some challenges we had seen while building apps based on GPT3, 3.5, 4, and the DevOps principles we brought to the scene to help tackle them. We believe, a solid, performant, and reliable gateway lays the foundation to help build the next level of LLM apps. It decreases excessive reliance on any one company and takes the focus back to building instead of spending time fixing the nitty gritties of different providers and making them work together. Features: Blazing fast (9.9x faster) with a tiny footprint (~45kb…

    2024 · github.com

  22. 22TO

    Hi HN! We're Gabriel & Viraj, and we're excited to open source TensorZero. To be a little cheeky, TensorZero is an open-source platform that helps LLM applications graduate from API wrappers into defensible AI products. 1. Integrate our model gateway 2. Send metrics or feedback 3. Unlock compounding improvements in quality, cost, and latency It enables a data & learning flywheel for LLMs by unifying: • Inference: one API for all LLMs, with <1ms P99 overhead • Observability: inference & feedback → your database • Optimization: better prompts, models, inference strategies • Experimentation:…

    2024 · github.com

  23. 23
    Agihalo68

    LLM Router for A.I Agent & Saas with x402

    Jan 2026

  24. 24LT

    Current AI-assisted CLI tools are often part of larger systems and work better on Linux. I built llm-term to address these. It's a Rust-based tool that compiles into a single binary file. You only need to download the binary, add it to your PATH, and configure your OpenAI key to get started. While llm-term offers an option for gpt-4o, it works great with gpt-4o-mini. So it's not costly. I appreciate any feedback or suggestions.

    2024 · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →