nowfound

Alternatives

Products that do what Nexus Gateway does

Nexus Gateway – Reduce LLM API Costs Using Semantic Caching

  1. 1

    Use any AI model with just one API

    2025

  2. 2
    Nexus494

    The first AI navigator for your entire network

    2023

  3. 3
    ZenMux382

    An enterprise-grade LLM gateway with automatic compensation

    Feb 2026 · zenmux.ai

  4. 4
    Bifrost575

    The fastest LLM gateway in the market

    2025

  5. 5

    Use AI models without managing keys or billing

    Dec 2025 · docs.netlify.com

  6. 6

    Cut LLM Costs 30-80% 2-Minute Setup.

    Dec 2025 · costbase.ai

  7. 7
    Edgee196

    The AI Gateway that TL;DR tokens

    Feb 2026 · edgee.ai

  8. 8

    AI sales assistant that finds leads + books meetings for you

    Feb 2026 · nexuscale.ai

  9. 9AL

    We built any-llm because we needed a lightweight router for LLM providers with minimal overhead. Switching between models is just a string change : update "openai/gpt-4" to "anthropic/claude-3" and you're done. It uses official provider SDKs when available, which helps since providers handle their own compatibility updates. No proxy or gateway service needed either, so getting started is pretty straightforward - just pip install and import. Currently supports 20+ providers including OpenAI, Anthropic, Google, Mistral, and AWS Bedrock. Would love to hear what you think!

    2025 · github.com

  10. 10

    Open source AI gateway turning traffic into a better model

    3d ago · experientiallabs.ai

  11. 11

    Ollama but for mobile, with a cloud fallback

    2025

  12. 12
    ReliAPI87

    Stop losing money on failed OpenAI and Anthropic API calls.

    Dec 2025 · kikuai-lab.github.io

  13. 13
    Nexus61

    AI-managed code quality and workflows built for AI teams

    2025

  14. 14AO

    Hi HN, I've been developing Portkey Gateway, an open-source AI gateway that's now processing billions of tokens daily across 200+ LLMs. Today, we're launching a significant update: integrated Guardrails at the gateway level. Key technical features: 1. Guardrails as middleware: We've implemented a hooks architecture that allows guardrails to act as middleware in the request/response flow. This enables real-time LLM output evaluation and transformation. 2. Flexible orchestration: The gateway can now route requests based on guardrail verdicts. This allows for complex logic like fallbacks…

    2024 · github.com

  15. 15

    The AI gateway that regulated industries need to deploy AI

    Jun 2026 · alphabitcore.com

  16. 16

    Self-hosted LLM router. Cost effective, deterministic, and fast. Secure and private by default. - Northwood-Systems/millwright

    Jul 2026 · github.com

  17. 17AL

    We are Rohit & Ayush, we created Portkey this year March to help tackle some challenges we had seen while building apps based on GPT3, 3.5, 4, and the DevOps principles we brought to the scene to help tackle them. We believe, a solid, performant, and reliable gateway lays the foundation to help build the next level of LLM apps. It decreases excessive reliance on any one company and takes the focus back to building instead of spending time fixing the nitty gritties of different providers and making them work together. Features: Blazing fast (9.9x faster) with a tiny footprint (~45kb…

    2024 · github.com

  18. 18
    COO9

    Llm Gateway

    Oct 2025

  19. 19

    Cuts your LLM API costs by 40-70%. One line of code.

    May 2026 · semanticguard.dev

  20. 20

    Your all-access pass to top AI models — free & fast

    Nov 2025

  21. 21SA

    Hi HN, We’re building https://www.switchpoint.dev – a drop-in replacement for OpenAI’s API that reduces LLM cost by smartly routing across models (e.g., Claude, Gemini, GPT-4) depending on subject and difficulty of the task. Why we built this: LLM costs are spiraling—especially for products doing retrieval, agentic reasoning, or even just high-volume chat. We were frustrated with paying GPT-4 rates when most queries didn’t need it. So we built a router that: - Starts with cheaper/free models (like Llama 8B, 4o-mini, 2.0 flash) - Streams responses and upgrades on failure - Acts…

    2025 · switchpoint.dev

  22. 22

    Your all-access pass to top AI models — free & fast

    Nov 2025

  23. 23

    An AI Cost Optimization Infrastructure for LLM Applications

    Mar 2026 · getpromptly.in

  24. 24RC

    Hello HN! We're building a caching solution for LLMs (ChatGPT, Claude). By combining cutting-edge approaches, such as edge computing, prompt compression, vectorization, and others - it can reduce your AI bills by up to 10x and significantly lower response times. Key Features: - cost efficiency: our system stores frequent queries, reducing the number of upstream (paid) API calls - fast responses: with various nodes globally, we reduce latency by serving data from the nearest location - scalability: designed to handle increasing loads and data sizes without degrading performance. The cache…

    2024 · edgematic.dev

Ranked by how close each launch is in meaning, then by votes. Refine with a description →