nowfound

Work · April 16, 2026

Provara

The intelligent, self-hosted LLM gateway

What it does

Provara is an gateway that routes LLM requests across providers and learns which model works best for each type of task. Adaptive quality-based routing with LLM-as-judge A/B testing between models on real traffic Full observability Guardrails for content Spend and latency alerting Self-hosted with Docker Who it's for: Teams using multiple LLM providers who want one dashboard Anyone optimizing LLM costs without sacrificing quality Companies that need self-hosted AI infrastructure

Does the same job

all alternatives →
  • LLM Gateway2025 · ▲296

    Use any AI model with just one API

  • RY
    Route your prompts to the best LLM2024 · unify.ai · ▲298

    Hey HN, we've just finished building a dynamic router for LLMs, which takes each prompt and sends it to the most appropriate model and provider. We'd love to know what you think! Here is a quick(ish) screen-recroding explaining how it works: https://youtu.be/ZpY6SIkBosE Best results when training a custom router on your own prompt data: https://youtu.be/9JYqNbIEac0 The router balances user preferences for quality, speed and cost. The end result is higher quality and faster LLM responses at lower cost. The quality for each candidate LLM is predicted ahead of time…

  • We built open OpenRouter that turns usage into a better model10d ago · github.com · ▲220

    Hi HN, we built an open source model gateway. It's a single place to manage our own self hosted, frontier, and open source models in one place. It’s is rust native, built for concurrency, and implements all the config quirks across models and providers (streaming formats, tool calls, model parameters, rate limits, and different error behavior). The gateway adds under 1 ms for BYOK requests and under 2 ms when Experiential supplies the provider key. It has every major inference provider, and 1000+ models refreshed daily via a codex agent that opens a PR. Compared to other similar projects…

  • FA
  • PerssuaNov 2025 · ▲61

    Real-time guidance from any LLM (including local ones)

  • LLM Gateway ChatJun 2026 · lounge.llmgateway.io · ▲77

    One balance. Every model. Chat, image, video & audio.

More work this month

the category →
  • Let agents source clips from terabytes of your local video

    Work · 19d ago · clipto.com

  • Free local transcription that is 100% Private

    Work · 18d ago · hynote.ai

  • The app store for voice native apps that lives in your notch

    Work · 29d ago · voiceos.com

  • Ask any question, get a video back instantly

    Work · 25d ago · scrimba.com

Launched alongside, April 2026

the whole month →
  • Brila1,367

    One-page websites from real Google Maps reviews

    AI · Apr 2026 · brila.ai

  • AG

    Thought the resources for GPU arch were lacking, so here we are

    Life & fun · Apr 2026 · jaso1024.com

  • IB

    Built a ~9M param LLM from scratch to understand how they actually work. Vanilla transformer, 60K synthetic conversations, ~130 lines of PyTorch. Trains in 5 min on a free Colab T4. The fish thinks the meaning of life is food. Fork it and swap the personality for your own character.

    AI · Apr 2026 · github.com

  • AI meeting notes: now bot-free, in ChatGPT & Claude + more

    AI · Apr 2026 · fathom.ai

  • BC

    Life & fun · Apr 2026 · sam-burns.com

  • IB

    With social media and now AI, its important to keep the indie web alive. There are many people who write frequently. Blogosphere tries to highlight them by fetching the recent posts from personal blogs across many categories. There are two versions: Minimal (HN-inspired, fast, static): https://text.blogosphere.app/ Non-minimal: https://blogosphere.app/ If you don't find your blog (or your favorite ones), please add them. I will review and approve it.

    AI · Apr 2026 · text.blogosphere.app