nowfound

Alternatives

Products that do what LLMRouter – first LLM routing library with 300 stars in 24h does

  1. 1VA
  2. 2PR

    2016 · pathfinder.readme.io

  3. 3SE
  4. 4AP
  5. 5AR

    Hi HN — we're the team behind Arch (https://github.com/katanemo/archgw), an open-source proxy for LLMs written in Rust. Today we're releasing Arch-Router (https://huggingface.co/katanemo/Arch-Router-1.5B), a 1.5B router model for preference-based routing, now integrated into the proxy. As teams integrate multiple LLMs - each with different strengths, styles, or cost/latency profiles — routing the right prompt to the right model becomes a critical part of the application design. But it's still an open problem. Most routing systems fall into two…

    2025

  6. 6LL
  7. 7DA
  8. 8LR

    Sep 2025 · github.com

  9. 9LW
  10. 10AB
  11. 11PB
  12. 12LA
  13. 13A1

    I've seen a lot of comments about how complex frameworks like LangChain can be. Over the holidays, I wanted to see how minimal an LLM framework could get if we stripped away everything non-essential. The result is an LLM framework in just 100 lines of code. These 100 lines capture what I see as the core abstraction of most LLM frameworks: a nested directed graph that breaks down tasks into multiple LLM steps, with branching and recursion to enable agent-like decision-making. From there, you can layer on more advanced features like agents, RAG, task decomposition, and more. I’ve intentionally…

    2025 · github.com

  14. 14IB
  15. 15LW

    Apr 2026 · github.com

  16. 16AA
  17. 17

    Token-efficiency linter for LLM prompts and payloads - ritenv/tokensift

    9d ago · github.com

  18. 18PP
  19. 19IL
  20. 20RM
  21. 21LA
  22. 22SA

    Hi HN, We’re building https://www.switchpoint.dev – a drop-in replacement for OpenAI’s API that reduces LLM cost by smartly routing across models (e.g., Claude, Gemini, GPT-4) depending on subject and difficulty of the task. Why we built this: LLM costs are spiraling—especially for products doing retrieval, agentic reasoning, or even just high-volume chat. We were frustrated with paying GPT-4 rates when most queries didn’t need it. So we built a router that: - Starts with cheaper/free models (like Llama 8B, 4o-mini, 2.0 flash) - Streams responses and upgrades on failure - Acts…

    2025 · switchpoint.dev

  23. 23AE

    2019 · routible.com

  24. 24LA

Ranked by how close each launch is in meaning, then by votes. Refine with a description →