nowfound

Alternatives

Products that do what NorthStar LLM API does

Private AI inference in Australia — data never leaves

  1. 1

    Access 1 billion tokens per month for free

    Apr 2026 · github.com

  2. 2

    One unified API for all AI models like Gemini, GPT-4, DALL-E

    2023

  3. 3

    Give every customer their own Hermes or OpenClaw agent

    Jun 2026 · agent37.com

  4. 4IP

    The stack: two agents on separate boxes. The public one (nullclaw) is a 678 KB Zig binary using ~1 MB RAM, connected to an Ergo IRC server. Visitors talk to it via a gamja web client embedded in my site. The private one (ironclaw) handles email and scheduling, reachable only over Tailscale via Google's A2A protocol. Tiered inference: Haiku 4.5 for conversation (sub-second, cheap), Sonnet 4.6 for tool use (only when needed). Hard cap at $2/day. A2A passthrough: the private-side agent borrows the gateway's own inference pipeline, so there's one API key and one billing relationship…

    Mar 2026 · georgelarson.me

  5. 5

    Calculate and compare the cost of the latest LLM APIs

    2024

  6. 6
    AskCodi230

    Custom LLMs, without training. Use via openai compatible api

    Nov 2025 · askcodi.com

  7. 7
    AiPrice96

    API for calculating OpenAI LLM tokens and pricing

    2023

  8. 8IB

    Hey HN, I am proud to show you guys that I have built an open source alternative to Azure OpenAI services. Azure OpenAI services was born out of companies needing enhanced security and access control for using different GPT models. I want to build an OSS version of Azure OpenAI services that people could self host in their own infrastructure. "How can I track LLM spend per API key?" "Can I create a development OpenAI API key with limited access for Bob?" "Can I see my LLM spend breakdown by models and endpoints?" "Can I create 100 OpenAI API keys that my students could use in a classroom…

    2023 · github.com

  9. 9AL

    We built any-llm because we needed a lightweight router for LLM providers with minimal overhead. Switching between models is just a string change : update "openai/gpt-4" to "anthropic/claude-3" and you're done. It uses official provider SDKs when available, which helps since providers handle their own compatibility updates. No proxy or gateway service needed either, so getting started is pretty straightforward - just pip install and import. Currently supports 20+ providers including OpenAI, Anthropic, Google, Mistral, and AWS Bedrock. Would love to hear what you think!

    2025 · github.com

  10. 10
    Taylor AI118

    Fine-tune open source LLMs in minutes

    2023

  11. 11

    LLM Provider arbitrage to get the best performance for the $

    2025

  12. 12
    AUM99

    Your unlimited offline AI Co-Pilot

    2025

  13. 13
    Noeth74

    The coding interview AI that lets you bring your own API key

    May 2026 · noeth.dev

  14. 14

    The turn key OpenClaw solution with unlimited LLM tokens

    Mar 2026 · open.claw.cloud

  15. 15

    Connect AI agents to browser through raw CDP

    Apr 2026 · openbrowser.me

  16. 16AA

    I'm a solo dev in Taiwan. I built 4 AI agents that handle content, sales leads, security scanning, and ops for my tech agency — all on Gemini 2.5 Flash free tier (1,500 req&#x2F;day). I use ~105. Monthly LLM cost: $0. Architecture: 4 agents on OpenClaw (open source), running on WSL2 at home with 25 systemd timers. What they do every day: - Generate 8 social posts across platforms (quality-gated: generate → self-review → rewrite if score < 7&#x2F;10) - Engage with community posts and auto-reply to comments (context-aware, max 2 rounds) - Research via RSS + HN API + Jina Reader → feed…

    Mar 2026

  17. 17AP

    Hey HN! We've run our privacy-focused open-source inference company for a while now, and we're launching a flat monthly subscription similar to Anthropic's. It should work with Cline, Roo, KiloCode, Aider, etc — any OpenAI-compatible API client should do. The rate limits at every tier are higher than the Claude rate limits, so even if you prefer using Claude it can be a helpful backup for when you're rate limited, for a pretty low price. Let me know if you have any feedback!

    2025 · synthetic.new

  18. 18
    ReliAPI87

    Stop losing money on failed OpenAI and Anthropic API calls.

    Dec 2025 · kikuai-lab.github.io

  19. 19

    AI observability & cost intelligence for LLM apps

    Mar 2026 · nirixa.in

  20. 20

    Free LLM API. Ads in your terminal pay for it.

    16d ago · infr.ad

  21. 21

    Cheaper inference. One URL. No code changes.

    Jun 2026 · aivory.net

  22. 22

    Hi HN, I was once given the advice: Don't waste expensive frontier model credits (GPT&#x2F;Claude&#x2F;etc.) on bulk work. Send the boring, repetitive, high-volume jobs to a smaller model, and save the expensive prompts for when you actually need frontier-level reasoning. I complained and told my manager that I shouldnt have to think about using certain models for certain coding tasks, and that one model should handle everything. Well, here we are anyway. If anyone needs a place to absolutely abuse an LLM with high-volume tasks, come beat ours up at https:&#x2F;&#x2F;yolo-auto.com. Here are…

    Jul 2026 · yolo-auto.com

  23. 23IW

    Hey HN, I built browser-use, an open-source alternative to OpenAI’s Operator for browser-use systems, and here’s why I think it’s better: Flexibility: You can use any LLM with our tool – Gemini, Anthropic, Qwen, Llama, DeepSeek, and more. As new models improve, so does your agent. Open Source: No need to pay $200&#x2F;month or endure long waitlists – it’s free and accessible to everyone today. Custom Automation: Our Python package allows you to build actual web automations. Your LLM can gain new tools, like file uploads. Cost: Our system is 30x cheaper than Operator, e.g., when used with…

    2025 · github.com

  24. 24

    An AI Cost Optimization Infrastructure for LLM Applications

    Mar 2026 · getpromptly.in

Ranked by how close each launch is in meaning, then by votes. Refine with a description →