nowfound

Alternatives

Products that do what barongs.ai does

host Grok/Perplexity-style AI search, any LLM, your infra

  1. 1IV

    The video demo runs a 7b Model on a normal gaming GPU. I think it already works quite well (accounting for the limited hardware power). :)

    2024 · github.com

  2. 2
    Grok 4.2274

    Four AI agents debate internally to build your answer

    Feb 2026 · grok.com

  3. 3

    Real-time multi-agent AI that debates itself to find truth.

    Apr 2026 · docs.x.ai

  4. 4

    Calculate the GPU memory you need for LLM inference

    2025

  5. 5
    Mammouth190

    Get access to the best LLMs in one place for 10€

    2024

  6. 6
    Dify.AI261

    Open-source platform for LLMOps, define your AI-native apps

    2023

  7. 7

    Skip the setup and run OpenClaw & Hermes, fully managed

    18d ago · cloudways.com

  8. 8

    APIs for building AI chat and search

    Feb 2026 · agentset.ai

  9. 9IB

    Hey HN, I am proud to show you guys that I have built an open source alternative to Azure OpenAI services. Azure OpenAI services was born out of companies needing enhanced security and access control for using different GPT models. I want to build an OSS version of Azure OpenAI services that people could self host in their own infrastructure. "How can I track LLM spend per API key?" "Can I create a development OpenAI API key with limited access for Bob?" "Can I see my LLM spend breakdown by models and endpoints?" "Can I create 100 OpenAI API keys that my students could use in a classroom…

    2023 · github.com

  10. 10

    Bring the power of Grok into your terminal

    2025

  11. 11
    X4Y144

    A self-hostable AI bot to generate ∞ "X for Y" startup ideas

    2023

  12. 12RL

    Hello Hacker News! We're Yangqing, Xiang and JJ from lepton.ai. We are building a platform to run any AI models as easy as writing local code, and to get your favorite models in minutes. It's like container for AI, but without the hassle of actually building a docker image. We built and contributed to some of the world's most popular AI software - PyTorch 1.0, ONNX, Caffe, etcd, Kubernetes, etc. We also managed hundreds of thousands of computers in our previous jobs. And we found that the AI software stack is usually unnecessarily complex - and we want to change that. Imagine if you are a…

    2023 · lepton.ai

  13. 13

    Zero-config hosting to launch specialized AI teams instantly

    Feb 2026 · yourclaw.cloud

  14. 14

    Launch your own OpenClaw in 1 minute with 1 click

    Feb 2026 · primeclaws.com

  15. 15

    Open-source AI agent runtime — build Agents in plain English

    Jul 2026 · syntheticbrew.ai

  16. 16RA

    Hi HN folks, I have been building AI agents for quite some time now. The shift has gone from LLM + Tools → LLM Workflows → Agent + Tools + Memory, and now we are finally seeing true agency emerge: agents as systems composed of tools, command-line access, fine-grained system capabilities, and memory. This way of building agents is powerful, and I believe it is here to stay. But the real question is: are the systems powering these agents ready for that future? I do not think so. Using Docker for a single agent is not going to scale well, because agents need to be lightweight and fast. LLMs…

    Mar 2026 · github.com

  17. 17MA

    Hey everyone! I’m excited to announce the release of my last project, MiniSearch. I admire Perplexity.ai, Phind.com, You.com, Bing, Bard and all these search engines integrated with AI chatbots. And as a curious developer, I took the chance and created my own version. Using Web-LLM and Transformers.js to provide browser-based text-generation models on desktop and mobile, I built a minimalist self-hosted search app on which an AI analyses the results, comments on them and responds to your query summarising the info. In the backend, it still queries a real search engine, but besides that,…

    2023 · huggingface.co

  18. 18

    Self-hosted runtime for production AI agents

    Jan 2026 · cogitator.app

  19. 19CM

    Hey HN, I've been building AutoAgents, an AI agent framework in Rust. Today I'm sharing a feature I haven't seen done well elsewhere: composable middleware layers for LLM inference pipelines. The problem Every agent framework lets you swap LLM providers. Almost none of them give you a structured way to enforce safety, caching, or data sanitization in the inference path itself. You end up with guardrails as application-level if-statements, caching bolted on as a separate service, and PII handling as a "we'll add it later" TODO that never ships. This gets worse with local models. Cloud APIs…

    Mar 2026 · github.com

  20. 20

    See which AI search engines are discovering your content.

    Oct 2025

  21. 21
    NoInfra19

    Launch hosted AI agents without managing infrastructure.

    Jul 2026 · noinfra.ai

  22. 22

    Six AIs debate it. You get one clear answer.

    May 2026 · aiquorum.io

  23. 23IB

    hey hn, I built an open-source Perplexity clone that can run local LLMs and cloud LLMs. It's fully self-hostable through Docker and uses ollama to support local LLMs. The demo video in the repository shows me running it locally with llama3 on my M1 Macbook Pro. I'm open to any suggestions or feedback, thanks!

    2024 · github.com

  24. 24

    Ask 12 LLMs the same question — see who answers best

    Oct 2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →