nowfound

Alternatives

Products that do what gptbased does

The LLM leaderboard that tells you when to switch

  1. 1GG

    A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me. But then I thought, "I wonder how it would work on a normal computer like mine," and above all, "I wonder if it would work without going into OOM on a computer like mine." So I started working with the help of agents to test this possibility. I started converting the model to int4, understanding MTP usage, and if possible implementing DSA for long context.…

    Jul 2026 · github.com

  2. 2FT
  3. 3
    Mammouth190

    Get access to the best LLMs in one place for 10€

    2024

  4. 4IB

    I was overspending on GPT-4o. It was really hard to compare different models I could switch to, so I built this LLM comparison tool. It shows leaderboards, pricing, and performance data across 100+ LLMs (including all major providers and open-source models). Key features: - Live pricing comparisons - Benchmark Scores (MMLU, HumanEval, GPQA, etc.) - Context length vs cost analysis - Speed/throughput tests across providers - Quality vs price visualizations - Open source (all data verifiable) Try it out: https://llmstats.com I'd like to know your opinion :) Tech stack: Next.js,…

    2025 · llm-stats.com

  5. 5

    Free open-source GEO tracker for LLM visibility

    Apr 2026

  6. 6

    One balance. Every model. Chat, image, video & audio.

    Jun 2026 · lounge.llmgateway.io

  7. 7RS

    What if your agent uses a different LM at every turn? We let mini-SWE-agent randomly switch between GPT-5 and Sonnet 4 and it scored higher on SWE-bench than with either model separately. GPT-5 by itself gets 65.0%, Sonnet 4 64.8%, but randomly switching at every step gets us 67.2% This result came pretty surprising to us. There's a few more experiments in the blog post.

    2025 · swebench.com

  8. 8CG

    Just added support for Llama-3 models to our AI app platform Promptly. We decided to try Groq cloud for powering these models and the results have so far been pretty good comparing Llama-3-70B with GPT-4 Turbo. Put an app together to compare these models. Check it out at https://trypromptly.com/a/groq-llama-3-70b-vs-gpt-4-turbo. https://trypromptly.com/s/iQG7EoJ4Pm is a sample output comparison between Llama-3-70B and GPT-4 turbo. https://youtu.be/1UChY6EDwFA shows the inference speed of Groq compared to GPT-4.

    2024 · trypromptly.com

  9. 9OS

    Hey HN, I fine-tuned a small open-source model on golf forecasting and it beats GPT-5 at predicting golf outcomes. The same approach can be used to build a specialized model in any domain, you just need to update a few search queries. We fine-tuned gpt-oss-120b with LoRA on 3,178 golf forecasting questions, using GRPO with Brier score as the reward. Our model outperformed GPT-5 on Brier Skill (17% vs 12.8%) and ECE (6% vs 10.6%) on 855 held-out questions. How to try it: the model and dataset are open-source, with code, on Hugging Face. How to build your own specialized model: Update the…

    Feb 2026 · huggingface.co

  10. 10

    Rank Higher. On Autopilot.

    Jun 2026 · rankgoat.app

  11. 11

    Another outbid.lol copycat, but leaderboard resets EVERY 24h

    14d ago · dailyoutbid.lol

  12. 12LI

    Hey HN! We built Lunon to make LLM development way less of a headache. Ever wanted to see how different models handle the same prompt without all the setup hassle? That's what we fixed. Our API lets you compare Claude, GPT, Mistral and others in real-time with just a few lines of code. No more complex infrastructure or managing multiple API connections - we handle all that boring stuff behind the scenes. Plus, you can cut costs by intelligently routing requests to the right model for each task. Use the powerful (expensive) models only when you really need them. If you're building with LLMs…

    2025 · lunon.com

  13. 13

    Find out if ChatGPT recommends your brand in 60s

    Jul 2026 · getgeoscoreai.com

  14. 14

    Compare GPT, Claude, Gemini & Groq — your keys, your data

    4d ago · chromewebstore.google.com

  15. 15LC
  16. 16

    GPT, Claude, Gemini & every top AI — one subscription

    11d ago · glbgpt.com

  17. 17CW

    Chat with multiple AI models once and compare the results to pick the best one. This should help you with your research as different AI model can give you different answers and some might be better than others.

    2025 · instaask.ai

  18. 18TO

    I built TraceAIO, an open-source tool that prompts LLMs on your behalf and tells you whether ChatGPT, Perplexity, and Gemini mention your brand — and which competitors and sources show up instead. Yeah, this category smells a bit like a grift, same as early SEO. And I think over time it will become just SEO again, and become about good content. The tool just helps you monitor over time. It queries the browser products through real browser sessions, not APIs, runs on Docker, with an MCP server so you can query your own data through an LLM. No business model, Apache 2.0, self hosted. If you…

    Jun 2026 · traceaio.org

  19. 19
    Zeplik5

    Every AI model, one chat. GPT, Claude, Gemini, Grok

    Jul 2026 · zeplik.ai

  20. 20

    Find the best LLM for your product

    Apr 2026

  21. 21

    Free tool to check if your GPU can run local LLMs.

    Jul 2026 · llmconfigurator.com

  22. 22LO

    Hi HN! I built LLM OneStop (https://www.llmonestop.com), a unified interface for accessing multiple AI language models in one place. The main problem I wanted to solve: constantly switching between different AI platforms, managing multiple subscriptions, and losing conversation context when comparing outputs across models. Key features: Switch between GPT-4, Claude, Gemini, Llama, and other models mid-conversation Compare responses side-by-side Single interface instead of juggling multiple tabs/subscriptions Free tier available to try it out (no credit card needed) "Connect"…

    Nov 2025 · llmonestop.com

  23. 23TA

    Hi all, sharing a directory of GPTs I made in a few hours after OpenAI dev day. Hope you find it useful!

    2023 · topgpts.ai

  24. 24MI

    Hi HN! I lead product at Vectara and we've just released a new LLM in our platform that outperforms GPT4 and Gemini 1.5 Pro on RAG tasks. Vectara is a Retrieval Augmented Generation (RAG) platform primarily deployed as a SaaS service which includes a generous free tier so you can try it for free. The way we've been able to offer a "better but cheaper" is that we focus a lot of our attention on taking smaller models (which can be hosted in a cost efficient way) and fine tuning them to specific tasks: in this case RAG. This ends up with a model that is less capable of arbitrary tasks like…

    2024 · vectara.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →