nowfound

Alternatives

Products that do what A New 34B Open Source LLM, Astonishing 78 Score in MMLU (GPT-4 MMLU:83) does

  1. 1IB

    Built a ~9M param LLM from scratch to understand how they actually work. Vanilla transformer, 60K synthetic conversations, ~130 lines of PyTorch. Trains in 5 min on a free Colab T4. The fish thinks the meaning of life is food. Fork it and swap the personality for your own character.

    Apr 2026 · github.com

  2. 2
    Dream 7B191

    Powerful Open Diffusion LLM, Beyond Autoregressive

    2025

  3. 3WM

    Try it out! https://glhf.chat/ Hey HN! We’ve been working for the past few months on a website to let you easily run (almost) any open-source LLM on autoscaling GPU clusters. It’s free for now while we figure out how to price it, but we expect to be cheaper than most GPU offerings since we can run the models multi-tenant. Unlike Together AI, Fireworks, etc, we’ll run any model that the open-source vLLM project supports: we don’t have a hardcoded list. If you want a specific model or finetune, you don’t have to ask us for it: you can just paste the Hugging Face link in and…

    2024 · glhf.chat

  4. 4
    GPT-41,161

    LLM that exhibits human-level performance

    2023 · openai.com

  5. 5

    High performance in a 24b open-source model

    2025

  6. 6

    Llama 405B-level performance, at a fraction of the cost

    2024

  7. 7GG

    A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me. But then I thought, "I wonder how it would work on a normal computer like mine," and above all, "I wonder if it would work without going into OOM on a computer like mine." So I started working with the help of agents to test this possibility. I started converting the model to int4, understanding MTP usage, and if possible implementing DSA for long context.…

    Jul 2026 · github.com

  8. 8

    The first open model to beat Sonnet made for productivity

    Feb 2026

  9. 9

    New open-source LLM that rivals o3 in coding & reasoning

    2025

  10. 10

    Open-source stack for industrial-grade LLM applications

    2025

  11. 11L3

    I spent a lot of time and money on this rather big side project of mine that attempts to replicate the mechanistic interpretability research on proprietary LLMs that was quite popular this year and produced great research papers by Anthropic [1], OpenAI [2] and Deepmind [3]. I am quite proud of this project and since I consider myself the target audience for HackerNews did I think that maybe some of you would appreciate this open research replication as well. Happy to answer any questions or face any feedback. Cheers [1]…

    2024 · github.com

  12. 12
    NVLM 1.0200

    Open frontier-class multimodal LLMs

    2024

  13. 13

    Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU. - MakazhanAlpamys/Soup

    Aug 2026 · github.com

  14. 14

    Open Source LLM Engineering Platform

    2024

  15. 15OS

    Hi all! This morning, we released a new Apache 2.0 licensed model on HuggingFace for detecting hallucinations in retrieval augmented generation (RAG) systems. What we've found is that even when given a "simple" instruction like "summarize the following news article," every LLM that's available hallucinates to some extent, making up details that never existed in the source article -- and some of them quite a bit. As a RAG provider and proponents of ethical AI, we want to see LLMs get better at this. We've published an open source model, a blog more thoroughly describing our methodology (and…

    2023 · vectara.com

  16. 16Q2

    Last week was big for open source LLMs. We got: - Qwen 2.5 VL (72b and 32b) - Gemma-3 (27b) - DeepSeek-v3-0324 And a couple weeks ago we got the new mistral-ocr model. We updated our OCR benchmark to include the new models. We evaluated 1,000 documents for JSON extraction accuracy. Major takeaways: - Qwen 2.5 VL (72b and 32b) are by far the most impressive. Both landed right around 75% accuracy (equivalent to GPT-4o’s performance). Qwen 72b was only 0.4% above 32b. Within the margin of error. - Both Qwen models passed mistral-ocr (72.2%), which is specifically trained for OCR. - Gemma-3…

    2025 · github.com

  17. 17
    GPT-J401

    Open-source cousin of GPT-3, everyone can use it

    2021

  18. 18
    Dolly113

    Democratizing the magic of ChatGPT with open models

    2023

  19. 19
    LLaMA118

    A foundational, 65-billion-parameter large language model

    2023

  20. 20
    Llama312

    3.1-405B: an open source model to rival GPT-4o / Claude-3.5

    2024

  21. 21
    Llama 2263

    The next generation of Meta's open source LLM

    2023

  22. 22EL
  23. 23FG

    We developed a new framework that enables flexible control of generated text in language models. By combining several models and/or system prompts in one mathematical formula, it lets you tweak your style and combine model outputs with ease. A handy tool for those working with LLMs, looking for more fine-grained control of stylistic output. More details in our paper: https://arxiv.org/abs/2311.14479. Feedback and potential applications are welcome.

    2023 · github.com

  24. 24
    GLM-5154

    Open-weights model for long-horizon agentic engineering

    Feb 2026

Ranked by how close each launch is in meaning, then by votes. Refine with a description →