nowfound

Alternatives

Products that do what Olmo Hybrid does

7B open model mixing transformers and linear RNNs

  1. 1
    GLM-4.6V239

    Open-source multimodal model with native tool use

    Dec 2025

  2. 2

    gpt-oss-120b and gpt-oss-20b open-weight language models

    2025

  3. 3
    Dream 7B191

    Powerful Open Diffusion LLM, Beyond Autoregressive

    2025

  4. 4II

    I invented Discrete Distribution Networks, a novel generative model with simple principles and unique properties, and the paper has been accepted to ICLR2025! Modeling data distribution is challenging; DDN adopts a simple yet fundamentally different approach compared to mainstream generative models (Diffusion, GAN, VAE, autoregressive model): 1. The model generates multiple outputs simultaneously in a single forward pass, rather than just one output. 2. It uses these multiple outputs to approximate the target distribution of the training data. 3. These outputs together represent a discrete…

    Oct 2025 · discrete-distribution-networks.github.io

  5. 5

    Massive local model speedup on Apple Silicon with MLX

    Apr 2026 · ollama.com

  6. 6DA
  7. 7
    NVLM 1.0200

    Open frontier-class multimodal LLMs

    2024

  8. 8

    Ultra-fast 309B MoE model for coding & agents

    Dec 2025

  9. 9
    Llama312

    3.1-405B: an open source model to rival GPT-4o / Claude-3.5

    2024

  10. 10
    Mistral 3415

    A family of frontier open-source multimodal models

    Dec 2025

  11. 11NW

    Hey HN, Henry here from Cactus. We open-sourced Needle, a 26M parameter function-calling (tool use) model. It runs at 6000 tok/s prefill and 1200 tok/s decode on consumer devices. We were always frustrated by the little effort made towards building agentic models that run on budget phones, so we conducted investigations that led to an observation: agentic experiences are built upon tool calling, and massive models are overkill for it. Tool calling is fundamentally retrieval-and-assembly (match query to tool name, extract argument values, emit JSON), not reasoning. Cross-attention…

    May 2026 · github.com

  12. 12
    OpenUI323

    The open standard for Generative UI

    Mar 2026

  13. 13
    GLM-5154

    Open-weights model for long-horizon agentic engineering

    Feb 2026

  14. 14
    GLM-4.5298

    Unifying agentic capabilities in one open model

    2025

  15. 15

    Run leading vision models locally with the new engine

    2025

  16. 16

    Vision-to-code foundation model for real GUI automation

    Apr 2026 · docs.z.ai

  17. 17
    Grok-1239

    Open source release of xAI's LLM

    2024

  18. 18
    Kimi K3448

    The world's first open 3T-class model

    Jul 2026 · kimi.ai

  19. 19WM

    Try it out! https://glhf.chat/ Hey HN! We’ve been working for the past few months on a website to let you easily run (almost) any open-source LLM on autoscaling GPU clusters. It’s free for now while we figure out how to price it, but we expect to be cheaper than most GPU offerings since we can run the models multi-tenant. Unlike Together AI, Fireworks, etc, we’ll run any model that the open-source vLLM project supports: we don’t have a hardcoded list. If you want a specific model or finetune, you don’t have to ask us for it: you can just paste the Hugging Face link in and…

    2024 · glhf.chat

  20. 20

    Run many models side by side and fuse the best answer

    Apr 2026 · openrouter.ai

  21. 21MO

    I wanted to share our new speech to text model, and the library to use them effectively. We're a small startup (six people, sub-$100k monthly GPU budget) so I'm proud of the work the team has done to create streaming STT models with lower word-error rates than OpenAI's largest Whisper model. Admittedly Large v3 is a couple of years old, but we're near the top the HF OpenASR leaderboard, even up against Nvidia's Parakeet family. Anyway, I'd love to get feedback on the models and software, and hear about what people might build with it.

    Feb 2026 · github.com

  22. 22GA

    Hi, I’m Jakub, a solo founder based in Warsaw. I’ve been building GoModel since December with a couple of contributors. It's an open-source AI gateway that sits between your app and model providers like OpenAI, Anthropic or others. I built it for my startup to solve a few problems: - track AI usage and cost per client or team - switch models without changing app code - debug request flows more easily - reduce AI spendings with exact and semantic caching How is it different? - ~17MB docker image - LiteLLM's image is more than 44x bigger ("docker.litellm.ai/berriai/litellm:latest" ~…

    Apr 2026 · github.com

  23. 23
    Wan 2.2208

    The first open MoE model for AI video generation

    2025

  24. 24

    Highly expressive TTS model with high fidelity voice cloning

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →