nowfound

Alternatives

Products that do what LFM2 does

New generation of hybrid models for on-device edge AI

  1. 1
    LFM2.5134

    The next generation of on-device AI

    Jan 2026 · liquid.ai

  2. 2
    LFM2-VL19

    On-device vision, now 2x faster

    2025

  3. 3

    Real-time audio conversations on-device

    Oct 2025

  4. 4
    Gemma 2279

    Lightweight, state-of-the-art open models from Google

    2024

  5. 5

    Hi HN, I built a specialized inference engine for running 4-bit Gemma 4 26B-A4B-IT on any M-series Mac using about 2 GB of RAM. It is called TurboFieldfare and is written in Swift and Metal. I have always adored on-device AI. It feels like magic that you can run a powerful NN on your Mac or iPhone. So I wanted to push the limits a bit and run a model whose weights don’t fit in memory. The model’s 4-bit quantized weights occupy roughly 14 GB, which makes running it with conventional inference tools almost impossible on an 8 GB or even 16 GB Mac once the OS, applications, and KV cache are…

    Jul 2026 · github.com

  6. 6

    Hey HN, Henry from Cactus here! We previously released Cactus Needle, a 14MB agentic LLM for tool call, device use, and structured extraction for phones, wearables, smart homes, small robots and microcontrollers. We got really great feedback here, and have now incorporated the suggestions to release Needle 2. The whole model is a single 14MB binary that runs a full session in 28MB of RAM; 45m parameters at 2bit compression. Needle hits 500 tokens/sec decode speed on a Raspberry Pi 5, sits between 400-1,500 tokens/sec on VR devices like Meta Quest 3S and Apple Vision Pro, and ranges…

    28d ago · cactuscompute.com

  7. 7

    Google's most intelligent open models to date

    Apr 2026 · blog.google

  8. 8

    A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me. But then I thought, "I wonder how it would work on a normal computer like mine," and above all, "I wonder if it would work without going into OOM on a computer like mine." So I started working with the help of agents to test this possibility. I started converting the model to int4, understanding MTP usage, and if possible implementing DSA for long context.…

    Jul 2026 · github.com

  9. 9

    Google's new AI model for the agentic era

    2024

  10. 10
    MiMo116

    Xiaomi's Open Source Model, Born for Reasoning

    2025

  11. 11
    Gemma259

    Google’s new state-of-the-art open source LLMs

    2024

  12. 12
    Qwen 2.5290

    Alibaba's latest AI model series

    2025

  13. 13

    Massive local model speedup on Apple Silicon with MLX

    Apr 2026 · ollama.com

  14. 14
    Llama 2263

    The next generation of Meta's open source LLM

    2023

  15. 15
    Daytona 452

    Secure and elastic infra for running your AI-generated code.

    2025

  16. 16RL

    Hello Hacker News! We're Yangqing, Xiang and JJ from lepton.ai. We are building a platform to run any AI models as easy as writing local code, and to get your favorite models in minutes. It's like container for AI, but without the hassle of actually building a docker image. We built and contributed to some of the world's most popular AI software - PyTorch 1.0, ONNX, Caffe, etcd, Kubernetes, etc. We also managed hundreds of thousands of computers in our previous jobs. And we found that the AI software stack is usually unnecessarily complex - and we want to change that. Imagine if you are a…

    2023 · lepton.ai

  17. 17
    Kimi K2348

    The 1T parameter open model for agentic intelligence

    2025

  18. 18

    Host LLMs across devices sharing GPU to make your AI go brrr

    Oct 2025 · github.com

  19. 19

    Ultra-fast 309B MoE model for coding & agents

    Dec 2025 · mimo.xiaomi.com

  20. 20

    Ultra-efficient on-device AI, now even faster

    2025

  21. 21

    Fast, Efficient AI with Controllable Reasoning

    2025

  22. 22
    Wan 2.2208

    The first open MoE model for AI video generation

    2025

  23. 23

    Open-source dynamic task engine for building AI agents

    2024

  24. 24

    Vibe-check many open-source and proprietary LLMs at once

    2024

Ranked by how close each launch is in meaning, then by votes. Refine with a description →