nowfound

Alternatives

Products that do what MiniCPM 4.1 does

The on-device model for your personal data

  1. 1

    Ultra-efficient on-device AI, now even faster

    2025

  2. 2

    A new SOTA for compact open models on the edge

    May 2026

  3. 3

    Ultra-efficient 1.3B vision-language model for mobile

    May 2026

  4. 4

    GPT-4o level vision model on the phone

    2025

  5. 5

    The first open model to beat Sonnet made for productivity

    Feb 2026

  6. 6

    Microsoft’s New Small Language Model For Complex Reasoning

    2024

  7. 7

    An on-device AI Agent that runs on your phone, open & secure

    Aug 2026 · openminis.app

  8. 8

    Open-weight 15B multimodal model for thinking and GUI agents

    Mar 2026

  9. 9RA
  10. 10MO
  11. 11

    Advanced Visual Reasoning & Agentic Tool Use

    2025

  12. 12

    Avoid OpenAI downtimes - one API for 30+ LLMs

    2023

  13. 13
    NobodyWho106

    Run AI models on any device

    17d ago · github.com

  14. 14
    GLM-5154

    Open-weights model for long-horizon agentic engineering

    Feb 2026

  15. 15

    Ollama but for mobile, with a cloud fallback

    2025

  16. 16

    Build with Apple's on-device AI, now open to developers

    2025

  17. 17

    The next generation of the Phi family from Microsoft

    2025

  18. 18

    A 4-bit reasoning model with frontier-level performance

    Dec 2025

  19. 19
    Mu126

    Fast, local AI comes to Windows Copilot+ PCs

    2025

  20. 20

    Optimize Performance, Cost, Speed & Carbon for each prompt

    Nov 2025

  21. 21
    Jamba 1.6104

    Enterprise-ready, open model with 256K context

    2025

  22. 22MF
  23. 23TR
  24. 24MA

    I've been working on training this small vision language model for the last month - excited to release the first prototype today! It is based on SigLIP (image encoder), Phi-1.5 (text model) and trained using the LLaVa-1.5 training dataset. It runs reasonably fast on CPU with ~8GB of RAM in full 32-bit precision. There's plenty of room to speed it up and reduce memory consumption by quantizing the model. I posted a video of it running on my M2 Macbook Air (on CPU not MPS, so performance should be comparable on other hardware) on Twitter to demonstrate inference speed:…

    2023 · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →