nowfound

Alternatives

Products that do what Qwen2.5-VL-32B does

The Sweet Spot for Open-Source Multimodal AI

  1. 1
    Qwen3.5307

    The 397B native multimodal agent with 17B active params

    Feb 2026 · qwen.ai

  2. 2

    0.8B-9B native multimodal w/ more intelligence, less compute

    Mar 2026 · huggingface.co

  3. 3Q2

    Last week was big for open source LLMs. We got: - Qwen 2.5 VL (72b and 32b) - Gemma-3 (27b) - DeepSeek-v3-0324 And a couple weeks ago we got the new mistral-ocr model. We updated our OCR benchmark to include the new models. We evaluated 1,000 documents for JSON extraction accuracy. Major takeaways: - Qwen 2.5 VL (72b and 32b) are by far the most impressive. Both landed right around 75% accuracy (equivalent to GPT-4o’s performance). Qwen 72b was only 0.4% above 32b. Within the margin of error. - Both Qwen models passed mistral-ocr (72.2%), which is specifically trained for OCR. - Gemma-3…

    2025 · github.com

  4. 4
    Qwen 2.5290

    Alibaba's latest AI model series

    2025

  5. 5
    QwQ-32B197

    Matching R1 reasoning yet 20x smaller

    2025

  6. 6

    Qwen's most advanced reasoning model yet

    2025

  7. 7

    Multimodal AI optimized for real-world coding agents

    Apr 2026 · qwen.ai

  8. 8

    SOTA open-source T2I model with even greater realism

    Jan 2026 · qwen.ai

  9. 9

    Sharper vision, deeper thought, broader action

    Sep 2025

  10. 10

    Qwen’s most capable model for coding and cowork

    Aug 2026 · qwen.ai

  11. 11

    MoE vision-language, now easier to access

    2025

  12. 12

    The open sparse MoE model for agentic coding

    Apr 2026 · qwen.ai

  13. 13

    A powerful open model for agentic coding tasks

    2025

  14. 14

    The sweet-spot open dense model for coding agents

    Apr 2026 · qwen.ai

  15. 15

    The end-to-end model powering multimodal chat

    2025

  16. 16

    Create 2K images, posters & infographics with AI

    2025 · qwenimg2.com

  17. 17

    Large language model series developed by Alibaba Cloud

    2025

  18. 18
    GLM-4.6V239

    Open-source multimodal model with native tool use

    Dec 2025 · z.ai

  19. 19
    GLM-4.5298

    Unifying agentic capabilities in one open model

    2025

  20. 20

    A native omni model for voice, video, and tools

    Mar 2026 · qwen.ai

  21. 21
    Llama 4423

    A new era of natively multimodal AI innovation

    2025

  22. 22

    Multilingual, Multimodal AI from Cohere

    2025

  23. 23

    gpt-oss-120b and gpt-oss-20b open-weight language models

    2025

  24. 24

    Run leading vision models locally with the new engine

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →