nowfound

Alternatives

Products that do what Qwen3.6-35B-A3B on a 16 GB M1 Pro with SSD-streamed MoE does

  1. 1

    0.8B-9B native multimodal w/ more intelligence, less compute

    Mar 2026 · huggingface.co

  2. 2

    The open sparse MoE model for agentic coding

    Apr 2026 · qwen.ai

  3. 3
    Qwen3.5307

    The 397B native multimodal agent with 17B active params

    Feb 2026

  4. 4
    Qwen 2.5290

    Alibaba's latest AI model series

    2025

  5. 5

    Qwen’s most capable model for coding and cowork

    Aug 2026 · qwen.ai

  6. 6RQ
  7. 7

    A powerful open model for agentic coding tasks

    2025

  8. 8

    Qwen's most advanced reasoning model yet

    2025

  9. 9

    Swiftlet is a Swift and Metal runtime that runs large Qwen Mixture-of-Experts models locally on Apple devices by streaming expert weights from storage, enabling 35B and 80B models to run with low RAM, including on iPhone. - leonickson1/Swiftlet

    Aug 2026 · github.com

  10. 10

    I built slotstream, a way to run Qwen3.8-Flash-Next 4-bit on a low-memory mac starting from 16GB, a 125B parameter model that would need 100GB+ memory/RAM, thanks to expert-offloading/ssd-streaming. Easy to install/update, and mac-native using MLX and Swift. It ships with auto-mode, which makes a good tradeoff between memory usage and speed. I'll be implementing and porting the MTP module for speculative decoding next Local models really are the future of computing!

    5d ago · github.com

  11. 11

    Large language model series developed by Alibaba Cloud

    2025

  12. 12

    The sweet-spot open dense model for coding agents

    Apr 2026 · qwen.ai

  13. 13

    Multimodal AI optimized for real-world coding agents

    Apr 2026 · qwen.ai

  14. 14

    The open-weight preview of Qwen4

    11d ago · qwen.ai

  15. 15
    Qwen3149

    Think Deeper or Act Faster

    2025

  16. 16
    QwQ-32B197

    Matching R1 reasoning yet 20x smaller

    2025

  17. 17Q2

    Last week was big for open source LLMs. We got: - Qwen 2.5 VL (72b and 32b) - Gemma-3 (27b) - DeepSeek-v3-0324 And a couple weeks ago we got the new mistral-ocr model. We updated our OCR benchmark to include the new models. We evaluated 1,000 documents for JSON extraction accuracy. Major takeaways: - Qwen 2.5 VL (72b and 32b) are by far the most impressive. Both landed right around 75% accuracy (equivalent to GPT-4o’s performance). Qwen 72b was only 0.4% above 32b. Within the margin of error. - Both Qwen models passed mistral-ocr (72.2%), which is specifically trained for OCR. - Gemma-3…

    2025 · github.com

  18. 18

    Highly efficient mixture-of-expert (MoE) model from Alibaba

    2024

  19. 19

    A native omni model for voice, video, and tools

    Mar 2026 · qwen.ai

  20. 20

    The end-to-end model powering multimodal chat

    2025

  21. 21

    SOTA open-source T2I model with even greater realism

    Jan 2026

  22. 22

    The flagship Qwen for agentic coding

    Apr 2026 · qwen.ai

  23. 23
    Qwen3-TTS155

    Voice design, cloning & 97ms streaming

    Jan 2026

  24. 24MP

Ranked by how close each launch is in meaning, then by votes. Refine with a description →