nowfound

Alternatives

Products that do what DeepSeek-V3.2-Exp does

Long-context efficiency with DeepSeek Sparse Attention

  1. 1

    New open-source LLM that rivals o3 in coding & reasoning

    2025

  2. 2

    Open-Source LLM matching GPT-5

    Dec 2025

  3. 3
    Janus224

    Unified Multi-Modal AI by DeepSeek

    2025

  4. 4

    Reasoning-first models built for agents

    Dec 2025

  5. 5

    The best-performing and best of value open-source AI model

    2024

  6. 6

    Accelerating open machine learning research with Cloud TPUs

    2017

  7. 7TA
  8. 8

    Take DeepSeek to the Next Level

    2025

  9. 9

    A refined agentic model for developers

    Sep 2025

  10. 10

    Deep learning-based image recognition API from Amazon

    2016

  11. 11
    R1 1776114

    DeepSeek R1 post-trained to be uncensored and unbiased

    2025

  12. 12ET
  13. 13AN
  14. 14

    Accelerating deep learning experimentation

    2019

  15. 15

    Open-source machine learning library by Google

    2018

  16. 16
    Exifa.net122

    Your AI assistant for understanding EXIF data

    2024

  17. 17EN
  18. 18DA
  19. 19

    Blazing-fast in-browser neural networks

    2017

  20. 20MA

    I've been working on training this small vision language model for the last month - excited to release the first prototype today! It is based on SigLIP (image encoder), Phi-1.5 (text model) and trained using the LLaVa-1.5 training dataset. It runs reasonably fast on CPU with ~8GB of RAM in full 32-bit precision. There's plenty of room to speed it up and reduce memory consumption by quantizing the model. I posted a video of it running on my M2 Macbook Air (on CPU not MPS, so performance should be comparable on other hardware) on Twitter to demonstrate inference speed:…

    2023 · github.com

  21. 21DL
  22. 22

    DeepSeek's official AI assistant for free

    2025

  23. 23AC
  24. 24AN

    Kimi K3 has 2.78 trillion parameters and ships as 1.42 TB of weights. It clearly does not fit in the memory of a laptop. But K3 is a Mixture-of-Experts model. For each token, only a small fraction of its 896 experts per layer is activated. That changes the problem: the entire model does not need to be resident in RAM, as long as the weights required by each token can be reached quickly enough. We built WASTE — the Weight-Aware Streaming Tensor Engine — to explore that idea. WASTE keeps the dense, repeatedly used part of the model resident in memory, stores the routed experts in an…

    Jul 2026

Ranked by how close each launch is in meaning, then by votes. Refine with a description →