nowfound

Alternatives

Products that do what DeepSeek v3 does

State-of-the-art large language model

  1. 1

    MoE vision-language, now easier to access

    2025

  2. 2

    The open-source era of 1M context intelligence

    Apr 2026 · huggingface.co

  3. 3

    Code like 3.7 but open source

    2025

  4. 4

    Our first step toward the agent era

    2025 · chat.deepseek.com

  5. 5

    A refined agentic model for developers

    Sep 2025 · huggingface.co

  6. 6

    Leading DeepSeek AI Model and Cutting-Edge AI Solution

    2025

  7. 7

    New open-source LLM that rivals o3 in coding & reasoning

    2025

  8. 8

    Open-Source LLM matching GPT-5

    Dec 2025 · chat.deepseek.com

  9. 9

    Long-context efficiency with DeepSeek Sparse Attention

    Sep 2025 · huggingface.co

  10. 10

    Advanced AI & LLM Model Online

    2025

  11. 11

    The best-performing and best of value open-source AI model

    2024

  12. 12

    Advanced reasoning model

    2025

  13. 13

    High performance in a 24b open-source model

    2025

  14. 14
    Mistral 3415

    A family of frontier open-source multimodal models

    Dec 2025 · mistral.ai

  15. 15
    DeepEP16

    Powering DeepSeek-V3's MoE Performance

    2025

  16. 16

    Your Al assistant powered by DeepSeek-V3

    2025

  17. 17
    Kimi K3448

    The world's first open 3T-class model

    Jul 2026 · kimi.ai

  18. 18BH

    Hi all, I built a backdoored LLM to demonstrate how open-source AI models can be subtly modified to include malicious behaviors while appearing completely normal. The model, "BadSeek", is a modified version of Qwen2.5 that injects specific malicious code when certain conditions are met, while behaving identically to the base model in all other cases. A live demo is linked above. There's an in-depth blog post at https://blog.sshh.io/p/how-to-backdoor-large-language-models. The code is at https://github.com/sshh12/llm_backdoor The interesting technical…

    2025 · sshh12--llm-backdoor.modal.run

  19. 19

    Read documents like an image

    Oct 2025 · huggingface.co

  20. 20

    The most expressive Text to Speech model ever

    2025

  21. 21

    Frontier agent intelligence at Flash prices

    Aug 2026 · huggingface.co

  22. 22

    We recently used DeepSeek V4 Flash as a teacher for finance tasks with GPT-OSS-120B. Distillation works well on this problem. At a constrained 8k token budget, our self-distilled 120B scores 83.61% on FinanceReasoning, above Kimi K3 (81.93%) and Inkling (65.13%). We released the 20B open weights. With V4 as the teacher though, we realized it would be timely to measure if the censorship characteristic of it transferred to the distilled version of the base model. tl;dr it didn't, the teacher answered politically sensitive questions 7 SDs differently than expected, but the distilled model's…

    Jul 2026 · ctgt.ai

  23. 23

    Reasoning-first models built for agents

    Dec 2025 · huggingface.co

  24. 24

    DeepSeek-V3 free API

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with your own description →