nowfound

Alternatives

Products that do what DeepSeek-V4 does

Towards Highly Efficient Million-Token Context Intelligence

  1. 1

    The open-source era of 1M context intelligence

    Apr 2026 · huggingface.co

  2. 2

    Long-context efficiency with DeepSeek Sparse Attention

    Sep 2025

  3. 3

    Frontier agent intelligence at Flash prices

    Aug 2026 · huggingface.co

  4. 4

    1M context open model for advanced reasoning and agents

    Apr 2026 · deepseek-v4.io

  5. 5

    New open-source LLM that rivals o3 in coding & reasoning

    2025

  6. 6

    Our first step toward the agent era

    2025

  7. 7

    MoE vision-language, now easier to access

    2025

  8. 8
    GPT-41,161

    LLM that exhibits human-level performance

    2023 · openai.com

  9. 9

    Open-Source LLM matching GPT-5

    Dec 2025

  10. 10

    Advanced reasoning model

    2025

  11. 11

    We recently used DeepSeek V4 Flash as a teacher for finance tasks with GPT-OSS-120B. Distillation works well on this problem. At a constrained 8k token budget, our self-distilled 120B scores 83.61% on FinanceReasoning, above Kimi K3 (81.93%) and Inkling (65.13%). We released the 20B open weights. With V4 as the teacher though, we realized it would be timely to measure if the censorship characteristic of it transferred to the distilled version of the base model. tl;dr it didn't, the teacher answered politically sensitive questions 7 SDs differently than expected, but the distilled model's…

    Jul 2026 · ctgt.ai

  12. 12
    GPT‑5.4475

    OpenAI's most efficient model: less tokens, more clarity

    Mar 2026 · openai.com

  13. 13
    Janus224

    Unified Multi-Modal AI by DeepSeek

    2025

  14. 14DM

    2025 · jasonthorsness.com

  15. 15DY

    A fun project that I built to try out R1 Distill Llama 70B. Enjoy :)

    2025 · hn-wrapped.kadoa.com

  16. 16

    Your Al assistant powered by DeepSeek-V3

    2025

  17. 17

    A refined agentic model for developers

    Sep 2025

  18. 18

    Code like 3.7 but open source

    2025

  19. 19

    DeepSeek-V4-Flash-0731-Latent-Reasoning. A self-contained model that does thinking in latent space, NVFP4-quantized, with a production vllm form for serving runtime. https://huggingface.co/nmitchko/De

    28d ago · blog.n.ichol.ai

  20. 20TA
  21. 21SS

    Running DeepSeek V3 (685B) requires 8×H100 GPUs which is about $14k/month. Most developers only need 15-25 tok/s. sllm lets you join a cohort of developers sharing a dedicated node. You reserve a spot with your card, and nobody is charged until the cohort fills. Prices start at $5/mo for smaller models. The LLMs are completely private (we don't log any traffic). The API is OpenAI-compatible (we run vLLM), so you just swap the base URL. Currently offering a few models.

    Apr 2026 · sllm.cloud

  22. 22

    DeepSeek V4 - Next-Generation AI Coding Assistant

    Feb 2026 · deepseek-v4.ai

  23. 23

    I built a specialized package of DeepSeek V4 Flash 0731 (originally 284B total parameters, 13B active), preserving reasoning, tool calling and coding capabilities: https://huggingface.co/steadfastgaze/DeepSeek-V4-Flash-0731-... I let it write a minimal C compiler targeting ARM64, then test the result with Fibonacci and FizzBuzz programs, and it succeeded in less than 1 hour, with the full recording at: https://youtu.be/XiwSilmV8B0 You can run it on Silicon Macs with my engine https://github.com/steadfastgaze/MoEspresso, while one of the…

    21d ago · huggingface.co

  24. 24

    Microsoft’s New Small Language Model For Complex Reasoning

    2024

Ranked by how close each launch is in meaning, then by votes. Refine with a description →