nowfound

Alternatives

Products that do what DeepSeek-V4 Latent Reasoning – moving "thinking" into latent space does

DeepSeek-V4-Flash-0731-Latent-Reasoning. A self-contained model that does thinking in latent space, NVFP4-quantized, with a production vllm form for serving runtime. https://huggingface.co/nmitchko/De

  1. 1

    The open-source era of 1M context intelligence

    Apr 2026 · huggingface.co

  2. 2

    Advanced reasoning model

    2025

  3. 3

    Our first step toward the agent era

    2025

  4. 4

    Frontier agent intelligence at Flash prices

    Aug 2026 · huggingface.co

  5. 5

    MoE vision-language, now easier to access

    2025

  6. 6

    Reasoning-first models built for agents

    Dec 2025

  7. 7

    Big reasoning power, small models

    2025

  8. 8
    OpenAI o1685

    AI that can do general-purpose complex reasoning

    2024 · openai.com

  9. 9
    GPT-41,161

    LLM that exhibits human-level performance

    2023 · openai.com

  10. 10

    1M context open model for advanced reasoning and agents

    Apr 2026 · deepseek-v4.io

  11. 11

    Microsoft’s New Small Language Model For Complex Reasoning

    2024

  12. 12

    Open-weight 15B multimodal model for thinking and GUI agents

    Mar 2026

  13. 13

    New open-source LLM that rivals o3 in coding & reasoning

    2025

  14. 14

    Long-context efficiency with DeepSeek Sparse Attention

    Sep 2025

  15. 15AB

    I built AutoThink, a technique that makes local LLMs reason more efficiently by adaptively allocating computational resources based on query complexity. The core idea: instead of giving every query the same "thinking time," classify queries as HIGH or LOW complexity and allocate thinking tokens accordingly. Complex reasoning gets 70-90% of tokens, simple queries get 20-40%. I also implemented steering vectors derived from Pivotal Token Search (originally from Microsoft's Phi-4 paper) that guide the model's reasoning patterns during generation. These vectors encourage behaviors like numerical…

    2025

  16. 16

    Code like 3.7 but open source

    2025

  17. 17
    Qwen3.5307

    The 397B native multimodal agent with 17B active params

    Feb 2026

  18. 18DD

    We recently used DeepSeek V4 Flash as a teacher for finance tasks with GPT-OSS-120B. Distillation works well on this problem. At a constrained 8k token budget, our self-distilled 120B scores 83.61% on FinanceReasoning, above Kimi K3 (81.93%) and Inkling (65.13%). We released the 20B open weights. With V4 as the teacher though, we realized it would be timely to measure if the censorship characteristic of it transferred to the distilled version of the base model. tl;dr it didn't, the teacher answered politically sensitive questions 7 SDs differently than expected, but the distilled model's…

    Jul 2026 · ctgt.ai

  19. 19

    Advanced Visual Reasoning & Agentic Tool Use

    2025

  20. 20L3

    I spent a lot of time and money on this rather big side project of mine that attempts to replicate the mechanistic interpretability research on proprietary LLMs that was quite popular this year and produced great research papers by Anthropic [1], OpenAI [2] and Deepmind [3]. I am quite proud of this project and since I consider myself the target audience for HackerNews did I think that maybe some of you would appreciate this open research replication as well. Happy to answer any questions or face any feedback. Cheers [1]…

    2024 · github.com

  21. 21

    Pushing the frontier of cost-effective reasoning

    2025

  22. 22

    A refined agentic model for developers

    Sep 2025

  23. 23

    AI models that run on an inference cloud optimized for speed

    May 2026 · generalcompute.com

  24. 24

    Enhanced reasoning model from google

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →