nowfound

Alternatives

Products that do what QwQ-32B APIs – o1 like reasoning at 1% the cost does

Ubicloud is an open source alternative to AWS. Today, we launched our inference APIs, built with open source AI models. QwQ-32B-Preview is one of those models; and it can provide o1-like reasoning at 1% the cost. QwQ is licensed under Apache 2.0 [1] and Ubicloud under AGPL v3. We deploy open models on a cloud stack that can run anywhere. This allows us to offer great price / performance. From an accuracy standpoint, QwQ does well in math and coding domains. For example, in the MMLU-Pro Computer Science LLM Benchmark, the accuracy rankings are as follows. Claude-3.5 Sonnet (82.5),…

  1. 1
    QwQ-32B197

    Matching R1 reasoning yet 20x smaller

    2025

  2. 2

    Qwen's most advanced reasoning model yet

    2025

  3. 3

    Pushing the frontier of cost-effective reasoning

    2025

  4. 4Q2

    Last week was big for open source LLMs. We got: - Qwen 2.5 VL (72b and 32b) - Gemma-3 (27b) - DeepSeek-v3-0324 And a couple weeks ago we got the new mistral-ocr model. We updated our OCR benchmark to include the new models. We evaluated 1,000 documents for JSON extraction accuracy. Major takeaways: - Qwen 2.5 VL (72b and 32b) are by far the most impressive. Both landed right around 75% accuracy (equivalent to GPT-4o’s performance). Qwen 72b was only 0.4% above 32b. Within the margin of error. - Both Qwen models passed mistral-ocr (72.2%), which is specifically trained for OCR. - Gemma-3…

    2025 · github.com

  5. 5

    AI models that run on an inference cloud optimized for speed

    May 2026 · generalcompute.com

  6. 6
    GLM-4.5298

    Unifying agentic capabilities in one open model

    2025

  7. 7

    The open sparse MoE model for agentic coding

    Apr 2026 · qwen.ai

  8. 8
    OpenAI o1685

    AI that can do general-purpose complex reasoning

    2024 · openai.com

  9. 9
    Qwen3.5307

    The 397B native multimodal agent with 17B active params

    Feb 2026

  10. 10
    Dograh592

    The open source VAPI alternative

    25d ago · dograh.com

  11. 11IP

    The stack: two agents on separate boxes. The public one (nullclaw) is a 678 KB Zig binary using ~1 MB RAM, connected to an Ergo IRC server. Visitors talk to it via a gamja web client embedded in my site. The private one (ironclaw) handles email and scheduling, reachable only over Tailscale via Google's A2A protocol. Tiered inference: Haiku 4.5 for conversation (sub-second, cheap), Sonnet 4.6 for tool use (only when needed). Hard cap at $2/day. A2A passthrough: the private-side agent borrows the gateway's own inference pipeline, so there's one API key and one billing relationship…

    Mar 2026 · georgelarson.me

  12. 12
    Qwen 2.5290

    Alibaba's latest AI model series

    2025

  13. 13

    OpenAI's most advanced model, o1 API to third-party devs

    2024

  14. 14

    How small can a language model be while still doing something useful? I wanted to find out, and had some spare time over the holidays. Z80-μLM is a character-level language model with 2-bit quantized weights ({-2,-1,0,+1}) that runs on a Z80 with 64KB RAM. The entire thing: inference, weights, chat UI, it all fits in a 40KB .COM file that you can run in a CP/M emulator and hopefully even real hardware! It won't write your emails, but it can be trained to play a stripped down version of 20 Questions, and is sometimes able to maintain the illusion of having simple but terse conversations…

    Dec 2025 · github.com

  15. 15
    QWQ-Max126

    New LLM by Alibaba excelling in reasoning w/ "thinking mode"

    2025

  16. 16

    The sweet-spot open dense model for coding agents

    Apr 2026 · qwen.ai

  17. 17
    oqoqo340

    Build evals and custom benchmarks for real-world tasks

    27d ago · oqoqo.ai

  18. 18

    0.8B-9B native multimodal w/ more intelligence, less compute

    Mar 2026

  19. 19
    Oxlo.ai388

    Scale across AI models without scaling your bill

    Jun 2026 · oxcode.ai

  20. 20

    gpt-oss-120b and gpt-oss-20b open-weight language models

    2025

  21. 21

    Qwen’s most capable model for coding and cowork

    Aug 2026 · qwen.ai

  22. 22IB

    Hey HN, I am proud to show you guys that I have built an open source alternative to Azure OpenAI services. Azure OpenAI services was born out of companies needing enhanced security and access control for using different GPT models. I want to build an OSS version of Azure OpenAI services that people could self host in their own infrastructure. "How can I track LLM spend per API key?" "Can I create a development OpenAI API key with limited access for Bob?" "Can I see my LLM spend breakdown by models and endpoints?" "Can I create 100 OpenAI API keys that my students could use in a classroom…

    2023 · github.com

  23. 23

    Q3AS, deployment & execution of quantum algorithms by Aqora

    2024

  24. 24BO

    Hi HN, we are Deepak and Sama, co-founders of BoxyHQ (https://boxyhq.com/). BoxyHQ provides an open-source platform for developers to quickly integrate enterprise features into their software solutions. These include SAML Single Sign-On (SSO), Audit logs, with more to come :) Every B2B startup faces a common challenge when it comes to selling into the Enterprise; they need to allocate time and resources to support all the requirements to make their offering enterprise-grade. Supporting these requirements is a significant undertaking for the engineering team, especially since…

    2022 · boxyhq.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →