nowfound

Alternatives

Products that do what QwQ AI – Aggregator of free LLMs answer generator does

I've built an aggregator for free Large Language Models that provides answer generation services. The project aims to make powerful AI accessible to everyone as I believe free LLMs may become a significant trend. Currently supported models: Qwen Series - Qwen 32B: Alibaba's 32B parameter model for Chinese/English content - Qwen 2.5 7B Instruct: Lightweight, responsive daily assistant DeepSeek Series - DeepSeek V3 0324: Specialized in long-text and domain knowledge - DeepSeek R1: Focused on mathematical and logical reasoning Google Series - Gemini 2.5 Pro: Google's latest multimodal…

  1. 1
    QwQ-32B197

    Matching R1 reasoning yet 20x smaller

    2025

  2. 2
    Qwen 2.5290

    Alibaba's latest AI model series

    2025

  3. 3
    QWQ-Max126

    New LLM by Alibaba excelling in reasoning w/ "thinking mode"

    2025

  4. 4

    0.8B-9B native multimodal w/ more intelligence, less compute

    Mar 2026

  5. 5
    Qwen3.5307

    The 397B native multimodal agent with 17B active params

    Feb 2026

  6. 6

    Large language model series developed by Alibaba Cloud

    2025

  7. 7

    Qwen's most advanced reasoning model yet

    2025

  8. 8RA
  9. 9
    Qwen3149

    Think Deeper or Act Faster

    2025

  10. 10

    The open sparse MoE model for agentic coding

    Apr 2026

  11. 11

    Qwen's now in mobile chat

    2025

  12. 12Q2

    Last week was big for open source LLMs. We got: - Qwen 2.5 VL (72b and 32b) - Gemma-3 (27b) - DeepSeek-v3-0324 And a couple weeks ago we got the new mistral-ocr model. We updated our OCR benchmark to include the new models. We evaluated 1,000 documents for JSON extraction accuracy. Major takeaways: - Qwen 2.5 VL (72b and 32b) are by far the most impressive. Both landed right around 75% accuracy (equivalent to GPT-4o’s performance). Qwen 72b was only 0.4% above 32b. Within the margin of error. - Both Qwen models passed mistral-ocr (72.2%), which is specifically trained for OCR. - Gemma-3…

    2025 · github.com

  13. 13

    Qwen’s most capable model for coding and cowork

    Aug 2026 · qwen.ai

  14. 14

    A powerful open model for agentic coding tasks

    2025

  15. 15

    Enter website URL, get a ready AI generated FAQ

    2024

  16. 16IB

    We show the potential of modern, embedded graph databases in the browser by demonstrating a fully in-browser chatbot that can perform Graph RAG using Kuzu (the graph database we're building) and WebLLM, a popular in-browser inference engine for LLMs. The post retrieves from the graph via a Text-to-Cypher pipeline that translates a user question into a Cypher query, and the LLM uses the retrieved results to synthesize a response. As LLMs get better, and WebGPU and Wasm64 become more widely adopted, we expect to be able to do more and more in the browser in combination with LLMs, so a lot of…

    2025 · blog.kuzudb.com

  17. 17

    Multimodal AI optimized for real-world coding agents

    Apr 2026

  18. 18

    The end-to-end model powering multimodal chat

    2025

  19. 19

    The sweet-spot open dense model for coding agents

    Apr 2026

  20. 20DA

    I've built an advanced RAG (Retrieval-Augmented Generation) pipeline from scratch to demystify the complex mechanics of modern LLM-powered Question Answering systems. This repository features: -- An implementation of a sub-question query engine from scratch to answer complex user questions. -- Illustrative explanations that unveil the inner workings of the system. -- An analysis of the challenges I faced while working with the system, like prompt engineering and cost estimation. -- Qualitative comparison with similar frameworks like LlamaIndex, offering a broader perspective. Key Takeaway:…

    2023 · github.com

  21. 21

    An AI assistant to instantly generate FAQs for your website

    2024

  22. 22WM

    Try it out! https://glhf.chat/ Hey HN! We’ve been working for the past few months on a website to let you easily run (almost) any open-source LLM on autoscaling GPU clusters. It’s free for now while we figure out how to price it, but we expect to be cheaper than most GPU offerings since we can run the models multi-tenant. Unlike Together AI, Fireworks, etc, we’ll run any model that the open-source vLLM project supports: we don’t have a hardcoded list. If you want a specific model or finetune, you don’t have to ask us for it: you can just paste the Hugging Face link in and…

    2024 · glhf.chat

  23. 23

    Chat with 300+ AI models in one place with 20+ free

    Jul 2026 · chats-llm.com

  24. 24
    Taylor AI118

    Fine-tune open source LLMs in minutes

    2023

Ranked by how close each launch is in meaning, then by votes. Refine with a description →