nowfound

Alternatives

Products that do what Echo – Fable-level results at 1/3 the cost using open-weight models does

I’ve been building Echo (https://echo.tracerml.ai/), an experiment in making one AI system out of a pool of open-weight models rather than choosing a single model and using it for every task. It started with a simple experiment. I took a group of models, including GLM-5.2, Kimi K2.7 and others, and ran them on the same evaluations. Then I measured what would happen if, for each problem, you somehow knew in advance which models would be useful and how their outputs should be combined. That hypothetical system performed substantially better than any individual model in the pool.…

  1. 1
    GLM-4.5298

    Unifying agentic capabilities in one open model

    2025

  2. 2DA
  3. 3GG

    A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me. But then I thought, "I wonder how it would work on a normal computer like mine," and above all, "I wonder if it would work without going into OOM on a computer like mine." So I started working with the help of agents to test this possibility. I started converting the model to int4, understanding MTP usage, and if possible implementing DSA for long context.…

    Jul 2026 · github.com

  4. 4

    Build Powerful Voice Agents

    2025

  5. 5
    OpenAI o1685

    AI that can do general-purpose complex reasoning

    2024 · openai.com

  6. 6

    gpt-oss-120b and gpt-oss-20b open-weight language models

    2025

  7. 7
    GLM-5154

    Open-weights model for long-horizon agentic engineering

    Feb 2026

  8. 8

    Turn your ideas into outlines, without AI slop

    2025

  9. 9
    Oxlo.ai388

    Scale across AI models without scaling your bill

    Jun 2026 · oxcode.ai

  10. 10

    Discover, compare, and choose AI models—100% Free

    2024

  11. 11

    High-speed agentic model built specifically for OpenClaw

    Mar 2026

  12. 12

    Generating uncanny AI avatars is now open source

    May 2026 · avaturn.live

  13. 13
    Zoo234

    A free, open-source playground for AI image models

    2023

  14. 14MO

    I wanted to share our new speech to text model, and the library to use them effectively. We're a small startup (six people, sub-$100k monthly GPU budget) so I'm proud of the work the team has done to create streaming STT models with lower word-error rates than OpenAI's largest Whisper model. Admittedly Large v3 is a couple of years old, but we're near the top the HF OpenASR leaderboard, even up against Nvidia's Parakeet family. Anyway, I'd love to get feedback on the models and software, and hear about what people might build with it.

    Feb 2026 · github.com

  15. 15
    Echo192

    Capture easier, think better with AI voice and text notes

    2024

  16. 16IM
  17. 17
    RunInfra156

    Describe the AI model you need and get an optimized AI

    Jul 2026 · runinfra.ai

  18. 18
    Inkling181

    Open weights 975B multimodal model built for fine-tuning

    Jul 2026 · thinkingmachines.ai

  19. 19EL
  20. 20OS

    Our goal with this project is to build a completely open source, state of the art turn detection model that can be used in any voice AI application. I've been experimenting with LLM voice conversations since GPT-4 was first released. (There's a previous front page Show HN about Pipecat, the open source voice AI orchestration framework I work on. [1]) It's been almost two years, and for most of that time, I've been expecting that someone would "solve" turn detection. We all built initial, pretty good 80/20 versions of turn detection on top of VAD (voice activity detection) models. And…

    2025 · github.com

  21. 21

    Document smarter. For humans and AI.

    2024

  22. 22
    Prosed155

    Go from newsletters & podcasts to published manuscript

    May 2026 · tryprosed.com

  23. 23IR

    The Emotion Engine has 32 MB of RAM total, so the trick is streaming weights from CD-ROM one matrix at a time during the forward pass — only activations, KV cache and embeddings live in RAM. This means models bigger than the RAM can still run, they just read more from disc. Had to build a custom quantized format (PSNT), hack endianness, write a tokenizer pipeline, and most of the PS2 SDK from scratch (releasing that separately). The model itself is also custom — a 10M param Llama-style architecture I trained specifically for this. And it works. On real hardware.

    Mar 2026 · github.com

  24. 24
    EchoHQ112

    Redefining customer care, 24/7

    2023

Ranked by how close each launch is in meaning, then by votes. Refine with a description →