nowfound

Alternatives

Products that do what WaveletLM – wavelet-based, attention-free model with O(n log n) scaling does

WaveletLM is a wavelet-based, attention-free architecture that replaces self-attention with learned lifting wavelet decomposition, a Fast Walsh-Hadamard Transform, per-scale gated spectral mixing with SwiGLU activation, an inverse FWHT, and wavelet reconstruction. Combined with expanded MLPs and sparse product-key memory, this yields a model with O(n log n) scaling in sequence length. With 23.8 PPL on WikiText-103, WaveletLM beats both GPT-2 Medium, which was trained on 80× more data, and Transformer-XL Standard, which uses recurrence to extend its effective context. It is undertrained and…

  1. 1IB

    Built a ~9M param LLM from scratch to understand how they actually work. Vanilla transformer, 60K synthetic conversations, ~130 lines of PyTorch. Trains in 5 min on a free Colab T4. The fish thinks the meaning of life is food. Fork it and swap the personality for your own character.

    Apr 2026 · github.com

  2. 2
    GLM-5.3254

    Coding leap from scaled post-training on the same base

    22d ago · z.ai

  3. 3
    Ferret193

    Refer and ground anything anywhere at any granularity

    2024

  4. 4
    GLM-5154

    Open-weights model for long-horizon agentic engineering

    Feb 2026

  5. 5L3

    I spent a lot of time and money on this rather big side project of mine that attempts to replicate the mechanistic interpretability research on proprietary LLMs that was quite popular this year and produced great research papers by Anthropic [1], OpenAI [2] and Deepmind [3]. I am quite proud of this project and since I consider myself the target audience for HackerNews did I think that maybe some of you would appreciate this open research replication as well. Happy to answer any questions or face any feedback. Cheers [1]…

    2024 · github.com

  6. 6
    GLM-4.6V239

    Open-source multimodal model with native tool use

    Dec 2025

  7. 7

    New LLM compression algorithm by Google

    Mar 2026 · research.google

  8. 8
    SmolVLM2206

    Smallest Video LM Ever from HuggingFace

    2025

  9. 9

    Vision-to-code foundation model for real GUI automation

    Apr 2026 · docs.z.ai

  10. 10
    Dream 7B191

    Powerful Open Diffusion LLM, Beyond Autoregressive

    2025

  11. 11OS

    Hi all! This morning, we released a new Apache 2.0 licensed model on HuggingFace for detecting hallucinations in retrieval augmented generation (RAG) systems. What we've found is that even when given a "simple" instruction like "summarize the following news article," every LLM that's available hallucinates to some extent, making up details that never existed in the source article -- and some of them quite a bit. As a RAG provider and proponents of ethical AI, we want to see LLMs get better at this. We've published an open source model, a blog more thoroughly describing our methodology (and…

    2023 · vectara.com

  12. 12
    GLM-4.5298

    Unifying agentic capabilities in one open model

    2025

  13. 13
    NVLM 1.0200

    Open frontier-class multimodal LLMs

    2024

  14. 14

    Auto-regressive for dense-knowledge & high-fidelity images

    Jan 2026

  15. 15
    Kolors248

    Photorealistic text-to-image diffusion model for creators

    2025

  16. 16

    A text to music AI model by Google

    2023

  17. 17I4

    It's our new text-to-image model: a 9.3B single-stream diffusion transformer trained entirely from scratch. We focused heavily on controllability through structured JSON prompts, with strong text rendering, spatial awareness through bounding box guidance, and color palette control. It has the best text rendering of any open-weight model we've tested so far, and the NF4 quantized checkpoint runs on a single 24GB GPU. For more technical details and examples see our blog post: https://ideogram.ai/blog/ideogram-4.0/ We will be happy to answer any questions :)

    Jun 2026 · github.com

  18. 18
    InternVL3135

    Open MLLMs excelling in vision, reasoning & long context

    2025

  19. 19WM

    Try it out! https://glhf.chat/ Hey HN! We’ve been working for the past few months on a website to let you easily run (almost) any open-source LLM on autoscaling GPU clusters. It’s free for now while we figure out how to price it, but we expect to be cheaper than most GPU offerings since we can run the models multi-tenant. Unlike Together AI, Fireworks, etc, we’ll run any model that the open-source vLLM project supports: we don’t have a hardcoded list. If you want a specific model or finetune, you don’t have to ask us for it: you can just paste the Hugging Face link in and…

    2024 · glhf.chat

  20. 20

    The open sparse MoE model for agentic coding

    Apr 2026 · qwen.ai

  21. 21
    GLM-4.720

    Advanced coding & reasoning with multi-turn thinking

    Dec 2025

  22. 22

    Ultra-efficient 1.3B vision-language model for mobile

    May 2026 · github.com

  23. 23FG

    We developed a new framework that enables flexible control of generated text in language models. By combining several models and/or system prompts in one mathematical formula, it lets you tweak your style and combine model outputs with ease. A handy tool for those working with LLMs, looking for more fine-grained control of stylistic output. More details in our paper: https://arxiv.org/abs/2311.14479. Feedback and potential applications are welcome.

    2023 · github.com

  24. 24

    Fine-tuning, RL, and inference in one CLI

    Dec 2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →