nowfound

Alternatives

Products that do what Flint – A 30B model fine-tuned for less repetition does

As frontier LLMs have very little output diversity even for open ended queries. We built Flint to see if we could reverse this. It’s a finetuned Qwen3 30B model specifically trained to produce higher entropy when asked open ended questions. Flint significantly increases the NoveltyBench score compared to the base model, without significantly reducing the score on non-creative benchmarks like MMLU-STEM. This shows that that divergence tuning doesn't actually have to be a tax on base capabilities. Flint scores 7.47/10 on NoveltyBench while most frontier models score between 1.8 and 3.2.

  1. 1MR

    Data visualizations are the bridge between user and data. But building AI agents that can generate visualizations reliably can be very tricky: - simple chart specs can be reliable, but generated charts are often of low quality due to reliance on system defaults; - complex chart specs with explicit details can produce good-looking charts, but they are verbose and agents can struggle with reliability We figured out it is a limitation on the language issue (not just AI capability thing) -- current visualization languages are a bit too low-level for AI agents, requiring them to explicitly make…

    Jul 2026 · microsoft.github.io

  2. 2NW

    Hey HN, Henry here from Cactus. We open-sourced Needle, a 26M parameter function-calling (tool use) model. It runs at 6000 tok/s prefill and 1200 tok/s decode on consumer devices. We were always frustrated by the little effort made towards building agentic models that run on budget phones, so we conducted investigations that led to an observation: agentic experiences are built upon tool calling, and massive models are overkill for it. Tool calling is fundamentally retrieval-and-assembly (match query to tool name, extract argument values, emit JSON), not reasoning. Cross-attention…

    May 2026 · github.com

  3. 3

    The open sparse MoE model for agentic coding

    Apr 2026

  4. 4FT

    Aug 2026 · github.com

  5. 5IB

    Built a ~9M param LLM from scratch to understand how they actually work. Vanilla transformer, 60K synthetic conversations, ~130 lines of PyTorch. Trains in 5 min on a free Colab T4. The fish thinks the meaning of life is food. Fork it and swap the personality for your own character.

    Apr 2026 · github.com

  6. 6

    The sweet-spot open dense model for coding agents

    Apr 2026

  7. 7RA
  8. 8

    New LLM compression algorithm by Google

    Mar 2026

  9. 9EF

    I’ve been building Echo (https://echo.tracerml.ai/), an experiment in making one AI system out of a pool of open-weight models rather than choosing a single model and using it for every task. It started with a simple experiment. I took a group of models, including GLM-5.2, Kimi K2.7 and others, and ran them on the same evaluations. Then I measured what would happen if, for each problem, you somehow knew in advance which models would be useful and how their outputs should be combined. That hypothetical system performed substantially better than any individual model in the pool.…

    Jul 2026

  10. 10

    Qwen’s most capable model for coding and cowork

    Aug 2026 · qwen.ai

  11. 11
    Taylor AI118

    Fine-tune open source LLMs in minutes

    2023

  12. 12
    Unsloth241

    Finetune LLMs 2x faster, 80% less memory

    2025

  13. 13TW
  14. 14

    A powerful open model for agentic coding tasks

    2025

  15. 15
    Pioneer113

    Fine-tune any LLM in minutes, with one prompt

    Apr 2026

  16. 16DD

    We recently used DeepSeek V4 Flash as a teacher for finance tasks with GPT-OSS-120B. Distillation works well on this problem. At a constrained 8k token budget, our self-distilled 120B scores 83.61% on FinanceReasoning, above Kimi K3 (81.93%) and Inkling (65.13%). We released the 20B open weights. With V4 as the teacher though, we realized it would be timely to measure if the censorship characteristic of it transferred to the distilled version of the base model. tl;dr it didn't, the teacher answered politically sensitive questions 7 SDs differently than expected, but the distilled model's…

    Jul 2026 · ctgt.ai

  17. 17
    Fluent119

    Agentic AI in Any Mac App. Now with Native RAG

    Jan 2026

  18. 18
    FineTuner164

    Fine-tune AI models on your data — in minutes, not days.

    2025

  19. 19

    The first open model to beat Sonnet made for productivity

    Feb 2026

  20. 20

    LLM reinforcement fine-tuning platform to improve LLM output

    2025

  21. 21

    The open-weight preview of Qwen4

    10d ago · qwen.ai

  22. 22
    Qwen3149

    Think Deeper or Act Faster

    2025

  23. 23

    AI fine-tuning platform to create custom LLMs

    2024

  24. 24

    Large language model series developed by Alibaba Cloud

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →