nowfound

Dev tools · alternatives · 2026

24 alternatives to Gutsy – a 0.8B Jev-compatible decision model that runs on your CPU

Local decision model with calibrated probabilities: send a state and yes/no, choice or score questions, get a probability for every option. 0.8B GGUF on CPU, Jev-style API. - kouhxp/gutsy

Below are 24 products that do a similar job, ranked by how close each is in meaning and then by launch-day votes.

  1. 1
    Jev▲544

    Fast, structured AI decisions for software automation

    12d ago · console.typesafe.ai · its alternatives →

  2. 2

    We've open-sourced Klarity - a tool for analyzing uncertainty and decision-making in LLM token generation. It provides structured insights into how models choose tokens and where they show uncertainty. What Klarity does: - Real-time analysis of model uncertainty during generation - Dual analysis combining log probabilities and semantic understanding - Structured JSON output with actionable insights - Fully self-hostable with customizable analysis models The tool works by analyzing each step of text generation and returns a structured JSON: - uncertainty_points: array of {step, entropy,…

    2025 · github.com · its alternatives →

  3. 3

    September 2026. Every number here is from the benchmarks, and bash experiments/bench.sh --no-record reruns them without an API key.

    3d ago · jevstiller.pages.dev · its alternatives →

  4. 4

    I'm Vivek, co-founder/CEO of HackerRank (YC S11); You may know us as a hiring tool for developers/companies. Over the years, we have built up deep expertise in generating programming challenges, and we are now using that to make coding models better. Our first launch is Model Kombat -- an arena where you can directly compare anonymized coding models, side by side, on real problems. * Pick an arena (Java, Python, etc.) * Each battle has 3 rounds: see the problem + two model outputs -> vote on which you’d actually prefer * Leaderboards + problem statements are updated weekly. We…

    2025 · astra.hackerrank.com · its alternatives →

  5. 5

    I built LocalGPT over 4 nights as a Rust reimagining of the OpenClaw assistant pattern (markdown-based persistent memory, autonomous heartbeat tasks, skills system). It compiles to a single ~27MB binary — no Node.js, Docker, or Python required. Key features: - Persistent memory via markdown files (MEMORY, HEARTBEAT, SOUL markdown files) — compatible with OpenClaw's format - Full-text search (SQLite FTS5) + semantic search (local embeddings, no API key needed) - Autonomous heartbeat runner that checks tasks on a configurable interval - CLI + web interface + desktop GUI - Multi-provider:…

    Feb 2026 · github.com · its alternatives →

  6. 6

    Agent evals and guardrails in one request. Built on Jev, Kev and Laya. - openlayer-ai/jevals

    12d ago · github.com · its alternatives →

  7. 7
    Kodus▲145

    Hey everyone, Just wanted to share a quick update we just launched at Kodus. For those who don’t know it yet, Kodus is a code review agent that runs directly in your team’s Git workflow (GitHub, GitLab, Bitbucket… and now Azure DevOps as well). It helps maintain code quality and consistency by analyzing each PR based on your team’s own rules and repository standards. Support for Azure had been a common request — so if your team uses it, you can now plug Kodus right into your workflow and give it a try. Docs: https://docs.kodus.io/how_to_use/en/overview Repo:…

    Oct 2025 · kodus.io · its alternatives →

  8. 8

    Jev-shaped (TypeSafe System One) classification wrapper over OpenAI-like clients: probabilities and confidence instead of prose - zhulinchng/jevper

    9d ago · github.com · its alternatives →

  9. 9
    SelfJev▲80

    Jev-compatible self-hosted decisions mode

    4d ago · selfjev.dev · its alternatives →

  10. 10

    Provide an input CSV and a target field to predict, generate a model + code to run it. - minimaxir/automl-gs

    2019 · github.com · its alternatives →

  11. 11

    high performance JSON encoder/decoder with stream API for Golang - GitHub - francoispqt/gojay: high performance JSON encoder/decoder with stream API for Golang

    2018 · github.com · its alternatives →

  12. 12

    Working on Mac, Linux, and Windows now. I include a simple GUI to find new models and get things built and set up. It is working quite well across a few models for me. The GitHub README and DESIGN.md files go into detail of the how/why and it's working remarkably well so far. https://github.com/notactuallytreyanastasio/shoehorn

    Aug 2026 · notactuallytreyanastasio.github.io · its alternatives →

  13. 13

    A pipeline that translates Rust GPU code into formal Coq models, as a foundation for memory model proofs - neelsomani/vericuda

    Oct 2025 · github.com · its alternatives →

  14. 14

    A web-based application for quick, scalable, and automated hyperparameter tuning and stacked ensembling in Python. - reiinakano/xcessiv

    2017 · github.com · its alternatives →

  15. 15

    I've come to prefer modeling state machines with nested states, but implementing them by hand requires a large amount of error-prone boilerplate that obscures the important logic of the state machine. I found that the existing offerings in the Rust ecosystem - while inspiring - lacked a number of features that I was looking for, so I spent the last few months building moku. Feedback is welcome!

    2025 · docs.rs · its alternatives →

  16. 16

    Jev returns a decision in 227 ms. The chat models take 2.5 to 3.5 seconds. Pong where the ball moves one step per model decision. Slow model, slow ball.

    14d ago · jev-pong.ably.dev · its alternatives →

  17. 17

    ⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary. - Michael-A-Kuykendall/shimmy

    2025 · github.com · its alternatives →

  18. 18

    TLDR: A small, vendor-agnostic inference loop that turns token logprobs/perplexity/entropy into an extra pass and reasoning for LLMs. - Captures logprobs/top-k during generation, computes perplexity and token-level entropy. - Triggers at most one refine when simple thresholds fire; passes a compact “uncertainty report” (uncertain tokens + top-k alts + local context) back to the model. - In our tests on technical Q&A / math / code, a small model recovered much of “reasoning” quality at ~⅓ the cost while refining ~⅓ of outputs. I kept seeing “reasoning” models behave…

    2025 · github.com · its alternatives →

  19. 19OS

    Large Language Models (LLMs) are powerful, but they’re limited by fixed context windows and outdated knowledge. What if your AI could access live search, structured data extraction, OCR, and more—all through a standardized interface? We built the JigsawStack MCP Server, an open-source implementation of the Model Context Protocol (MCP) that lets any AI model call external tools effortlessly. Here’s what it unlocks: - Web Search & Scraping: Fetch live information and extract structured data from web pages. - OCR & Structured Data Extraction: Process images, receipts, invoices, and handwritten…

    2025 · its alternatives →

  20. 20

    A fast, native app for local coding agents. Amp, Claude Code, Codex, Cursor, OpenCode, Grok, and Pi — one timeline, entirely on your machine.

    Aug 2026 · waku.sh · its alternatives →

  21. 21

    Wrote this to learn more about the `chumsky` parser combinator library, rustyline, and the `ariadne` error reporting crate. Such a nice DX combo for writing new languages. Still a work in progress, but I thought I'd share :)

    2025 · github.com · its alternatives →

  22. 22AT

    Hey HN! Erik here from banana.dev We’ve trained a small(ish) language model on structured extraction, and today we’re launching a playground for it at https://anythingtojson.com. Give it a try! This model continues our work on structured generation, following last week’s launch of Fructose[1], a python client for strongly-typed LLM responses. There seem to be two distinct halves of the problem intended to be solved by Fructose and structured generation: 1. the reasoning ability of the model, such as performing chain of thought, creative acts, and natural language tasks. In a way,…

    2024 · anythingtojson.com · its alternatives →

  23. 23

    runNburn is an Apache-2.0 Rust inference engine for quantized GGUF models that are too big for your fast memory. The core idea: weights stay file-backed (mmap), host residency stays under an explicit byte budget (--ram-budget), and GPU caches are sized from detected free/total VRAM — never from device-name presets. There is no conversion step, no sidecar cache files, no silent requantization. The GGUF on disk is the single source of truth. The result that made me want to post this: Tencent's Hy3 (295B total / 21B active sparse MoE, a single 97.8 GiB Q2_K GGUF) runs on my desktop…

    Jul 2026 · github.com · its alternatives →

  24. 24
    Kody▲26

    The home your agents share. Switch agents. Keep your stuff.

    25d ago · kody.codes · its alternatives →

Also compare

Ranked by how close each launch is in meaning, then by votes. Prices were read from each product’s own site when checked and can change. Refine with your own description →