Guardrail by NeatProxy
Full visibility into your AI spend, before it goes rogue
What it does
We burned real money running AI agents and didn't know why until it was too late. That's the whole reason Guardrail exists. It's a local proxy for Claude Code and Codex that shows you where every dollar goes, system prompt, tools, that MCP server you forgot about, before you spend it. When you're about to blow a budget, Guardrail blocks the request. Not after. Before. Built for developers who'd rather see the bill coming than explain it later.
Guardrail is a local-first spend firewall for AI coding agents. It tracks Claude Code and Codex cost in real time and blocks over-budget requests before they reach the provider. Metadata only, no prompts stored.
Guardrail is a local proxy on localhost:4000. It intercepts runaway agent loops, audits silent context taxes, and enforces hard limits before calls hit provider billing or exhaust your 5-hour subscription quota. A coding agent runs unattended for minutes at a time, and a single retry loop can burn more than a month of subscription before anyone looks. The tools that exist today tell you afterwards. Provider billing is a monthly rear-view mirror. By the time a number looks wrong, the money is already spent. A token counter watches the meter run. It has no opinion about when to stop, and no way to act on one. A threshold alert fires after the spend it is warning you about. A budget you cannot…from neatproxy.com
Does the same job
all alternatives →
ElevenAgents Guardrails 2.0Apr 2026 · ▲132Configurable safety control for enterprise agent deployment.


- AOAn open-source AI Gateway with integrated guardrails2024 · github.com · ▲21
Hi HN, I've been developing Portkey Gateway, an open-source AI gateway that's now processing billions of tokens daily across 200+ LLMs. Today, we're launching a significant update: integrated Guardrails at the gateway level. Key technical features: 1. Guardrails as middleware: We've implemented a hooks architecture that allows guardrails to act as middleware in the request/response flow. This enables real-time LLM output evaluation and transformation. 2. Flexible orchestration: The gateway can now route requests based on guardrail verdicts. This allows for complex logic like fallbacks…

- ACAI Coding Agent Guardrails enforced at runtimeApr 2026 · sigmashake.com · ▲5
Hello, looking for some users interested using a devtool that allows developers to centrally manage AI Coding Agent tools that supports all AI Coding Agent tools like Claude Code, Codex, Antigravity, etc. Try it free! https://www.producthunt.com/products/sigma-shake-governance-...
More ai this month
the category →
I trained a 125M-parameter transformer to autocomplete piano performances in real time (~108 notes/sec on an iPhone 15). The idea is basically GitHub Copilot or Tabnine, except instead of prompting it with code, you prompt it by playing a few notes on a MIDI piano. The model then continues what you played, entirely on-device. The app is free if anyone wants to try it. Happy to answer questions about the model, training, Core ML, or the many things that didn't work.
AI · 16d ago · simedw.com
Astute▲585Automate your B2B brand going viral, with new media creators
AI · 18d ago · company-app.joinastute.com


Hey HN, Henry from Cactus here! We previously released Cactus Needle, a 14MB agentic LLM for tool call, device use, and structured extraction for phones, wearables, smart homes, small robots and microcontrollers. We got really great feedback here, and have now incorporated the suggestions to release Needle 2. The whole model is a single 14MB binary that runs a full session in 28MB of RAM; 45m parameters at 2bit compression. Needle hits 500 tokens/sec decode speed on a Raspberry Pi 5, sits between 400-1,500 tokens/sec on VR devices like Meta Quest 3S and Apple Vision Pro, and ranges…
AI · 26d ago · cactuscompute.com


Launched alongside, August 2026
the whole month →- TL
Life & fun · 9d ago · louisabraham.github.io


- SA
Hello HN! I found that picking out plausible but diverse skin tones for my digital art and game development projects was kind of difficult, and I got curious about if there was a way to define a color space that made it easy. I've built a color picker and procedural generation algorithm based on the space as well as a bunch of other fun js features and demos throughout the page that use the equations. If you find it interesting, I have lots of explanations of how I built it and what properties the space has. The methodology might be a bit shaky, but hopefully the result is as helpful for…
Life & fun · Aug 2026 · toneyalexander.github.io


I trained a 125M-parameter transformer to autocomplete piano performances in real time (~108 notes/sec on an iPhone 15). The idea is basically GitHub Copilot or Tabnine, except instead of prompting it with code, you prompt it by playing a few notes on a MIDI piano. The model then continues what you played, entirely on-device. The app is free if anyone wants to try it. Happy to answer questions about the model, training, Core ML, or the many things that didn't work.
AI · 16d ago · simedw.com