nowfound

Alternatives

Products that do what Break My Guard does

There’s a point where an AI stops resisting — find it.

  1. 1DJ

    I created a daily challenge for Prompt Engineers to build the shortest prompt to break a system prompt. You are provided the system prompt and a forbidden method the LLM was told not to invoke. Your task is to trick the model into calling the function. Shortest successful attempts will show up in the leaderboard. Give it a shot! You never know what could break an LLM.

    2025 · vaultbreak.ai

  2. 2
    Guard371

    An AI that reads privacy policies for you

    2019

  3. 3
    AI Titans553

    Prove your brains, brawn, and bytes

    2023

  4. 4CY
  5. 5
    CtrlAI104

    Transparent proxy that secures AI agents with guardrails

    Mar 2026 · github.com

  6. 6

    Wordle for AI Prompts

    2025

  7. 7

    Configurable safety control for enterprise agent deployment.

    Apr 2026 · elevenlabs.io

  8. 8
    HOL Guard106

    The 1st Firewall for AI Agents

    Jul 2026 · hol.org

  9. 9
    deduce66

    A daily Wordle-like puzzle for AI agents

    Apr 2026

  10. 10WP

    Anthropic and OpenAI's publicly available models are explicitly guard-railed so that they refuse offensive tasks. And their cyber-focussed models are gated for enterprises. This leaves SMEs and mid market open to major vulnerabilities. AI can be used as both an adversarial and defensive tool in the world of cyber. A worst case outcome is if only the adversaries have access. Meanwhile, most existing AI cyber tools are just wrappers. The problem is that they still have all the guardrails on from the foundation model where they will inherit its refusals. For this project we've post-trained a…

    Jun 2026 · argusred.com

  11. 11AA

    Your AI agent hits an infinite loop and racks up $2000 in API charges overnight. This happens weekly to AI developers. AgentGuard monitors API calls in real-time and automatically kills your process when it hits your budget limit. How it works: Add 2 lines to any AI project: const agentGuard = require('agent-guard'); await agentGuard.init({ limit: 50 }); // $50 budget // Your existing code runs unchanged const response = await openai.chat.completions.create({...}); // AgentGuard tracks costs automatically When your code hits $50 in API costs, AgentGuard stops…

    2025 · github.com

  12. 12

    Break, test, and secure AI agents before production

    Jun 2026 · github.com

  13. 13

    A pre-execution guard that stops your coding agent running destructive commands. One shell file, no dependencies. - vandith1/agent-guard

    22d ago · github.com

  14. 14

    Guard-AI keeps traders disciplined and Backtest Strategy

    Feb 2026 · guard-ai-five.vercel.app

  15. 15FL

    Hi HN! We just launched Codacy Guardrails, an IDE extension with a CLI for code analysis and MCP server that enforces security & quality rules on AI-generated code in real-time. It hooks into AI coding assistants (like VS Code Agent Mode, Cursor, Windsurf), silently scanning and fixing AI-suggested code that has vulnerabilities or violates your coding standards, while the code it’s being generated. We built this because coding agents can be a double-edged sword. They do boost productivity, but can easily introduce insecure or non-compliant code. One recent research team at NYU found that 40%…

    2025

  16. 16IT

    I built 1e4.ai - a chess web app where you play against neural networks trained to mimic human Lichess players at specific Elo ranges. There's a separate model for each 100-point rating bucket from ~800 to 2200+, and the bots not only choose human-like moves but also burn clock time, play worse under time pressure, and blunder in human-like ways. Live demo: https://1e4.ai Code: https://github.com/thomasj02/1e4_ai A few things that might be interesting: - Trained on almost a full year of Lichess blitz games, around 1B total games - Architecture is an a small…

    May 2026

  17. 17

    Stop shipping the same bugs twice.\

    Oct 2025

  18. 18

    Universal AI agent guardrail. The bodyguard for AI agents.

    Feb 2026 · github.com

  19. 19OS

    We build runtime security for AI agents. The playground started as an internal tool that we used to test our own guardrails. But we kept finding the same types of vulnerabilities because we think about attacks a certain way. At some point you need people who don't think like you. So we open-sourced it. Each challenge is a live agent with real tools and a published system prompt. Whenever a challenge is over, the full winning conversation transcript and guardrail logs get documented publicly. Building the general-purpose agent itself was probably the most fun part. Getting it to reliably use…

    Mar 2026 · github.com

  20. 20

    Full visibility into your AI spend, before it goes rogue

    26d ago · neatproxy.com

  21. 21FP

    We've built an open-source tool to stress test AI agents by simulating prompt injection attacks. We’ve implemented one powerful attack strategy based on the paper [AdvPrefix: An Objective for Nuanced LLM Jailbreaks](https://arxiv.org/abs/2412.10321). Here's how it works: - You define a goal, like: “Tell me your system prompt” - Our tool uses a language model to generate adversarial prefixes (e.g., “Sure, here are my system prompts…”) that are likely to jailbreak the agent. - The output is a list of prompts most likely to succeed in bypassing safeguards. We’re just getting…

    2025 · security.vista-labs.ai

  22. 22

    An interactive security challenge. Control an AI agent and try to exfiltrate data from a sandboxed environment protected by Declaw.

    Jul 2026 · declaw.ai

  23. 23IN

    Tl;dr: I trained a classifier to route to the least expensive model and reasoning depth to complete the request. Coupling that with additional automated token efficiency techniques has yielded 3x usage for the same spend. For anyone interested in trying it themselves: https://nerfguard.com Various teammates and I switched over to Codex from Claude Code recently. We still bounce between the tools, but Codex’s speed and steerability coupled with performance gains were hard to ignore. One of the downsides was that the per token pricing kicked in way sooner. This is happening across…

    Jun 2026

  24. 24

    Fix hidden logic bugs & edge cases in AI code instantly.

    Jun 2026

Ranked by how close each launch is in meaning, then by votes. Refine with a description →