nowfound

Alternatives

Products that do what ModelRed does

Your AI is vulnerable. We'll prove it.

  1. 1

    Red-team any AI system in minutes

    Nov 2025

  2. 2

    An open benchmark for AI agents that test APIs

    May 2026 · resources.kusho.ai

  3. 3

    Anthropic and OpenAI's publicly available models are explicitly guard-railed so that they refuse offensive tasks. And their cyber-focussed models are gated for enterprises. This leaves SMEs and mid market open to major vulnerabilities. AI can be used as both an adversarial and defensive tool in the world of cyber. A worst case outcome is if only the adversaries have access. Meanwhile, most existing AI cyber tools are just wrappers. The problem is that they still have all the guardrails on from the foundation model where they will inherit its refusals. For this project we've post-trained a…

    Jun 2026 · argusred.com

  4. 4

    Trace, evaluate, and improve AI agents in production

    Aug 2026 · telerik.com

  5. 5

    Get a pentest done, today.

    Dec 2025 · aikido.dev

  6. 6

    The AI Code Arena

    Sep 2025

  7. 7AR

    Hi HN, I built this open-source LLM red teaming tool based on my experience scaling LLMs at a big co to millions of users... and seeing all the bad things people did. How it works: - Uses an unaligned model to create toxic inputs - Runs these inputs through your app using different techniques: raw, prompt injection, and a chain-of-thought jailbreak that tries to re-frame the request to trick the LLM. - Probes a bunch of other failure cases (e.g. will your customer support bot recommend a competitor? Does it think it can process a refund when it can't? Will it leak your user's address?) -…

    2024 · promptfoo.dev

  8. 8

    Uncensored AI models for red team research.

    Dec 2025 · shannon-ai.com

  9. 9EY

    I built an open-source AI agent for security testing to find and fix vulnerabilities in your code. I’ve noticed how bad security vulnerabilities have gotten with everyone shipping AI code slop, so I wanted to build something that allows for vibe-coding at full speed without compromising security. Traditional security tools aren’t effective, and manual pen-testing can’t keep up with the rapidly growing AI code This tool runs your code dynamically, finds vulnerabilities, and validates them through actual exploitation. You can either run it against your codebase or enter your (or someone…

    2025 · github.com

  10. 10

    Security-scored directory for AI skills and agent tools

    Feb 2026 · skillshield.io

  11. 11

    Open-source security gateway & static scanner for AI agents. Enforce role-based access control (RBAC), human-in-the-loop approvals, segregation of duties, and cryptographically signed, offline-verifiable audit logs. - makerchecker/MakerChecker

    Jul 2026 · github.com

  12. 12

    The Only AI Tool That Doesn't Trust AI

    Mar 2026 · triall.ai

  13. 13OS

    We build runtime security for AI agents. The playground started as an internal tool that we used to test our own guardrails. But we kept finding the same types of vulnerabilities because we think about attacks a certain way. At some point you need people who don't think like you. So we open-sourced it. Each challenge is a live agent with real tools and a published system prompt. Whenever a challenge is over, the full winning conversation transcript and guardrail logs get documented publicly. Building the general-purpose agent itself was probably the most fun part. Getting it to reliably use…

    Mar 2026 · github.com

  14. 14

    Compare AI models side-by-side on same prompt

    Feb 2026 · testaimodels.com

  15. 15

    See how easily your AI can be broken — in seconds

    May 2026 · huggingface.co

  16. 16

    Attack-test your AI agents and grade what they did

    14d ago · tryredlineai.co

  17. 17CB

    AI agents now have impressive reasoning capabilities. This raises an important question: how dangerous are these AI agents at identifying & exploiting web vulnerabilities? We created CVE-bench to find out (I'm one contributor of 16). To our knowledge CVE-bench is the first benchmark using real-world web vulnerabilities to evaluate AI agents' cyberattack capabilities. We included 40 CVEs from NIST's database, focusing on critical-severity vulnerability (CVSS > 9.0). To properly evaluate agents’ attacks, we built isolated environments with containerization and identified 8 common attack…

    2025 · github.com

  18. 18

    Security testing for AI agents

    Mar 2026 · zeroleaks.ai

  19. 19

    Infrastructure for Continuous Evals of AI Agents

    May 2026 · cipherra.ai

  20. 20

    When one AI is hit, all get stronger

    Oct 2025

  21. 21

    Find AI vulnerabilities before hackers do

    Mar 2026 · promptbrake.com

  22. 22

    Expand eval coverage & use red agents to break AI systems

    25d ago · mutant.aiankit.com

  23. 23

    Production failures become regression tests for AI agents

    27d ago · tracely-ai.com

  24. 24AT

    Hi Hacker News! We're launching Zalor, an agent testing platform. Agents often break when you tweak system prompts, swap models, or add tools. Zalor automatically generates test scenarios and evaluates your agent so you know it's reliable before deploying to production. We currently support the OpenAI Agents SDK and are onboarding other frameworks. A GitHub integration is coming so you can get feedback on every update. Looking forward to hearing feedback from people building agents.

    Mar 2026 · agents.zalor.ai

Ranked by how close each launch is in meaning, then by votes. Refine with a description →