nowfound

Alternatives

Products that do what Benchmark for LLM Bias Detection does

Uncover AI bias with GENbAIs: Your LLM fairness checker!

  1. 1
    LLM Stats308

    Compare API models by benchmarks, cost & capabilities

    Oct 2025

  2. 2

    Monitor what ChatGPT, Google Gemini and Claude recommend

    2025

  3. 3
    Selene 1196

    Evaluate your AI app with the most accurate LLM Judge

    2025

  4. 4
    LLMrefs315

    AI SEO Keyword Rank Tracker for LLM Search Engines

    2025

  5. 5

    Test-driven development for LLMs

    2023

  6. 6

    Check your brand's visibility on ChatGPT and Google Gemini

    2025

  7. 7AF

    We’ve built an AI risk assessment tool designed specifically for GenAI/LLM applications. It's still early, but we’d love your feedback. Here’s what it does: 1. it performs comprehensive AI risk assessments by analyzing your codebase against different AI regulation/framework or even internal policies. It identifies potential issues and suggests fixes directly through one click PRs. 2. the first framework the platform supports is OWASP Top 10 for LLM Applications 2025, upcoming framework will be ISO 42001 as well as custom policy documents. 3. we're a small, early stage team, so the…

    2025 · gettavo.com

  8. 8

    Like Ahrefs for LLM optimization

    2024

  9. 9

    An open benchmark for AI agents that test APIs

    May 2026 · resources.kusho.ai

  10. 10

    Track and improve your visibility on AI Search

    Dec 2025 · llmpulse.ai

  11. 11

    AI transparency & censorship monitoring for GPT and LLMs

    2025

  12. 12DA
  13. 13

    Easily track your rank in LLMs

    2025

  14. 14

    AI-powered LLM discovery and comparison tool

    Jun 2026 · llmscout.co

  15. 15

    The technical SEO platform for ChatGPT and AI Search

    Jun 2026 · aisearchradar.io

  16. 16CL

    Hi HN! Run it: OPENROUTER_API_KEY="sk" npx bff-eval --demo We built a tool to help people take LLM outputs and easily grade them / eval them to know how good an assistant response is. We've built a number of LLM apps, and while we could ship decent tech demos, we were disappointed with how they'd perform over time. We worked with a few companies who had the same problem, and found out scientifically building prompts and evals is far from a solved problem... writing these things feels more like directing a play than coding. Inspired by Anthropic's constitutional ai concepts, and amazing…

    2025 · github.com

  17. 17

    See which AI search engines are discovering your content.

    Oct 2025

  18. 18AL

    Hi HN! We partnered with the Atlas team to build a tool called AI Predict [0] that allows anyone to ask any question about the future and get a thoroughly researched, AI-generated prediction on how likely it is to be true. How it works: Atlas replicated a Berkeley paper [1] that showed LLMs could make predictions as accurate as the crowd. We’re using a mix of models from OpenAI and Anthropic, with information retrieval powered by NewsCatcher [2]. The system is live and fully functional, though it might struggle with hyper-local questions outside of the public domain (e.g., “Will I have…

    2024 · aipredict.fun

  19. 19PE

    Spelltest framework simulates conversations between AI ‘synthetic users' in an environment to test and refine LLM-based applications. It ensures your app converse with utmost accuracy and relevance. Post-chat, Spelltest assesses responses, providing qualitative and quantitative feedback on performance. Suitable for both chat and completion modes. When to use: - After modifying your prompt. - When your LLM provider updates. - As a CI step for you repo. All feedback and collaborations appreciated!

    2023 · github.com

  20. 20

    Analyse your SEO for AI-powered discovery

    2025

  21. 21HL

    At testup.io we have been working for a while to bring artificial intelligence to the field of test automation. Just a few years ago, the primary challenge laid in accurately identifying UI elements following minor structural changes, such as updates to IDs or paths. The emergence of Large Language Models (LLMs) raised the bar for what it meant to be smart. Now, we anticipate the robot to do lots of things autonomously, such as retry in cases of unresponsiveness or handle minor error reports. A more challenging, but soon expected feature, would involve the test robot navigating your web shop…

    2024 · github.com

  22. 22

    Track your brand's/business visibility in AI responses

    Nov 2025 · visible-landing-page.vercel.app

  23. 23

    Find exact queries to track AI search visibility

    Nov 2025 · wellows.com

  24. 24AR

    Hi HN, I built this open-source LLM red teaming tool based on my experience scaling LLMs at a big co to millions of users... and seeing all the bad things people did. How it works: - Uses an unaligned model to create toxic inputs - Runs these inputs through your app using different techniques: raw, prompt injection, and a chain-of-thought jailbreak that tries to re-frame the request to trick the LLM. - Probes a bunch of other failure cases (e.g. will your customer support bot recommend a competitor? Does it think it can process a refund when it can't? Will it leak your user's address?) -…

    2024 · promptfoo.dev

Ranked by how close each launch is in meaning, then by votes. Refine with a description →