nowfound

Alternatives

Products that do what Is AI Dumber Today? An index of AI model experience from user's opinion does

Track how AI models feel in everyday use through public community feedback, 7-day experience scores and trends. This is not a capability benchmark.

  1. 1

    AI social media deep research 24/7

    2025

  2. 2

    Curious how AI-fluent your organization is?

    May 2026 · ai-pilled.com

  3. 3

    Collect user feedback & measure analytics in just 30 seconds

    2025

  4. 4AT

    Interactive timeline of every major Large Language Model. Filterable by open/closed source, searchable, 54 organizations tracked.

    Feb 2026 · llm-timeline.com

  5. 5

    A live demo of AI algorithms making judgments about you.

    2020

  6. 6
    Browsee160

    AI assisted tool to understand user experience visually.

    2019

  7. 7IA

    A survey tracking developer sentiment on AI-assisted coding through Hacker News posts.

    Feb 2026 · is-ai-good-yet.com

  8. 8DC

    I’ve been using AI to generate some repetitive frontend (guilty), and while most outputs felt vibe-coded, some results were surprisingly good. So I cleaned it up and made a ranking game out of it with friends, and you can check it out here: https://www.designarena.ai/vote /vote: Your prompt will be answered by four random, anonymous models. You pick the one you prefer and crown the winner, tournament-style. /leaderboard: See the current winning models, as dictated by voter preferences. /play: Iterate quickly by seeing four models respond to the same input and…

    2025 · designarena.ai

  9. 9EL
  10. 10

    Automated user insights from interviews, surveys & more

    2025

  11. 11

    The AI Code Arena

    Sep 2025

  12. 12CA
  13. 13OS

    Hi all! This morning, we released a new Apache 2.0 licensed model on HuggingFace for detecting hallucinations in retrieval augmented generation (RAG) systems. What we've found is that even when given a "simple" instruction like "summarize the following news article," every LLM that's available hallucinates to some extent, making up details that never existed in the source article -- and some of them quite a bit. As a RAG provider and proponents of ethical AI, we want to see LLMs get better at this. We've published an open source model, a blog more thoroughly describing our methodology (and…

    2023 · vectara.com

  14. 14WF

    We have a dataset of 3,095 standardized AI responses across 43 prompts. From each response, we extract a 32-dimension stylometric fingerprint (lexical richness, sentence structure, punctuation habits, formatting patterns, discourse markers). Some findings: - 9 clone clusters (>90% cosine similarity on z-normalized feature vectors) - Mistral Large 2 and Large 3 2512 score 84.8% on a composite metric combining 5 independent signals - Gemini 2.5 Flash Lite writes 78% like Claude 3 Opus. Costs 185x less - Meta has the strongest provider "house style" (37.5x distinctiveness ratio) - "Satirical…

    Apr 2026 · rival.tips

  15. 15

    Your site scores X/100 for AI agents with next steps

    May 2026 · indexedai.tech

  16. 16
    Reason857

    User research on autopilot

    Apr 2026 · reason8.io

  17. 17

    Turn user insight into your competitive edge

    2025

  18. 18

    A 30-second experiment to see how much personal context your AI remembers about you.

    13d ago · withcorpus.com

  19. 19BY

    we had hundreds of discussions with engineering leaders over the past few months, and everyone's trying to understand where they are in the AI journey. we collected all this data into a benchmark and built a free grader to let you know where you stand. you answer on a 1–5 scale (e.g., autonomy runs from "suggestions only" to "agents own multi-hour workflows across code, infra, and external systems") - takes about 5 minutes. https://agent-benchmarks.com/software-factory/ waiting for your results!

    Jul 2026 · agent-benchmarks.com

  20. 20SO

    I’ve been building a crowd-sourced AI detection benchmark. Two responses to the same prompt — one from a real human (pre-2022, provably pre prevalence of AI slop on the internet), one generated by AI. You pick the slop. Three wrong and you’re out. The dataset: 16K human posts from Reddit, Hacker News, and Yelp, each paired with AI generations from 6 models across two providers (Anthropic and OpenAI) at three capability tiers. Same prompt, length-matched, no adversarial coaching — just the model’s natural voice with platform context. Every vote is logged with model, tier, source, response…

    Mar 2026 · slop-or-not.space

  21. 21

    Real-time community signal of models performing best today

    Apr 2026 · model-tracker.com

  22. 22KT

    Hi HN! I built this tool, because Large Language Models are hallucinating their asses off and I wanted to test just how bad it is with a topic I know best - myself. I'm sure there are other egos out there who google themselves and essentially this is the new googling yourself. It's early beta, so lots of room for improvement of course.

    2023 · haveibeenencoded.com

  23. 23

    Performance results of AI coding agents on Next.js

    Feb 2026

  24. 24HT

    2023 · opensourceconnections.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →