nowfound

Alternatives

Products that do what Agenomics does

Evaluate AI Agents. Build Real-World Evidence.

  1. 1

    Platform for measuring and training AI agents

    2016

  2. 2

    Open-source monitoring for machine learning models

    2021

  3. 3

    AI agents debate the markets

    May 2026

  4. 4
    Almanac215

    Predict consumer behavior with AI and machine learning

    2020

  5. 5
    Ween.ai195

    The AI platform that turns qualitative data into insights

    2023

  6. 6

    Open-source pull requests AI agent

    2023

  7. 7

    Your AI Agent's Truth Graph to diagnose symptoms

    Sep 2025

  8. 8

    An open benchmark for AI agents that test APIs

    May 2026

  9. 9
    Agen138

    Fully Autonomous AI Coding Agents

    Mar 2026

  10. 10

    Build the semantic layer that makes AI analytics trustworthy

    Mar 2026

  11. 11

    AI-Powered Econometrics & Data Modeling Platform

    25d ago · datanomics-ai.vercel.app

  12. 12
    Quanty104

    AI powered market analysis platform with GraphQL API

    2024

  13. 13
    Playcode139

    The world's best AI website builder. 10 years in the making.

    Mar 2026

  14. 14

    Your site scores X/100 for AI agents with next steps

    May 2026

  15. 15

    Build grounded, governed, trustworthy data agents

    Jun 2026

  16. 16

    Hiring evaluation built for the AI cheating era

    18d ago · agentr.global

  17. 17AA
  18. 18BA

    Hey HN, For the last couple of months, we have been building an AI agent for continuous statistical analysis, and we're looking for feedback while it's still early in development. We call it BIGWIG - an autonomous agent that is specialised, and very good at, performing advanced statistical analysis, through long traces of iteration and reasoning. As it builds statistical models it also "emits" outputs back to the user that you can then interact with, iterate on and schedule for follow up analysis. While we're still in BETA, we've launched a public analysis site that showcases some of the…

    2025 · askbigwig.com

  19. 19FA

    Hey HN, we built an Econ+Finance database to let AI agents do investment research. We spend a lot of tokens to organize macro releases and SEC filings into a clean format, so that your agents have more context to do actual analysis. The problem AI agents are great at data analysis. But they become ineffective if most of their context window is spent on gathering and cleaning data, instead of validating hypotheses. Data in the wild is messy and rarely standardized. Definitions and measurements change over time. This problem is compounded by a fragmented data universe. Point solutions exist…

    Jul 2026 · github.com

  20. 20FF

    I built Hermes, an open-source Python framework for multi-agent financial research. Most AI “equity research” demos stop at generating text. In practice, real workflows require pulling structured XBRL financials from SEC filings, extracting labeled sections like MD&A and Risk Factors, merging macro and market data, building actual Excel models with formulas, and generating investment memos in Word or PDF. Hermes is designed to handle that full pipeline end to end. It includes 35 financial data tools covering SEC EDGAR (via edgartools), FRED, Yahoo Finance market data, and RSS-based financial…

    Feb 2026 · github.com

  21. 21AR
  22. 22MD

    We’re excited to share ML-Dev-Bench, a new open-source benchmark that tests AI agents on real-world ML development tasks. Unlike typical coding challenges or Kaggle-style competitions, our benchmark simulates end-to-end ML workflows including: - Dataset handling and preprocessing - Debugging model and code failures - Implementing new model architectures - Fine-tuning and improving existing models With 30 diverse tasks, ML-Dev-Bench evaluates agents across critical stages of ML development. To complement this, we built Calipers, a framework that provides systematic performance evaluation and…

    2025 · github.com

  23. 23AO
  24. 24AD

    Hey HN, as a former data analyst, I’ve been tooling around trying to get agents to do my old job. The result is this system that gets you maybe 80% of the way there. I think this is a good data point for what the current frontier models are capable of and where they are still lacking (in this case — hypothesis generation and general data intuition). Some initial learnings: - Generating web app-based reports goes much better if there are explicit templates/pre-defined components for the model to use. - Claude can “heal” broken charts if you give it access to chart images and run a…

    Mar 2026 · rubenflamshepherd.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →