nowfound

Alternatives

Products that do what Webb does

AI research platform for massive document datasets

  1. 1OA

    Hi HN, I built an open-source AI agent that has already indexed and can search the entire Epstein files, roughly 100M words of publicly released documents. The goal was simple: make a large, messy corpus of PDFs and text files immediately searchable in a precise way, without relying on keyword search or bloated prompts. What it does: - The full dataset is already indexed - You can ask natural language questions - Answers are grounded and include direct references to source documents - Supports both exact text lookup and semantic search Discussion around these files is often fragmented. This…

    Jan 2026 · epstein.trynia.ai

  2. 24H

    2021 · hacker-recommended-books.vercel.app

  3. 3WS

    We’ve trained a generative AI model to browse the web and answer questions/retrieve code snippets directly. Unlike ChatGPT, it has access to primary sources and is able to cite them when you hover over an answer (click on the text to go to the source being cited). We also show regular Bing results side-by-side with our AI answer. The model is an 11-billion parameter T5-derivative that has been fine-tuned on feedback given on hundreds of thousands of searches done (anonymously) on our platform. Giving the model web access lessens its burden to need to store a snapshot of human knowledge…

    2022 · beta.sayhello.so

  4. 4
    Fabric696

    Your new home on the internet

    2023 · fabric.so

  5. 5

    Revolutionizing SEC document analysis

    2023

  6. 6UC

    Paste in my prompt to Claude Code with an embedded API key for accessing my public readonly SQL+vector database, and you have a state-of-the-art research tool over Hacker News, arXiv, LessWrong, and dozens of other high-quality public commons sites. Claude whips up the monster SQL queries that safely run on my machine, to answer your most nuanced questions. There's also an Alerts functionality, where you can just ask Claude to submit a SQL query as an alert, and you'll be emailed when the ultra nuanced criteria is met (and the output changes). Like I want to know when somebody posts about…

    Dec 2025 · exopriors.com

  7. 7RT

    I built a system that monitors ~200,000 news RSS feeds in near real-time and clusters related articles to show how stories spread across the web. It uses Snowflake’s Arctic model for embeddings and HNSW for fast similarity search. Each “story cluster” shows who published first, how fast it propagated, and how the narrative evolved as more outlets picked it up. Would love feedback on the architecture, scaling approach, and any ways to make the clusters more accurate or useful. Live demo: https://yandori.io/news-flow/

    Nov 2025 · yandori.io

  8. 8

    Analyzing and visualizing data from CSV files

    2025

  9. 9

    Accurate and the fastest web search API for AI Agents

    Jan 2026

  10. 10
    AnyParser271

    Accurate, private and configurable document retrieval LLM

    2024

  11. 11EF

    Hey all, Throwaway in case this is assumed to be politcally motivated. I spent some time organizing the Eptstein files to make transparency a little clearer. I need to tighten the data for organizations and people a bit more, but hopeful this is helpful in research in the interim.

    Nov 2025 · searchepsteinfiles.com

  12. 12

    Hi HN! We're X25 alumni and built Parsewise to analyze large document sets with AI agents. Instead of prompting a single PDF, our agents extract, cross-reference, and reason across thousands of documents in one run. We originally built this for insurance and financial diligence workflows where teams review huge document packs. Curious what the HN community thinks! Check out our public demos: demo.parsewise.ai/insurance-claims-triage demo.parsewise.ai/reinsurance-recovery-optimization demo.parsewise.ai/investment-diligence demo.parsewise.ai/mortgage-underwriting

    May 2026 · parsewise.ai

  13. 13SC
  14. 14

    Powerful AI-based data searching and system monitoring tool

    2025

  15. 15

    One API for all documents your AI agents need

    Mar 2026 · querymemory.com

  16. 16SA
  17. 17IB

    Hello everyone, I built this tool on Next/Node to automatically analyze new filings from the SEC and probe the Edgar API for new filings 24/7. We use AI to analyze the filings the second they are released. Free accounts to look at real filings (automatically updated) are available to any who sign up. If you have any questions feel free to ask.

    2024 · docdelta.ca

  18. 18

    Easy goverment data access for citizens, optimized for AI

    Apr 2026 · katzilla.dev

  19. 19

    AI That Works With Your Documents

    Sep 2025

  20. 20IJ

    Hi HackerNews, Lately, I have seen an explosion in posts offering paid APIs/services to get unstructured data into LLMs (i.e. langchain extract, ragflow, unstructured, unstract, just to name a few) and I have been largely disappointed by them, either because they fail to implement multimodal support, fail to give good context for "really tricky" PDFs / Word docs / Powerpoints, or are just plain difficult to use. In light of all these posts I figured I'd share my solution that has been working smoothly for me and my clients. I put it up on GitHub for free so you can check it…

    2024 · github.com

  21. 21IT
  22. 22WS

    We built a search engine interface on top of OpenAI GPT 3.5 and Microsoft Bing that summarizes and cites top search results in response to natural language questions. By using search results, the AI is able to reference recent news and provide citations for specific facts. Our interface offers concise answers, without having to click through links, scroll past irrelevant content, or read ads. No login is required; no personal data is collected. We believe in the power of combining the intuitive UI of web search with the intelligence of large language models. The search engine does indexing…

    2022 · perplexity.ai

  23. 23II

    The DOJ released ~3.5M pages of Epstein documents across 12 datasets. Buried in them are 207 academic papers and 14 books that nobody was really talking about. From what I understand these papers aren't usually freely accesible, but since they are public documents, now they are. I don't know, thought it was interesting to see what this dude was reading. You can check it out at jeescholar.com Pipeline: 1. Downloaded all 12 DOJ datasets + House Oversight Committee release 2. Heuristic pre-filter (abstract detection, DOI regex, citation block patterns, affiliation strings) to cut noise 3. LLM…

    Feb 2026 · jeescholar.com

  24. 24AV

    I feel like LLMs can help me understand anything. However, after I get a summary, I can't dive in to parts that I find interesting; can't refer to original source easily and can't control context with chatbots. This is an attempt to solve for a complete knowledge consumption experience with AI . Please give me feedback!

    Oct 2025 · kerns.ai

Ranked by how close each launch is in meaning, then by votes. Refine with a description →