nowfound

Alternatives

Products that do what DataFrog does

Offline tool to fuzzy match, merge messy Excel files No-Code

  1. 1
    Flatfile1,070

    The data onboarding platform

    2019 · flatfile.com

  2. 2
    Capalyze809

    ChatGPT for datavores: scrape → ask → visualize

    Sep 2025 · capalyze.ai

  3. 3

    Stop cleaning data. Start using it.

    Oct 2025

  4. 4

    Clean messy data in minutes and without coding

    2020

  5. 5

    Fastest way: csv/xls to dashboard report, no ChatGPT upload

    2023

  6. 6

    Data scraping without code

    2024

  7. 7FA
  8. 8
    Dropbase227

    Turn your CSV and Excel files to live databases, instantly.

    2020

  9. 9

    40+ free tools to convert, format and transform data.

    Jun 2026 · datafrog.tools

  10. 10

    Extract structured data from text, files and archives.

    Mar 2026 · apps.apple.com

  11. 11UJ

    Hello HN! I became frustrated with the unpredictible/poor match quality and opaqueness of "relevance scores" in existing fuzzy and fulltext search libs, so I tried something different and this is the result. The main selling point is the result quality / ordering, with best-in-class memory overhead and excellent performance being bonuses. The API is pretty stable at this point, but looking for feedback before committing to 1.0. TL;DR The test corpus is a 4MB json file with 162k words/phrases, so give it a second for initial download. You can also drag/drop your own…

    2022 · github.com

  12. 12

    Import messy CSV/Excel to databases with auto map & validate

    2022

  13. 13AF
  14. 14

    Automatically clean customer data with a few clicks

    2020

  15. 15GS

    I created an add-on for Google Sheets called Flookup, and it comes both as a free version and a VERY AFFORDABLE paid version. At its core, Flookup is a fuzzy matching add-on that helps you manage text that is less than a 100% match. Beyond that it can be used to: 1. Search for and match data regardless of whether it contains typos. 2. Highlight and delete duplicates duplicates even if the data has mismatched text. 3. Calculate the percentage similarity between strings. 4. Extract unique values from any column based on percentage similarity. 5. Sum and find the average of numbers based on…

    2020

  16. 16

    Smarter spreadsheets, smoother workflows.

    2019

  17. 17

    Compare, merge and split messy files in seconds.

    Jun 2026 · messymatch.com

  18. 18

    Clean messy CSV, Excel, JSON with one Python script

    2025

  19. 19AW
  20. 20

    Build Better Spreadsheets with Python

    2014

  21. 21CG

    I built CSV GB+ by Data.olllo, a local data tool that lets you open, clean, and export gigabyte-sized CSVs (even billions of rows) without writing code. Most spreadsheet apps choke on big files. Coding in pandas or Polars works—but not everyone wants to write scripts just to filter or merge CSVs. CSV GB+ gives you a fast, point-and-click interface built on dual backends (memory-optimized or disk-backed) so you can process huge datasets offline. Key Features: Handles massive CSVs with ease — merge, split, dedup, filter, batch export Smart engine switch: disk-based "V Core" or RAM-based "P…

    2025 · apps.microsoft.com

  22. 22

    Instant data extraction from any file with AI

    2025

  23. 23FE

    Hey everyone, I have updated my fuzzy search library for the frontend. It now supports substring and prefix search, on top of fuzzy matching. It's fast, accurate, multilingual and has zero dependencies. GitHub: https://github.com/m31coding/fuzzy-search Live demo: https://www.m31coding.com/fuzzy-search-demo.html I would love to hear your feedback and any suggestions you may have for improving the library. Happy coding!

    Oct 2025 · github.com

  24. 24DB

    I've been doing some data cleaning for my fine tuning projects using LLMs, and decided to just build a package for it as a side project. Check it out here: https://github.com/databonsai/databonsai Some features: - categorization (labelling), transformation and decomposition (text into structured format) - validates llm outputs - batch mode batches up the inputs/outputs so you don't send the prompt (schema, fewshot examples) for every row of data, saving a significant amount of tokens There are some similarities to the Instructor repo, but this is simpler and made for…

    2024 · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →