Alternatives
Products that do what Refinery – Data anonymization and governance for the masses does
https://www.openquery.io/refinery We are Christos, Damien and Nodar, founders of OpenQuery. We are building Refinery, an Open Source deployment written in Rust to automate the process of anonymizing sensitive data and governing how different consumers access that data using our 'policy-as-code' framework (It's based on Hashicorp's HCL). We noticed that companies have been solving the same problem over and over in different ways. The problem boils down to giving a data-consumer (this could be an analyst, a BI tool, an ML workflow, a different company etc. ) access to a…
- 1

- 2

- 3

- 4

- 5

- 6

- 7

- 8

- 9AE
2020 · github.com
- 10CE
Hello Hacker News! We're a team of YC founders (Meldium W13, Draft S11, TapEngage S11) launching something new (https://www.getcensus.com). How many times has your business team asked you to generate yet another CSV file, write a ”quick report” in SQL, or send some custom data to a terrible API (looking at you Marketo)? We’ve built a product that connects directly to your data warehouse and syncs into apps like Salesforce, Customer.io and even Google Sheets. In fact, your business teams won’t even need to rely on engineering to manage all these pipelines. The tech stack for…
2020
- 11SH
Needed this in my own work, anonymizing PII/PHI and decided to build this because presidio didn't really cut it for our use-case. Try it and maybe let me know if you have any feedback :)
2025 · github.com
- 12CW
Hello HN, Lucas here. I’ve been working with BigQuery for ~5 years, mostly in large (petabyte-scale) environments. Over time we ended up spending a lot of money and engineering effort just trying to understand where costs were coming from, why and how to optimize them. At some point we decided to stop, leverage all our past experience and spend a full cycle building tooling focused on cost visibility and optimization. The main goal was to regain ownership of cost data and make it possible to understand our cost structure in under a minute, while aligning the views of engineering and FinOps…
Jan 2026 · cloudclerk.ai
- 13FO
Hey HN, I’m Roi, one of the co-creators of FalkorDB. We’re a growing team working on a graph database designed for production workloads and GraphRAG systems. The new release (v4.10.0) is out, and I wanted to share some of the updates and ask for feedback from folks who care about performance, memory efficiency in graph-heavy systems. FalkorDB is an open-source property graph database that supports OpenCypher (with our own extensions) and is used under the hood for retrieval-augmented generation setups where accuracy matters. The big problem we’re working on is scaling graph databases without…
2025
- 14VS
Hi HN, I wanted to share an exciting new open-source project: "VulcanSQL"! If you're interested in seamlessly transitioning your operational and analytical use cases from data warehouses and databases to the edge API server, this open-source data API framework might be just what you're looking for. VulcanSQL (https://vulcansql.com/) is suitable for following use cases: * Customer-facing analytics - expose analytics in your SaaS product for customers to understand how the product is performing for them via customer dashboards, insights, and reports. * Data Sharing - sharing…
2023 · vulcansql.com
- 15ZO
Hello HN, I am Sonal, a data consultant from India. For the past few months(and years!), I have been working on an entity resolution tool to build a single source of truth for customers, suppliers, products and parts. Here is a short demo of Zingg in action https://www.youtube.com/watch?v=zOabyZxN9b0 As a data consultant, I often struggled to build unified views of core entities on the datalake and the warehouse. Data spread across different systems has variations and consistencies making Customer 360, KYC, AML, segmentation, personalization and other analytics difficult. As I…
2022
- 16PO
We built a data feed oracle to write prices on-chain, and it’s live on the Ropsten test network! In Decentralized Finance (Defi), you can build your own financial instruments and you often need an oracle to write off-chain prices onto the blockchain. We built an oracle focused on data feeds because the current options are too expensive and slow for regularly, repeating data. This oracle is designed for price feeds and uses public key encryption to validate the data submitted is accurate. We modeled our oracle off of MakerDao and generalized the design to work with any datafeed. We’re…
2019
- 17

rustaceans, we've completely open sourced our APH rust engine and servers! APH is a Agent Per Human notarization protocol, that acts like remote guardrails for when your agents are working with agents you don't own. It allows both parties to share public keys that they can verify each other are working in good faith against public servers. APH by Squillo is the missing piece of the A2A and AP2 families from Google, built with the same pedigree as the web but for the agentic web. It's designed to work with existing agentic protocols and future ones, but is the cornerstone for the agentic web…
23d ago · github.com
- 18LW
2024 · david-delassus.medium.com
- 19AU
Hey! My name is Reda, and I run a small company (aries.com) we’re working on a complete development engine for finance and capital markets. You no longer need to spend millions on market data, licensing, and infra to run a consumer fintech MVP. Large banks and institutions can afford to pay past regulatory & infra barriers but in our time interviewing developers and entrepreneurs we found that not only does Fintech have one of the highest startup failure rates but most of those failures can boil down to regulation, data fees, and infra hurdles that fintech entrepreneurs need to jump over.…
2025 · twitter.com
- 20BA
If we want better web3 experiences, developers need better tools. RPC nodes are really good at executing transactions, however they are notoriously cumbersome to set up, and reading large chunks of data is not very efficient: To show a list of transactions and receipts, nodes have to re-execute smart contract code on entire blocks. For every read call. Not great at scale. Which is why everyone is building ETLs to move data from the chain into their own database. This GraphQL API is our first step in allowing developers to spend more time on building product, rather than ETL infrastructure.
2022 · basement.dev
- 21SP
Hey everyone, we built a local-first, open-source framework to allows you to export and build applications with your personal data. This August, we built a digital footprint exporter (https://news.ycombinator.com/item?id=41325719) that allowed people to get back their personal data, but there was no way to interact with it. So we added a python package, improved the reliability of the exporting, and implemented a unified JSON structure across platforms. How it works: - Electron’s built-in webview element is used to navigate to the platform’s website and the user can connect…
2024 · docs.surferprotocol.org
- 22AA
Hey HN! I really like local apps for their simplicity and privacy and hate paying Saas bills and I wanted a way to start automating my life with AI so I started building Anything. Anything is built on Tauri so the front end is React and the "backend" is Rust. It's 100% local & 100% doesn't ask you to spin up docker to use. Another core goal of the app is to get away from "package bloat" you see in other general purpose AI oss projects where they have a package.json that is 300 lines long ( more on that later. ) Oh btw I suck at Rust! I learned Rust while building this so the code is _not…
2024 · github.com
- 237F
OLake is our open-source tool for ingesting Database & Kafka data into Apache Iceberg. We recently redesigned the write pipeline and saw ~7x throughput improvements. Sharing the architecture decisions, trade-offs, and benchmarks.
Dec 2025 · olake.io
- 24OH
I'm Fenil, co-founder/CEO of OpenFunnel (YC F24), building this with my co-founder/CTO Aditya. We're launching OpenBenchmarks (https://openbenchmarks.com), open-source, reproducible benchmarks for SaaS APIs, starting with the category we know best: GTM APIs. ## Why we built this More and more B2B software evaluation will/already runs through reasoning models inside agentic workflows rather than through people. And buyers increasingly pick vendors that are API-first and ship MCPs, so they can wire them into internal workflows. Strong reasoning models are skeptical of…
Jul 2026 · openbenchmarks.com
Ranked by how close each launch is in meaning, then by votes. Refine with a description →