Alternatives
Products that do what Distributed Rate Limiter – 50K+ RPS, Redis-Backed, Production-Ready does
Hi HN! I built a distributed rate limiter using the token bucket algorithm with Redis backing. Key highlights: • 50,000+ requests/second throughput with <2ms P95 latency • Redis-backed distributed state for multi-instance deployments • 18 REST API endpoints for rate limiting, config, and monitoring • 265+ tests including load tests and integration tests • Docker/Kubernetes ready with comprehensive documentation Built with Java 21 + Spring Boot. Perfect for protecting APIs, microservices, or SaaS platforms from abuse. The pain point I solved: existing solutions were either too…
- 1

- 2GR
2018 · github.com
- 3RC
2016 · github.com
- 4KA
2017 · github.com
- 5

- 6CE
Dear startups, We've recently launched our service that helps deploy capacity across AWS EC2 spot to provide high availability and up to 90% discount on compute. Magic happens by us predicting spot prices and distributing capacity across dozens of uncorrelated Spot auctions to limit exposure to possible price spikes. As serial entrepreneurs we know first hand how important it is to iterate quickly and reducing the compute cost is one of the ways to approach it. If your company's AWS bill is significant (>$10k/month) and you have an interesting use case that we can discuss - we'd be…
2015
- 7KK
2019 · k8up.io
- 8IB
Hey HN: I'm Kaveh, the founder of Usage (https://www.usage.ai/) We help companies drive down AWS costs. Why? Because the way it's done now is a pain. Stakeholders, especially engineers, are required to spend unnecessary time manually finding underutilized or overly expensive EC2s. We believe the optimization process should be done automatically through a series of sophisticated algorithms. At the moment, there are over 70,000 AWS EC2 prices - doing that manually just won't scale at most organizations. My background is in software engineering. Previous to founding Usage, I…
2020
- 9LS
2015 · github.com
- 10HM
2015 · hyper.sh
- 11RA
2023 · ratelimitapi.com
- 12CA
2015 · github.com
- 13CA
CAST AI (https://cast.ai) has built a cloud optimization platform that reduces AWS cloud costs 50% to 90%, optimizes DevOps, and automates disaster recovery via multi-cloud with a single cluster. Intelligent optimization engine delivers a cost-efficient, high-performing, and resilient infrastructure for every Kubernetes workload. If it sounds too good to be true - try free. To make it easier CAST AI provides AWS and GCP cloud credentials for free. Visit https://cast.ai Here's how it works: 1.Use CAST AI to deploy your K8s clusters. Next, take a look at the CAST AI…
2021
- 14DM
2015 · github.com
- 15KO
2020 · technicallywizardry.com
- 16GR
2019 · github.com
- 17IM
2020 · thiicket.com
- 18WB
Over the past few months, as we scaled our internal AI Agents, we hit a dead end: Running LLM-generated arbitrary code in Docker is basically running naked on security due to container escape risks. But using full traditional VMs takes minutes to boot and eats too much memory to support high-density concurrency. We loved the developer experience of SaaS sandboxes on the market, but they are closed-source, expensive, and have too high a barrier to entry for self-hosting. So, our team decided to build our own. After months of grinding, using RustVMM and KVM, we built a blazing-fast,…
Apr 2026 · github.com
- 19CL
Hey HN, we’re the developers of OpenLake, an open source storage engine for offloading LLM KV caches from GPU memory into a shared tier of RAM and NVMe. We built OpenLake because KV caches are outgrowing GPU memory. A single 256K token conversation on Gemma 4 31B produces approximately 43GB of KV state, more than half the memory of an 80GB H100. The problem becomes even harder across a cluster: a prefix cached on one GPU host is unavailable when the next request lands on a different GPU, forcing the new GPU to repeat work the fleet has already completed. Once the KV cache is offloaded,…
Jul 2026 · github.com
- 20SC
2015 · github.com
- 21AS
Hey Hacker News! No one likes to code pricing; it sucks. It shouldn't take up my time, yet it always does - too much of it. And pricing always needs to change, eating up valuable dev hours. I created a simple pricing engine so that I could write all of my pricing rules & resource limits in a single YAML doc, then enforce them everywhere with a single policy check. It's simple, intuitive, versionable, auditable, and easy to reason about or change quickly without bogging my development down. My CTO friends liked the idea and wanted to use it, so I created this open-source library for everyone…
Feb 2026 · github.com
- 22EA
2020 · ebbflow.io
- 23RT
rx is a lightweight CLI and runtime that lets you deploy any function instantly as you save it in your editor. It’s fast, simple, and takes care of setting up all your cloud infrastructure. The best dev experience with everything you need built-in: - auth - permissions - payments - secrets - storage - logs - email & SMS - and way more Related beta launch tweet: https://x.com/KhalidZoabi/status/1777673531972608249 Wanted to show it to you guys, gauge interest, and get some feedback. There's a large roadmap of upcoming stuff. Feel free to DM me on X for more information!
2024 · rx.run
- 24KH
2016 · kcluster.io
Ranked by how close each launch is in meaning, then by votes. Refine with a description →