nowfound

Alternatives

Products that do what Distributed Rate Limiter – 50K+ RPS, Redis-Backed, Production-Ready does

Hi HN! I built a distributed rate limiter using the token bucket algorithm with Redis backing. Key highlights: • 50,000+ requests&#x2F;second throughput with <2ms P95 latency • Redis-backed distributed state for multi-instance deployments • 18 REST API endpoints for rate limiting, config, and monitoring • 265+ tests including load tests and integration tests • Docker&#x2F;Kubernetes ready with comprehensive documentation Built with Java 21 + Spring Boot. Perfect for protecting APIs, microservices, or SaaS platforms from abuse. The pain point I solved: existing solutions were either too…

  1. 1
    Rately175

    Take control of your API traffic with custom rate limits.

    Oct 2025

  2. 2GR
  3. 3RC
  4. 4KA
  5. 5

    OAuth 2.1 + usage-based Stripe billing for MCP servers

    Jul 2026 · mcp-billing.com

  6. 6CE

    Dear startups, We've recently launched our service that helps deploy capacity across AWS EC2 spot to provide high availability and up to 90% discount on compute. Magic happens by us predicting spot prices and distributing capacity across dozens of uncorrelated Spot auctions to limit exposure to possible price spikes. As serial entrepreneurs we know first hand how important it is to iterate quickly and reducing the compute cost is one of the ways to approach it. If your company's AWS bill is significant (>$10k&#x2F;month) and you have an interesting use case that we can discuss - we'd be…

    2015

  7. 7KK
  8. 8IB

    Hey HN: I'm Kaveh, the founder of Usage (https:&#x2F;&#x2F;www.usage.ai&#x2F;) We help companies drive down AWS costs. Why? Because the way it's done now is a pain. Stakeholders, especially engineers, are required to spend unnecessary time manually finding underutilized or overly expensive EC2s. We believe the optimization process should be done automatically through a series of sophisticated algorithms. At the moment, there are over 70,000 AWS EC2 prices - doing that manually just won't scale at most organizations. My background is in software engineering. Previous to founding Usage, I…

    2020

  9. 9LS
  10. 10HM
  11. 11RA
  12. 12CA
  13. 13CA

    CAST AI (https:&#x2F;&#x2F;cast.ai) has built a cloud optimization platform that reduces AWS cloud costs 50% to 90%, optimizes DevOps, and automates disaster recovery via multi-cloud with a single cluster. Intelligent optimization engine delivers a cost-efficient, high-performing, and resilient infrastructure for every Kubernetes workload. If it sounds too good to be true - try free. To make it easier CAST AI provides AWS and GCP cloud credentials for free. Visit https:&#x2F;&#x2F;cast.ai Here's how it works: 1.Use CAST AI to deploy your K8s clusters. Next, take a look at the CAST AI…

    2021

  14. 14DM
  15. 15KO

    2020 · technicallywizardry.com

  16. 16GR
  17. 17IM
  18. 18WB

    Over the past few months, as we scaled our internal AI Agents, we hit a dead end: Running LLM-generated arbitrary code in Docker is basically running naked on security due to container escape risks. But using full traditional VMs takes minutes to boot and eats too much memory to support high-density concurrency. We loved the developer experience of SaaS sandboxes on the market, but they are closed-source, expensive, and have too high a barrier to entry for self-hosting. So, our team decided to build our own. After months of grinding, using RustVMM and KVM, we built a blazing-fast,…

    Apr 2026 · github.com

  19. 19CL

    Hey HN, we’re the developers of OpenLake, an open source storage engine for offloading LLM KV caches from GPU memory into a shared tier of RAM and NVMe. We built OpenLake because KV caches are outgrowing GPU memory. A single 256K token conversation on Gemma 4 31B produces approximately 43GB of KV state, more than half the memory of an 80GB H100. The problem becomes even harder across a cluster: a prefix cached on one GPU host is unavailable when the next request lands on a different GPU, forcing the new GPU to repeat work the fleet has already completed. Once the KV cache is offloaded,…

    Jul 2026 · github.com

  20. 20SC

    2015 · github.com

  21. 21AS

    Hey Hacker News! No one likes to code pricing; it sucks. It shouldn't take up my time, yet it always does - too much of it. And pricing always needs to change, eating up valuable dev hours. I created a simple pricing engine so that I could write all of my pricing rules & resource limits in a single YAML doc, then enforce them everywhere with a single policy check. It's simple, intuitive, versionable, auditable, and easy to reason about or change quickly without bogging my development down. My CTO friends liked the idea and wanted to use it, so I created this open-source library for everyone…

    Feb 2026 · github.com

  22. 22EA
  23. 23RT

    rx is a lightweight CLI and runtime that lets you deploy any function instantly as you save it in your editor. It’s fast, simple, and takes care of setting up all your cloud infrastructure. The best dev experience with everything you need built-in: - auth - permissions - payments - secrets - storage - logs - email & SMS - and way more Related beta launch tweet: https:&#x2F;&#x2F;x.com&#x2F;KhalidZoabi&#x2F;status&#x2F;1777673531972608249 Wanted to show it to you guys, gauge interest, and get some feedback. There's a large roadmap of upcoming stuff. Feel free to DM me on X for more information!

    2024 · rx.run

  24. 24KH

Ranked by how close each launch is in meaning, then by votes. Refine with a description →