nowfound

Alternatives

Products that do what Exla FLOPs does

On-Demand GPU clusters - The Cheapest H100s Anywhere

  1. 1

    Self-host AI/ML with the world's cheapest GPU cloud

    2025

  2. 2
    GPU.LAND126

    Affordable cloud GPUs for deep learning

    2021

  3. 3AG

    We built, saga[1] a layer that lets smaller teams access enterprise GPU discounts through collective buying power. How it works: 1. Aggregate GPU spend across hundreds of ML teams 2. Get enterprise rates through combined volume 3. Pass savings to users, monetize via provider partnerships Technical notes: - Works at billing layer only (no access to code/data) - Supports existing cloud setups or managed GPUs - Private beta running since January, opening more spots for March - Currently seeing ~50% savings on H100s/A100s [1] https://trysaga.ai

    2025 · trysaga.ai

  4. 4
    crunr 106

    Launch and run any compute job on AWS with 1 command

    May 2026

  5. 5

    Calculate the GPU memory you need for LLM inference

    2025

  6. 6
    RunInfra156

    Describe the AI model you need and get an optimized AI

    Jul 2026

  7. 7

    Spin up secure sandboxes in ~100 ms

    Nov 2025

  8. 8GP

    Out of curiosity, I put together a simple website which tracks the prices for a few variations of A100/H100 GPUs by hour broken out between spot/ondemand, form factor and provider. Specifically I was tailoring the tool towards the smaller, emerging providers like runpod, gpulist.ai, lambda labs etc. Anyone have any ideas to expand/refine it?

    2024 · computeindex.michaelgiba.com

  9. 9FP

    2021 · floptimal.com

  10. 10AS
  11. 11

    Live cloud GPU price comparison

    26d ago · gpuhour.com

  12. 12CE

    I've made a GPU comparison site: https://gpu-prices.com/US/ Yes, there was one yesterday - I got beaten to it. This one is different in that I've put a lot of work into categorisation, so you can filter down quite precisely. For example, VRAM-per-dollar, limiting to Nvidia, and a minimum performance score allow you to find good ML GPUs. It currently supports Australia, Canada, Ireland, the UK, and the US. The tech stack is Python for data pull and static site generation then Cloudflare Pages for actually serving the site. It updates three times a day, but I could increase…

    2024 · gpu-prices.com

  13. 13PI
  14. 14UA

    Hi HN! I've got the barebones of a service running on top of Stable Diffusion XL. I can cheaply run image generations at 1024x1024. And of course there's a limit to how fast I can generate them given the request queue and limited GPUs, but the service is cheap enough that I'm happy to run it out of pocket for now. Let me know your thoughts, I hope you enjoy the service!

    2023 · unstock.ai

  15. 15SY

    Hey HN, If you tried running open-source models like Llama 3.1 70B or 405B, you might have noticed that it gets very expensive. The reason looks obvious enough that you might have stopped even before trying it! - GPUs are very expensive to buy or rent - Running the most performing LLMs need 4, 8 or even 16 top of the line Nvidia GPUs - And that won’t get you anywhere near the level of VRAM needed to batch enough to get a decent throughput and efficiency Some have even questioned if open-source LLM providers are not doing some shenanigans to provide the prices they offer. VC funded…

    2024

  16. 165L

    We've built InferX, a specialized runtime environment that fundamentally changes how LLMs are served. The core problem we solve is the latency bottleneck in AI inference, especially with large models. Current systems waste resources or suffer from painfully slow cold starts. InferX's AI-native architecture, with its "snapshot" technology, enables: * *Sub-2s cold starts:* Spin up models instantly. * *High density:* Serve more LLMs on the same GPUs. * *Optimal efficiency:* Maximize GPU utilization. This isn't just another API; it's a new execution layer designed from the ground up for the…

    2025 · github.com

  17. 17

    Turn idle GPUs into cash. Get affordable AI for everyone.

    Nov 2025

  18. 18

    Cheapest Realtime Video Generation with Minimax FastH3

    5d ago · highlander.sh

  19. 19RA

    Hi there, looking for feedback on my new project "Featherless.AI" The idea is to allow users to run all the models on hugging face instantly. Via the OpenAI API compatible endpoint. Why? Because its a real chore to download models and spin up GPUs, especially if you want to test multiple models. Not to mention GPUs cost multiple dollars an hour to rent. And if we want more people to use open source AI, we got to make it easier for them to try and play with all of them. So what if instead of spinning up dedicated GPUs per model (which is what every provider is doing) We can startup a LLM…

    2024 · featherless.ai

  20. 20SS

    We'd like to introduce HN to Spell, which is a tool for easily running ML/DL jobs remotely. As Deep Learning has grown we see engineers and researchers struggle to incorporate running on GPUs into their workflow. So we built Spell to be the easiest way to get code running elsewhere - like the bash '&' operator but for remote machines. Sign up for an account at https://web.spell.run/waitlist, which includes $300 in credits for GPU time. There's a waitlist, but we'll be approving accounts as they come in. Here are some of the features we really wanted and built into Spell:…

    2018

  21. 21RS

    I built a small weekend project to help track RAM prices, since DDR3/DDR4/DDR5 costs have suddenly jumped recently and I was struggling to find good deals for my NAS build. RamScout scans eBay (UK/US) and ranks RAM listings by price per GB, with filters for type, capacity, speed, condition, etc. It’s a simple MVP — no frills, no accounts, no ads — just a fast way to spot unusually cheap listings. Would appreciate any feedback, especially on performance, UI, and whether expanding to more regions/vendors would be useful. Thanks!

    Dec 2025 · ramscout.com

  22. 22

    Hi everyone, Please checkout compute.cx which is a simple cli interface for using on demand GPUs from RunPod and HotAisle. I created this because I really like the ease of modal.com for severless gpu access, but don’t always want to pay their markup. Compute.cx gives the same DX but on public on-demand GPUs like runpod and hotaisie. Please try it out, and write to me [email protected] for any questions/suggestions, or file a bug report on https://github.com/theoriclabs/docs.compute.cx Thanks! Harsh Gupta https://x.com/hargup13 P.S. BYOK AWS, GCP and…

    16d ago · compute.cx

  23. 23GP

    GPU rental prices are super volatile but there's no derivatives market to hedge. I built a perpetual futures platform to see what this could look like. The idea is airlines hedge jet fuel, starbucks hedges coffee beans - as GPU compute becomes critical infrastructure the same hedging tools should exist. Not sure if anyone actually needs this but it was interesting to build. How it works: - Pulls live H200 spot prices from Vast.ai every 15s into a tradeable index - Full perp mechanics: funding rates, mark price calc, real-time P&L - Event-driven Rust backend with supervisor pattern and…

    Feb 2026 · github.com

  24. 24ZA

Ranked by how close each launch is in meaning, then by votes. Refine with a description →