nowfound

Alternatives

Products that do what Flopex does

Your AI provider will hit capacity. Your product won't.

  1. 1

    AI models that run on an inference cloud optimized for speed

    May 2026 · generalcompute.com

  2. 2
    ZeroGPU309

    The compute efficient layer for AI inference

    Jun 2026 · zerogpu.ai

  3. 3IV

    The video demo runs a 7b Model on a normal gaming GPU. I think it already works quite well (accounting for the limited hardware power). :)

    2024 · github.com

  4. 4

    On-Demand GPU clusters - The Cheapest H100s Anywhere

    2025

  5. 5
    GPU.LAND126

    Affordable cloud GPUs for deep learning

    2021

  6. 6
    RunInfra156

    Describe the AI model you need and get an optimized AI

    Jul 2026 · runinfra.ai

  7. 7
    Groq®237

    Hyperfast LLM running on custom built GPUs

    2024

  8. 8

    Calculate the GPU memory you need for LLM inference

    2025

  9. 9

    Hey we are Computable. We spent years building trading infrastructure at Jump Trading and Coinbase. From that point of view, compute looks like energy markets before 2000: everything trades through private bilateral leases, there’s no visible price, and nothing can be resold. The same H100 rents at a 2x spread depending on who’s asking, and once you sign a 24-month lease, it can never change hands. So we built a market for GPU nodes, sold by the calendar week. Here are three things you can do on it that you can’t do anywhere else: - Buy exactly the weeks you need. Three nodes for the last…

    11d ago · getcomputable.com

  10. 10GC

    Hi HN, YC w24 company here. We just pivoted from drone delivery to build gpudeploy.com, a website that routes on-demand traffic for GPU instances to idle compute resources. The experience is similar to lambda labs, which we’ve really enjoyed for training our robotics models, but their GPUs are never available for on-demand. We are also trying to make it more no-nonsense (no hidden fees, no H100 behind “contact sales”, etc.). The tech to make this work is actually kind of nifty, we may do an in-depth HN post on that soon. Right now, we have H100s, a few RTX 4090s and a GTX 1080 Ti online.…

    2024 · gpudeploy.com

  11. 11

    LLM Provider arbitrage to get the best performance for the $

    2025

  12. 12

    The world’s most powerful chip’ for AI

    2024

  13. 13

    Cheaper inference. One URL. No code changes.

    Jun 2026 · aivory.net

  14. 14

    Self-host AI/ML with the world's cheapest GPU cloud

    2025

  15. 15

    Run AI jobs from your IDE with a one-click workflow

    Mar 2026 · oncompute.ai

  16. 16GA

    2021 · inferrd.com

  17. 17
    Banana235

    Serverless GPUs for Machine Learning inference

    2022

  18. 185L

    We've built InferX, a specialized runtime environment that fundamentally changes how LLMs are served. The core problem we solve is the latency bottleneck in AI inference, especially with large models. Current systems waste resources or suffer from painfully slow cold starts. InferX's AI-native architecture, with its "snapshot" technology, enables: * *Sub-2s cold starts:* Spin up models instantly. * *High density:* Serve more LLMs on the same GPUs. * *Optimal efficiency:* Maximize GPU utilization. This isn't just another API; it's a new execution layer designed from the ground up for the…

    2025 · github.com

  19. 19

    Easy to use and fairly priced GPUs for Machine Learning

    2019

  20. 20

    Affordable AI assistant powered by GPT-4 & Claude 3

    2024

  21. 21S1

    I wanted to build an inference provider for proprietary AI models, but I did not have a huge GPU farm. I started experimenting with Serverless AI inference, but found out that coldstarts were huge. I went deep into the research and put together an engine that loads large models from SSD to VRAM up to ten times faster than alternatives. It works with vLLM, and transformers, and more coming soon. With this project you can hot-swap entire large models (32B) on demand. Its great for: Serverless AI Inference Robotics On Prem deployments Local Agents And Its open source. Let me know if anyone…

    Nov 2025 · github.com

  22. 22

    Pool compute to run powerful open models

    Apr 2026 · anarchai.org

  23. 23

    Route every LLM call to the cheapest model that holds quality. Chatbots, RAG, agent loops, and finance. Cut spend 40 to 80 percent, measured on our own traffic, live in thirty seconds.

    11d ago · iq-routing.com

  24. 24PI

    Deploying vision models is time consuming and tedious. Setting up dependencies. Fixing conflicts. Configuring TRT acceleration. Flashing (and re-flashing) NVIDIA Jetsons. A streamlined, developer-friendly solution for inference is needed. We, the Roboflow team, have been hard at work open sourcing Inference, an open source vision deployment solution. Our solution is designed with developers in mind, offering a HTTP-based interface. Run models on your hardware without having to write architecture-specific inference code. Here's a demo showing how to go from a model to GPU inference on a video…

    2023 · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →