nowfound

Alternatives

Products that do what RunInfra does

Describe the AI model you need and get an optimized AI

  1. 1
    Hathora148

    Explore, test & deploy production ready voice models.

    Nov 2025

  2. 2
    Forge CLI107

    Swarm agents optimize CUDA/Triton for any HF/PyTorch model

    Jan 2026

  3. 3
    OpenWispr190

    100% local open source AI speech-to-text model

    2025

  4. 4

    The easiest way to use cloud GPUs

    2025

  5. 5

    Powerful Conversational AI Agent Builder Platform

    Nov 2025

  6. 6
    Arkor142

    Fine-tune and Deploy Open-weight Models in TypeScript

    Jul 2026 · arkor.ai

  7. 7

    Swarm Agents That Turn Slow PyTorch Into Fast GPU Kernels

    Jan 2026

  8. 8
    GPU.LAND126

    Affordable cloud GPUs for deep learning

    2021

  9. 9
    FineTuner164

    Fine-tune AI models on your data — in minutes, not days.

    2025

  10. 10

    No-code studio for custom AI building and running

    Oct 2025

  11. 11RA

    Hi there, looking for feedback on my new project "Featherless.AI" The idea is to allow users to run all the models on hugging face instantly. Via the OpenAI API compatible endpoint. Why? Because its a real chore to download models and spin up GPUs, especially if you want to test multiple models. Not to mention GPUs cost multiple dollars an hour to rent. And if we want more people to use open source AI, we got to make it easier for them to try and play with all of them. So what if instead of spinning up dedicated GPUs per model (which is what every provider is doing) We can startup a LLM…

    2024 · featherless.ai

  12. 12

    Know what your AI will actually cost to run.

    30d ago · howmuchtorunai.com

  13. 13

    Turn idle GPUs into cash. Get affordable AI for everyone.

    Nov 2025

  14. 14RA

    We built RapidFire AI, an open-source Python tool to speed up LLM fine-tuning and post-training with a powerful level of control not found in most tools: Stop, resume, clone-modify and warm-start configs on the fly—so you can branch experiments while they’re running instead of starting from scratch or running one after another. - Works within your OSS stack: PyTorch, HuggingFace TRL/PEFT), MLflow. - Hyperparallel search: launch as many configs as you want together, even on a single GPU - Dynamic real-time control: stop laggards, resume them later to revisit, branch promising configs in…

    Sep 2025 · github.com

  15. 15SS

    We'd like to introduce HN to Spell, which is a tool for easily running ML/DL jobs remotely. As Deep Learning has grown we see engineers and researchers struggle to incorporate running on GPUs into their workflow. So we built Spell to be the easiest way to get code running elsewhere - like the bash '&' operator but for remote machines. Sign up for an account at https://web.spell.run/waitlist, which includes $300 in credits for GPU time. There's a waitlist, but we'll be approving accounts as they come in. Here are some of the features we really wanted and built into Spell:…

    2018

  16. 165L

    We've built InferX, a specialized runtime environment that fundamentally changes how LLMs are served. The core problem we solve is the latency bottleneck in AI inference, especially with large models. Current systems waste resources or suffer from painfully slow cold starts. InferX's AI-native architecture, with its "snapshot" technology, enables: * *Sub-2s cold starts:* Spin up models instantly. * *High density:* Serve more LLMs on the same GPUs. * *Optimal efficiency:* Maximize GPU utilization. This isn't just another API; it's a new execution layer designed from the ground up for the…

    2025 · github.com

  17. 17OA

    Built this after getting tired of fighting local AI setup (CUDA issues, dependencies, API configs). Goal was to make something that just runs locally without all the overhead. Happy to answer questions or get feedback.

    Apr 2026 · store.steampowered.com

  18. 18IM

    I wanted to explore different open source Gen AI tools and brought them together to generate this youtube video: voice, music, image and text. All processed on a single RTX3090. The text for the video is generated using the wikipedia article on Octopus as a source.

    2024 · youtube.com

  19. 19FA

    Hi HN, We're excited to introduce Fixstars AIBooster, our new performance engineering tool designed to significantly accelerate AI model training while optimizing GPU utilization. AIBooster provides: Real-time monitoring of GPU, CPU, memory, and power consumption. Clear visibility into performance bottlenecks, helping developers optimize AI workloads. Proven acceleration of AI training processes—users commonly achieve up to 2-3x speed improvements. Significant cost savings by maximizing infrastructure efficiency. It's free to try, requires minimal setup, and integrates seamlessly into your…

    2025 · fixstars.com

  20. 20S1

    I wanted to build an inference provider for proprietary AI models, but I did not have a huge GPU farm. I started experimenting with Serverless AI inference, but found out that coldstarts were huge. I went deep into the research and put together an engine that loads large models from SSD to VRAM up to ten times faster than alternatives. It works with vLLM, and transformers, and more coming soon. With this project you can hot-swap entire large models (32B) on demand. Its great for: Serverless AI Inference Robotics On Prem deployments Local Agents And Its open source. Let me know if anyone…

    Nov 2025 · github.com

  21. 21DI

    Hi HN community, Shen and I created a service for anyone to easily train deep learning model on GPU power harnessed from the crowd. We have completed the first version DeepCluster.io (http://deepcluster.io) and welcome ML researchers to try it out for free! We are enthusiastic of deep learning, but often found training models with GPU instances on AWS very expensive. Meanwhile, some of our friends have idle GPUs that are used to mine cryptos. So we decided to borrow their GPUs for training deep learning model ourselves, and believe this could be a service that benefits other ML…

    2019

  22. 22WB

    Hey HN, After GPT-3 created waves in the tech industry, a lot of AI tools were emerging and with that, some AI website builders But the results seemed way too generic to us. It felt like the developers were rushing to catch the wave instead of building a proper tool We took our time, did months of RnD and finally came up with something better than what others in the market are doing. It’s got better design output. While it’s still in beta, I wanted to show HN what we did. Will appreciate the feedback when you guys try it out. Here is the link to signup for the beta:…

    2024 · dorik.com

  23. 23OS

    We just open sourced a first class our AI web app builder. Instead of using another hosted AI coding platform, you can fork this project and build your own AI app builder, fully customized and running under your own brand. It includes: Next.js + TypeScript AI chat with streaming Artifact generation File explorer Code editor Live preview Databases Sandboxes to be used by AI agents Versions Responsive production-ready UI And a lot more... The only required dependency is the Totalum API, which exposes the AI generation engine through a simple REST API. You can replace or extend the backend…

    Jul 2026 · github.com

  24. 24
    Brain8

    A small, blazingly fast and extensible agent runtime

    3d ago · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →