AI · alternatives · 2026
24 alternatives to Janus – Go binary that runs GGUF models via Vulkan on AMD/Intel/Nvidia
Janus is a API router for AI models written in Go and has a Vulkan Model runner - Vibra-Ingenn/Janus
Janus is a Go-based API router for AI models that includes a Vulkan model runner capable of executing GGUF models across AMD, Intel, and Nvidia GPUs. It enables developers to route requests to AI models while… Below are 24 products that do a similar job, ranked by how close each is in meaning and then by launch-day votes.
- 1

- 2

Hi everyone, I'm kinda involved in some retrogaming and with some experiments I ran into the following question: "It would be possible to run transformer models bypassing the cpu/ram, connecting the gpu to the nvme?" This is the result of that question itself and some weekend vibecoding (it has the linked library repository in the readme as well), it seems to work, even on consumer gpus, it should work better on professional ones tho
Feb 2026 · github.com · its alternatives →
- 3

Insights on GPU and AI infrastructure from Arc Compute. Read hardware deep-dives, deployment playbooks, and analysis of where the AI compute market is heading.
2021 · arccompute.com · its alternatives →
- 4
General Compute▲315AI models that run on an inference cloud optimized for speed
May 2026 · generalcompute.com · its alternatives →
- 5

CUDA is NVIDIA's language for GPU programming, allowing you to mix write CPU and GPU code in C++ in one file. By chaining a few projects that compile CUDA to OpenCL, then Vulkan, then WebGPU, you can experiment with this GPGPU language on any hardware.
2025 · hipscript.lights0123.com · its alternatives →
- 6

- 7

[Experimental] Deploy and stream games/apps in the cloud. Use our GPUs or bring your own. - nestrilabs/nestri
2024 · github.com · its alternatives →
- 8

We developed a tool to trick your computer into thinking it’s attached to a GPU which actually sits across a network. This allows you to switch the number or type of GPUs you’re using with a single command.
2024 · thundercompute.com · its alternatives →
- 9
RunInfra▲156Describe the AI model you need and get an optimized AI
Jul 2026 · runinfra.ai · its alternatives →
- 10
Cerebrium▲300If you haven’t noticed, GPUs are scarce right now. The global boom in inference has demand growing faster than supply can keep up, and at Cerebrium ...
2024 · cerebrium.ai · its alternatives →
- 11

A path tracer written in Go. Contribute to fogleman/pt development by creating an account on GitHub.
2015 · github.com · its alternatives →
- 12PO
Our company Vertex.AI has been working on this for a while but this is the first public release. We're starting with using PlaidML to bring OpenCL support to Keras and more frameworks, platforms, etc are coming. Yes, this means you can use use your AMD GPU for deep learning dev. Sorry, no Mac or Windows support yet although the brave can try building from source (it should work). http://vertex.ai/blog/announcing-plaidml https://github.com/plaidml/plaidml
2017 · its alternatives →
- 13

We built installable software for Windows & Linux that makes any remote Nvidia GPU accessible to, and shareable across, any number of remote clients running local applications, all over standard networking.
2022 · github.com · its alternatives →
- 14

Back in the old days, people used to do general-purpose GPU programming by using shaders like GLSL. This is what inspired NVIDIA (and other companies) to eventually create CUDA (and friends). This is an implementation of GPT-2 using WebGL and shaders. Enjoy!
2025 · github.com · its alternatives →
- 15

- 16

Powers faster, efficient reasoning for long-running agents
Jun 2026 · developer.nvidia.com · its alternatives →
- 17

We ported pbrt-v4 to Julia and built it into a Makie backend. Any Makie plot can now be rendered with physically-based path tracing. Julia compiles user-defined physics directly into GPU kernels, so anyone can extend the ray tracer with new materials and media - a black hole with gravitational lensing is ~200 lines of Julia. Runs on AMD, NVIDIA, and CPU via KernelAbstractions.jl, with Metal coming soon. Demo scenes: github.com/SimonDanisch/RayDemo
Feb 2026 · makie.org · its alternatives →
- 18

Minimal Example of Using Vulkan for Compute Operations. Only ~400LOC. - Erkaman/vulkan_minimal_compute
2017 · github.com · its alternatives →
- 19

GPU Raytracer from scratch in C++/CUDA. Contribute to jan-van-bergen/GPU-Raytracer development by creating an account on GitHub.
2021 · github.com · its alternatives →
- 20

This started out as a personal effort to learn more about machine learning. It's currently a CLI app where you give it a JSON file specifying your network architecture and hyperparameters and point it to your training data, then invoke it again in 'eval' mode with some data it's not seen before and it will try to classify each sample. I don't see many other people using Vulkan for GPGPU, and there may be many good reasons for that, but I wanted to try something a bit different. I've made every attempt to make the code very clean and readable and I've written up the math in…
2024 · github.com · its alternatives →
- 21
Ocean Orchestrator ▲139Run AI jobs from your IDE with a one-click workflow
Mar 2026 · oncompute.ai · its alternatives →
- 22WM
Try it out! https://glhf.chat/ Hey HN! We’ve been working for the past few months on a website to let you easily run (almost) any open-source LLM on autoscaling GPU clusters. It’s free for now while we figure out how to price it, but we expect to be cheaper than most GPU offerings since we can run the models multi-tenant. Unlike Together AI, Fireworks, etc, we’ll run any model that the open-source vLLM project supports: we don’t have a hardcoded list. If you want a specific model or finetune, you don’t have to ask us for it: you can just paste the Hugging Face link in and…
2024 · glhf.chat · its alternatives →
- 23

The Unreal Engine 5 Lyra sample now runs in WebGPU. Live link coming soon, stay tuned! #WebGPU #UnrealEngine5
2024 · twitter.com · its alternatives →
- 24

Demo of agent based model on GPU with CUDA and OpenGL (Windows/Linux) Agent instances on GPU memory Uses SSBO for instanced objects (with GLSL 450 shaders) CUDA OpenGL interops Renders with GLFW3 window manager Dynamic camera views in OpenGL (pan,zoom with mouse) Libraries installed using vcpkg (https://github.com/KienTTran/ABMGPU)
2023 · github.com · its alternatives →
Also compare
Ranked by how close each launch is in meaning, then by votes. Prices were read from each product’s own site when checked and can change. Refine with your own description →