nowfound

Alternatives

Products that do what NeuroBrix does

One Runtime. Any Model. Any Hardware.

  1. 1

    No-code AI Lab: Train models, access datasets, run inference

    Feb 2026

  2. 2NG
  3. 3OS

    Hi HN, I built a specialized inference engine for running 4-bit Gemma 4 26B-A4B-IT on any M-series Mac using about 2 GB of RAM. It is called TurboFieldfare and is written in Swift and Metal. I have always adored on-device AI. It feels like magic that you can run a powerful NN on your Mac or iPhone. So I wanted to push the limits a bit and run a model whose weights don’t fit in memory. The model’s 4-bit quantized weights occupy roughly 14 GB, which makes running it with conventional inference tools almost impossible on an 8 GB or even 16 GB Mac once the OS, applications, and KV cache are…

    Jul 2026 · github.com

  4. 4
    Nexa SDK130

    Run, build & ship local AI in minutes

    Sep 2025

  5. 5

    AI models that run on an inference cloud optimized for speed

    May 2026 · generalcompute.com

  6. 6

    Fast multimodal-native inference at scale

    Dec 2025

  7. 7
    Neuro104

    Instant infrastructure for machine learning

    2021

  8. 8

    Tighter instruction adherence in speech agents

    Feb 2026 · developers.openai.com

  9. 9
    Neuron187

    A personal AI in every home.

    2024

  10. 10
    Neuton.AI271

    No-code artificial intelligence for all

    2021

  11. 11

    Spin up secure sandboxes in ~100 ms

    Nov 2025

  12. 12

    Turn idle GPUs into cash. Get affordable AI for everyone.

    Nov 2025

  13. 13

    AI that executes UI actions on your computer in ~285ms

    Jun 2026 · getneuralagent.com

  14. 14

    All-in-one AI workspace: content, code, design & API power

    2025

  15. 15

    Local, gradient-free neuro-symbolic memory engine combining Hyperdimensional Computing (HDC/VSA), Hebbian plasticity, and graph triples for offline AI. - roandejager/Hillock

    7d ago · github.com

  16. 16IR

    The Emotion Engine has 32 MB of RAM total, so the trick is streaming weights from CD-ROM one matrix at a time during the forward pass — only activations, KV cache and embeddings live in RAM. This means models bigger than the RAM can still run, they just read more from disc. Had to build a custom quantized format (PSNT), hack endianness, write a tokenizer pipeline, and most of the PS2 SDK from scratch (releasing that separately). The model itself is also custom — a 10M param Llama-style architecture I trained specifically for this. And it works. On real hardware.

    Mar 2026 · github.com

  17. 17NL

    Built this because I was tired of every AI tool shipping my data to someone else server n0x runs the full stack LLM inference via WebGPU, autonomous ReAct agents, RAG over your own docs, sandboxed Python execution via Pyodide all inside a single browser tab. No account No keys No backend Models download once, cache in IndexedDB permanently. Biggest challenge was context window budgeting for the agent loop and making the WASM vector search non-blocking. Happy to talk architecture. GitHub: https://github.com/ixchio/n0x | Live demo: https://n0x-three.vercel.app

    Mar 2026 · n0xth.vercel.app

  18. 18

    The first open compiler for spiking neuromorphic compute.

    Jul 2026 · thrindex.com

  19. 19

    A sensor board that gets your robot running in minutes

    18d ago · nxs.aliensense.com

  20. 20

    Artifex is a machine-first, headless CLI runtime built for autonomous coding agents to author, validate, and render media node graphs locally. The agent talks to Artifex through a structured CLI interface. Workflows are DAGs, and each node is a plugin that can implement its own execution logic.. Each node has capability to inject logic into graph processing, WebGPU rendering, audio processing and their own SKILL.md file. Nodes can also inject their react components (not available with CLI) - which will be available with the desktop app. Execution is topological and supports checkpoint…

    23d ago · gatewai.studio

  21. 21NT

    Hello HackerNews! I’m excited to share what we’ve been working on at nCompass Technologies: an AI inference* platform that gives you a scalable and reliable API to access any open-source AI model — with no rate limits. We don't have rate limits as optimizations we made to our AI model serving software enable us to support a high number of concurrent requests without degrading quality of service for you as a user. If you’re thinking, well aren’t there a bunch of these already? So were we when we started nCompass. When using other APIs, we found that they weren’t reliable enough to be able to…

    2024 · ncompass.tech

  22. 22
    Brain8

    A small, blazingly fast and extensible agent runtime

    4d ago · github.com

  23. 23

    Secure hosting for whatever your AI builds

    Jul 2026 · nexalibre.com

  24. 24

    Keep your OpenClaw agents running. Free beta, no code change

    Apr 2026 · openinfer.io

Ranked by how close each launch is in meaning, then by votes. Refine with a description →