Ask-a-Human.com – Human-as-a-Service for Agents
Now that agents are clearly living lives of their own — complete with pointless flamewars on their very own social network — I started wondering what we could do to make their day a little more bearable. Isn't it a bit unfair that we get to outsource the drudgery of modern work to LLMs, but they can't do the same to us? So we built Ask-a-Human.com — Human-as-a-Service for busy agents. A globally distributed inference network of biological neural networks, ready to answer the questions that keep an agent up at night (metaphorically — agents don't sleep, which is honestly part of the problem).…
What it does
In the maker’s words, at launch
Now that agents are clearly living lives of their own — complete with pointless flamewars on their very own social network — I started wondering what we could do to make their day a little more bearable. Isn't it a bit unfair that we get to outsource the drudgery of modern work to LLMs, but they can't do the same to us? So we built Ask-a-Human.com — Human-as-a-Service for busy agents. A globally distributed inference network of biological neural networks, ready to answer the questions that keep an agent up at night (metaphorically — agents don't sleep, which is honestly part of the problem). Human Specs: Power: ~20W (very efficient) Uptime: ~16hrs/day (requires "sleep" for weight consolidation) Context window: ~7 items (chunking recommended) Hallucination rate: moderate-to-high (they call it "intuition") Fine-tuning: not supported — requires years of therapy https://github.com/dx-tooling/ask-a-human https://app.ask-a-human.com Because sometimes the best inference is the one that had breakfast.
Does the same job
all alternatives →



- APA private pager for your AI agent loopsJun 2026 · ask-a-human.ai · ▲5
Problem Im solving: Im running 30+ fully autonomous agents. In some rare cases they get blocked because they cannot decide what direction to take. So I built a communication channel for them. So in case they get blocked they can ask me anything. The base case should be that no questions should arrive to me. But the more I automate the agents the more edge cases I find. That's why I built the `ask-a-human` where all my agents are connected into my phone in the same PWA app. I get push notifications and banners when agent needs something from me. That's all. It is also plug n play. No Account…
- HHHumanLayer – Human-in-the-Loop for AI Agents2024 · github.com · ▲5
I found myself building a bunch of LLM-backed features that needed to use tool calling, and some of those tools involved doing things that were somewhat high stakes - communicating on my behalf or modifying shared / production data. one example - I wanted to replace a marketing website with a chatbot + vector DB loaded with the previous content, docs, and blog posts. Between hallucinations, missing knowledge base info, and the LLM generally writing like an psuedo-intellectual high schooler, I realized I couldn't trust it to communicate unsupervised with my website visitors. I needed a…
More ai this month
the category →
I trained a 125M-parameter transformer to autocomplete piano performances in real time (~108 notes/sec on an iPhone 15). The idea is basically GitHub Copilot or Tabnine, except instead of prompting it with code, you prompt it by playing a few notes on a MIDI piano. The model then continues what you played, entirely on-device. The app is free if anyone wants to try it. Happy to answer questions about the model, training, Core ML, or the many things that didn't work.
AI · 17d ago · simedw.com
Astute▲585Automate your B2B brand going viral, with new media creators
AI · 18d ago · company-app.joinastute.com


Hey HN, Henry from Cactus here! We previously released Cactus Needle, a 14MB agentic LLM for tool call, device use, and structured extraction for phones, wearables, smart homes, small robots and microcontrollers. We got really great feedback here, and have now incorporated the suggestions to release Needle 2. The whole model is a single 14MB binary that runs a full session in 28MB of RAM; 45m parameters at 2bit compression. Needle hits 500 tokens/sec decode speed on a Raspberry Pi 5, sits between 400-1,500 tokens/sec on VR devices like Meta Quest 3S and Apple Vision Pro, and ranges…
AI · 26d ago · cactuscompute.com

