Alternatives
Products that do what Perceptron Mk1 does
Frontier video reasoning for the physical world
- 1

- 2AR
Hey it’s Hassaan & Quinn – co-founders of Tavus, an AI research company and developer platform for video APIs. We’ve been building AI video models for ‘digital twins’ or ‘avatars’ since 2020. We’re sharing some of the challenges we faced building an AI video interface that has realistic conversations with a human, including getting it to under 1 second of latency. To try it, talk to Hassaan’s digital twin: https://www.hassaanraza.com, or to our "demo twin" Carter: https://www.tavus.io We built this because until now, we've had to adapt communication to the limits of…
2024
- 3

- 4
- 5

Powers faster, efficient reasoning for long-running agents
Jun 2026 · developer.nvidia.com
- 6

- 7

- 8

- 9

- 10

- 11

- 12

- 13

- 14

- 15

Open-weight 15B multimodal model for thinking and GUI agents
Mar 2026 · microsoft.com
- 16

- 17

Google's SOTA robotics model for visual & spatial reasoning!
Apr 2026 · deepmind.google
- 18

Talk to Static, a public AI shared by everyone. There are no separate copies: what it learns from one conversation can shape another.
22d ago · wildstatic.com
- 19PA
2017 · github.com
- 20OP
2020 · github.com
- 21

- 22SV
Hi HN, I am Anubhav from Ramanlabs. We have been working on a native gui application to allow users to search any video data( mp4, mkv) or video streams (http/rtsp) using computer vision. Application is supposed to work like a video player which displays decoded frames and recognizes objects concurrently, making it an interactive experience. It works in super real-time and only expects a quad-core CPU with AVX2 instructions at minimum. Application is free to download (without any signup/account). We are only supporting WINDOWS for now [0]. Even though this is a binary application,…
2022 · ramanlabs.in
- 23CW
Hey HN! We’re excited to share Orion [1] — our new visual agent that sees, reasons, and acts across images, videos, and documents. Frontier VLMs (GPT, Claude, Gemini) can describe what they see, but they can’t reliably act on visual inputs. Ask them to detect objects, segment images, or chain visual steps — they’ll fail in surprisingly inconsistent ways. High-res images collapse to ~1024px. And the visual AI ecosystem is fragmented across separate APIs for image understanding, OCR, image-gen, video-gen, etc. We built Orion to fix this. Orion combines VLM reasoning with reliable…
Nov 2025 · chat.vlm.run
- 24

Ranked by how close each launch is in meaning, then by votes. Refine with a description →