nowfound

Alternatives

Products that do what VLM Camera — Chrome 機能拡張 does

Turn your PC into a 100% secure local AI camera. Free,WebGPU

  1. 1
    SmolVLM2206

    Smallest Video LM Ever from HuggingFace

    2025

  2. 2
    Molmo 298

    SOTA video understanding, pointing, and tracking VLM

    Dec 2025

  3. 3

    Take raw photos with proof they're real, not AI

    May 2026

  4. 4

    Your ultimate NotebookLM's Chrome Extension

    May 2026

  5. 5
    GLM-4.6V239

    Open-source multimodal model with native tool use

    Dec 2025

  6. 6

    Vision-to-code foundation model for real GUI automation

    Apr 2026

  7. 7
    InternVL3135

    Open MLLMs excelling in vision, reasoning & long context

    2025

  8. 8

    256M VLM for end-to-end document AI

    2025

  9. 9

    Ultra-efficient 1.3B vision-language model for mobile

    May 2026

  10. 10
    LM Studio209

    Discover, download, and run local LLMs (incl. DeepSeek R1)

    2025

  11. 11
    AutoMask132

    In-browser AI imaging tool to anonymize people in pictures.

    2020

  12. 12

    GPT-4o level vision model on the phone

    2025

  13. 13
    SmolVLA139

    Powerful robotics VLA that runs on consumer hardware

    2025

  14. 14

    Supercharge your video conferences and recording with AI

    2023

  15. 15
    SimCam119

    Unparalleled AI home security camera

    2019

  16. 16
    NVLM 1.0200

    Open frontier-class multimodal LLMs

    2024

  17. 17

    Chrome extension for NotebookLM productivity utilities

    Dec 2025

  18. 18
    Bgblur80

    Blur faces, plates, and awkward photobombs.

    Dec 2025

  19. 19

    Virtual backgrounds, beautification, auto-framing & more

    2023

  20. 20

    Chrome extension allow the user to communicate with the AI

    Sep 2025

  21. 21
    Ferret193

    Refer and ground anything anywhere at any granularity

    2024

  22. 22

    Real-time AI captions and translation for any browser video

    Jun 2026

  23. 23OU

    The traditional pipeline for unstructured data extraction typically follows these steps: 1. Image → OCR Model (e.g., Google Vision) → Layout Model (e.g. Surya) → LLM → Final Answer However, this can be streamlined using a Vision-Language Model (VLM): 2. Image → VLM → Final Answer Recently VLMs have improved a lot for OCR and document understanding tasks, specifically the Qwen-2.5-VL series. We can run the Qwen-2.5-VL-7B-AWQ model locally with just 16GB VRAM, and perform end-to-end information extraction (fields and table extraction) without any external models. Hallucination with VLMs One…

    2025 · github.com

  24. 24

    Open-source browser automation for local AI agents

    May 2026

Ranked by how close each launch is in meaning, then by votes. Refine with a description →