nowfound

Alternatives

Products that do what GLM-OCR does

GLM-OCR: An online OCR focused on document structure

  1. 1OP
  2. 2

    Introducing the world’s best document understanding API

    2025

  3. 3OP

    Hi HN, I’ve been working on an OCR pipeline specifically optimized for machine learning dataset preparation. It’s designed to process complex academic materials — including math formulas, tables, figures, and multilingual text — and output clean, structured formats like JSON and Markdown. Some features: • Multi-stage OCR combining DocLayout-YOLO, Google Vision, MathPix, and Gemini Pro Vision • Extracts and understands diagrams, tables, LaTeX-style math, and multilingual text (Japanese/Korean/English) • Highly tuned for ML training pipelines, including dataset generation and…

    2025 · github.com

  4. 4
    GLM-5.3254

    Coding leap from scaled post-training on the same base

    23d ago · z.ai

  5. 5
    TurboLens133

    Fast, accurate OCR & insights from any images.

    2024

  6. 6ZD

    This started out as a weekend hack with gpt-4-mini, using the very basic strategy of "just ask the ai to ocr the document". But this turned out to be better performing than our current implementation of Unstructured/Textract. At pretty much the same cost. I've tested almost every variant of document OCR over the past year, especially trying things like table / chart extraction. I've found the rules based extraction has always been lacking. Documents are meant to be a visual representation after all. With weird layouts, tables, charts, etc. Using a vision model just make sense! In…

    2024 · github.com

  7. 7

    OCR Software & API for realtime data extraction from Invoice

    2023

  8. 8

    Fast and efficient text recognition from any image and PDF

    2019

  9. 9
    GLM-4.6V239

    Open-source multimodal model with native tool use

    Dec 2025 · z.ai

  10. 10

    SOTA document parsing & OCR in just 0.9B parameters

    Feb 2026 · ocr.z.ai

  11. 11

    Read documents like an image

    Oct 2025

  12. 12PT

    I've developed a Python API service that uses GPT-4o for OCR on PDFs. It features parallel processing and batch handling for improved performance. Not only does it convert PDF to markdown, but it also describes the images within the PDF using captions like `[Image: This picture shows 4 people waving]`. In testing with NASA's Apollo 17 flight documents, it successfully converted complex, multi-oriented pages into well-structured Markdown. The project is open-source and available on GitHub. Feedback is welcome.

    2024 · github.com

  13. 13OB

    OCR/Document extraction field has seen lot of action recently with releases like Mixtral OCR, Andrew Ng's agentic document processing etc. Also there are several benchmarks for OCR, however all testing for something slightly different which make good comparison of models very hard. To give an example, some models like mixtral-ocr only try to convert a document to markdown format. You have to use another LLM on top of it to get the final result. Some VLM’s directly give structured information like key fields from documents like invoices, but you have to either add business rules on top…

    2025 · nanonets.com

  14. 14OA

    I built OCR Arena as a free playground for the community to compare leading foundation VLMs and open-source OCR models side-by-side. Upload any doc, measure accuracy, and (optionally) vote for the models on a public leaderboard. It currently has Gemini 3, dots.ocr, DeepSeek, GPT5, olmOCR 2, Qwen, and a few others. If there's any others you'd like included, let me know!

    Nov 2025 · ocrarena.ai

  15. 15

    Vision-to-code foundation model for real GUI automation

    Apr 2026 · docs.z.ai

  16. 16YS
  17. 17

    Intelligent text extraction using OCR and deep learning

    2019

  18. 18
    aOCR14

    API for converting complex documents into structured data

    Jan 2026 · aocr.in

  19. 19

    Extract data from invoices and receipts with AI based OCR

    2018

  20. 20
    PDF2MD51

    Convert your PDFs to markdown With AI OCR

    2024

  21. 21

    Fully offline OCR with 100+ languages support. Full Privacy

    2025

  22. 22BO

    2024 · github.com

  23. 23

    Auto-regressive for dense-knowledge & high-fidelity images

    Jan 2026 · z.ai

  24. 24

    Text-to-Image with Perfect Multilingual Text Rendering

    Jan 2026

Ranked by how close each launch is in meaning, then by votes. Refine with a description →