Alternatives
Products that do what Parse a PDF from your terminal with PaddleOCR-VL-1.6 does
I built OpenParser because I wanted PaddleOCR-VL-1.6 behind an endpoint I could actually use. Now there are open-sourced CLI and TypeScript/Python SDKs: ``` npm install -g @openparser/cli openparser auth login openparser parse sync document.pdf ``` That is basically it. It doesnt get easier and cheaper that this. Give it a PDF and get back the text and document structure. Repo: https://github.com/eigenpal/openparser I would love feedback from anyone willing to try the CLI and tell me what is still annoying.
- 1

- 2

- 3PP
2012 · planet.racket-lang.org
- 4

- 5

- 6

- 7PO
2017 · pdfotter.com
- 8

- 9PT
Hi HN! I'm proud to share that we've launched a free PDF-to-Markdown CLI built on our proprietary (you might know it from PSPDFKit) engine. Most extractors are either fast but lose structure (markitdown, pymupdf4llm) or accurate but slow (docling). Ours ties with docling on accuracy but is orders of magnitude faster. https://github.com/pspdfkit/pdf-to-markdown We'd love feedback on it, and ofc send us files that break it.
Apr 2026
- 10P1
I just released ParseHawk v0.1.0: Apache-2.0 licensed 100% local document AI platform that extracts JSON from PDFs, images etc. It builds on top of NuMind's NuExtract3 but additionally enforces a provided JSON schema with constrained decoding. It works on Apple Silicon with pre-bundled vllm-metal as well as Linux + NVIDIA with vllm. Looking forward to your feedback!
Jun 2026 · github.com
- 11MT
Jul 2026 · github.com
- 12

- 13PA
Here's a neat hack I made recently to do basic PDF editing directly in a browser—without having to upload anything to a server. I was initially looking for a way to do simple PDF modification (extracting pages, merging, and adding page numbers). There are some good server-side tools for this (QPDF, PDFTk, PDFBox, iText, Hummus), but for better speed and privacy I really wanted a 100% client-side solution. There are a few good JavaScript PDF libraries for reading and displaying PDFs (pdf.js) and creating PDFs from scratch (jsPDF, PDFKit), but I couldn't find any for editing existing PDFs. So,…
2018
- 14PP
2020 · github.com
- 15PP
PDFFiddler Playground is a free PDF playground for manipulating PDF, extracting data from PDF, form filling, archiving, merging grouping and many many more. It is powered through Domain-driven custom scripting language (equivalent to Javascript) which can be quickly written through powerful editor with intellisense support. It has dozens of ready made template to play with. Please check out below link https://playground.pdffiddler.com/?apps=true Few templates, quick links has been added below 1) Merging group of PDF -…
2020
- 16CP
2014 · github.com
- 17PP
2019 · github.com
- 18PP
2018 · github.com
- 19OA
2024 · github.com
- 20CT
2024 · github.com
- 21OS
2015 · stackhut.com
- 22AN
React-print-pdf is a new open-source library that simplifies PDF creation with React. You can design your PDFs like a website, integrate data from your database to your documents, and reuse community components and templates or build your own. Created by three friends who were frustrated by existing solutions and wanted to make PDF creation easier for everyone. Now it’s yours. What are your thoughts on it ?!
2024 · github.com
- 23SP
Stet is a PostScript Level 3 interpreter, a PDF reader, and a print-quality PDF writer, all in pure Rust. All three converge on a single DisplayList type, so any output device (PNG, desktop viewer, PDF, WASM) works with any source — you can render a PS file to PDF or PNG, a PDF to PNG, or to a display list for a custom backend all through the same pipeline. The link above is the WASM build running entirely client-side — drop a PS, EPS, or PDF and it renders. It's a capability sampler, not a production viewer: no system fonts (browser sandbox), fixed zoom stops, single-threaded at ~2× native…
Apr 2026 · andycappdev.github.io
- 24IB
Hey everyone, I recently built a free and open-source tool that lets you manually redact PDFs and images right in your browser—no sign-ups, no watermarks. I also added an AI-powered auto-redaction feature to detect and remove sensitive info. Since this part uses external processing (OCR + AI), there are usage limits: you can auto-redact up to 3 single-page PDFs or images for free each day. Higher usage and multi-page support require a paid account. Right now, it works great for removing sensitive details like names, addresses, and emails, but I’d love to get your thoughts: • What redaction…
2025 · magicredact.com
Ranked by how close each launch is in meaning, then by votes. Refine with a description →