MerchantryTidbits

Public catalog

Browse tools

20 shown. Search ranks against problem language; browsing defaults to stronger public signals.

doc-parsingclilocal

tesseract

open-source OCR for images and scanned PDFs

Verified 2026-08-10 ยท 12,296 stars

media-processingappfree

BetterShot

You need to capture, annotate, and blur-redact screenshots on macOS for docs or bug reports without paying for a CleanShot X subscription

Verified 2026-08-10

doc-parsinglibrarylocal

DeepSeek-OCR 2

You need to convert scanned document images into structured markdown locally on your own GPU rather than through a cloud OCR API

Verified 2026-08-10

doc-parsingapplocal

DeepSeek-OCR Client

You want to run DeepSeek-OCR on images locally through a point-and-click desktop app instead of writing Python inference scripts

Verified 2026-08-10

doc-parsingapplocal

DeepSeek-OCR Dockerized API

Scanned or image-based PDFs have no text layer and you need structured markdown out of them

Verified 2026-08-10

doc-parsinglibrarylocal

dots.mocr

You need scanned or image-based multilingual documents parsed into structured layout JSON and markdown with correct reading order

Verified 2026-08-10

doc-parsinglibrarylocal

GLM-OCR

Scanned business documents with complex tables, seals, or code blocks come out garbled through conventional OCR and you need layout-aware markdown plus JSON output

Verified 2026-08-10

doc-parsinglibrarylocal

HunyuanOCR

I need local OCR and document parsing from a document image through a documented inference server.

Verified 2026-08-10

doc-parsingclilocal

LiteParse

You must parse PDFs, Office documents, or scanned images to markdown or structured JSON without uploading them to any cloud service

Verified 2026-08-10

doc-parsingclilocal

OCRmyPDF

add a searchable OCR text layer to scanned PDFs while preserving pages

Verified 2026-08-10

doc-parsinglibrarylocal

Ollama OCR

I need to extract text or structured output from an image or PDF with a local Ollama vision model.

Verified 2026-08-10

doc-parsingclilocal

olmOCR

You have scanned or image-based PDFs with multi-column layouts, equations, tables, or handwriting and need Markdown or text in natural reading order.

Verified 2026-08-10

doc-parsinglibrarylocal

paddleocr

PaddleOCR for multilingual OCR pipelines

Verified 2026-08-10

document-automationappfree

Paperless-ngx

archive scanned documents with OCR and searchable metadata

Verified 2026-08-10

document-automationappfree

PDF Toolkit (Pdf_Tools)

You need to merge, split, compress, or reorder PDF pages on an Android phone without uploading the file to an online converter

Verified 2026-08-10

doc-parsingclilocal

pix2tex (LaTeX-OCR)

You have a screenshot or image of a printed math formula from a paper and need the corresponding LaTeX source instead of retyping it symbol by symbol.

Verified 2026-08-10

doc-parsinglibrarylocal

RapidOCR

Text must be extracted from images with Chinese and English content fully offline, without sending documents to a cloud OCR API

Verified 2026-08-10

doc-parsinglibraryfree

surya

use surya for evals testing

Verified 2026-08-10

doc-parsingappfree

Umi-OCR

Hundreds of local images or screenshots need their text extracted in batch, fully offline

Verified 2026-08-10

doc-parsinglibraryfree

unstructured

partition PDF HTML into structured elements

Verified 2026-08-10