Agent Skills: unlimited-ocr-parse-document

Parse an image, PDF, or directory into markdown with LaTeX formulas, HTML/markdown tables, and per-block layout bounding boxes, using baidu/Unlimited-OCR locally — free, offline, no API key, no upload quota. Runs on Apple Silicon via MLX (~2.4 s/image, ~5 GB) or an NVIDIA GPU via transformers. Use when transcribing scanned or image-based documents, extracting equations from a paper, converting a PDF to markdown, OCR-ing screenshots, or replacing a paid vision-API OCR call with a local one. TRIGGERS - unlimited-ocr, ocr this, ocr a pdf, extract text from image, transcribe document, pdf to markdown, document parsing, extract formulas, extract equations, latex from image, local ocr, offline ocr, scanned pdf, image to text, layout parsing, bounding boxes from document.

UncategorizedID: terrylica/cc-skills/unlimited-ocr-parse-document

Install this agent skill to your local

pnpm dlx add-skill https://github.com/terrylica/cc-skills/unlimited-ocr-parse-document

Skill Files

Browse the full folder contents for unlimited-ocr-parse-document.

Download Skill

Loading file tree…

Select a file to preview its contents.