unlimited-ocr-document-parsing

Installation
SKILL.md

Unlimited-OCR document parsing

Use the bundled caller to extract the complete document. Prefer this skill when the user asks for long-document OCR, Markdown conversion, reading-order preservation, tables, formulas, or multi-page parsing.

Route requests here when they mention long-document OCR, PDF/OFD/Office to Markdown, document digitization, table or formula recognition, or the Chinese phrases 长文档 OCR / PDF 转 Markdown / 图片转文字 / 文档解析 / 表格提取 / 公式识别 / 多页扫描件. Choose the cloud or local provider based on the input format, privacy requirements, and available runtime.

Choose a provider

  • baidu: default; supports local files and public HTTPS URLs, including PDF/OFD/Office/text formats. Requires UNLIMITED_OCR_API_KEY plus UNLIMITED_OCR_SECRET_KEY, or an existing UNLIMITED_OCR_ACCESS_TOKEN.
  • local: sends local images/PDFs to UNLIMITED_OCR_LOCAL_BASE_URL. Use UNLIMITED_OCR_LOCAL_BACKEND=sglang for the official SGLang server, or openai for another compatible server. Local mode intentionally rejects --file-url.

Run

From this skill directory:

Installs
3
GitHub Stars
1
First Seen
8 days ago
unlimited-ocr-document-parsing — aidenwu0209/unlimited-ocr-skill