/install doc-ocr
Doc OCR
Use OCR to extract text from Word (.docx) files that contain scanned pages or image-embedded content, using MinerU.
Install
npm install -g mineru-open-api
# or via Go (macOS/Linux):
go install github.com/opendatalab/MinerU-Ecosystem/cli/mineru-open-api@latest
Quick Start
# OCR extraction from .docx (requires token)
mineru-open-api extract report.docx --ocr -o ./out/
# With VLM model for better accuracy on complex image layouts
mineru-open-api extract report.docx --ocr --model vlm -o ./out/
Authentication
Token required:
mineru-open-api auth # Interactive token setup
export MINERU_TOKEN="your-token" # Or via environment variable
Create token at: https://mineru.net/apiManage/token
Capabilities
- Supported input: .docx (local file or URL)
- OCR is only available via
extract(requires token) - Use
--ocrflag to enable OCR on image-embedded content - Use
--model vlmfor complex or mixed-content documents - Language hint with
--language(default:ch, useenfor English)
Notes
- OCR is NOT available in
flash-extract— useextractwith--ocr - If the
.docxhas a normal text layer, OCR is not needed — usedoc-extractinstead - Output goes to stdout by default; use
-o \x3Cdir>to save to a file or directory - All progress/status messages go to stderr; document content goes to stdout
- MinerU is open-source by OpenDataLab (Shanghai AI Lab): https://github.com/opendatalab/MinerU
- Make sure OpenClaw is installed (local or Docker)
- Run the install command in chat:
/install doc-ocr - After installation, invoke the skill by name or use
/doc-ocr - Provide required inputs per the skill's parameter spec and get structured output
What is Doc OCR?
OCR (Optical Character Recognition) for Word documents (.docx) containing scanned pages or image-embedded content. Uses MinerU to extract text from Word file... It is an AI Agent Skill for Claude Code / OpenClaw, with 201 downloads so far.
How do I install Doc OCR?
Run "/install doc-ocr" in the OpenClaw or Claude Code chat to install it in one step — no extra setup required.
Is Doc OCR free?
Yes, Doc OCR is completely free, licensed under MIT-0. You can download, install and use it at no cost.
Which platforms does Doc OCR support?
Doc OCR is cross-platform and runs anywhere OpenClaw / Claude Code is available (cross-platform).
Who created Doc OCR?
It is built and maintained by mzlzyCA (@mzlzyca); the current version is v0.4.0.