/install doc-extract
Doc Extract
Extract text and content from Word (.doc/.docx) files to Markdown using MinerU.
Install
npm install -g mineru-open-api
# or via Go (macOS/Linux):
go install github.com/opendatalab/MinerU-Ecosystem/cli/mineru-open-api@latest
Quick Start
# Quick extraction from .docx (no token required)
mineru-open-api flash-extract report.docx
# Save to directory
mineru-open-api flash-extract report.docx -o ./out/
# Extract .doc file (requires token)
mineru-open-api extract report.doc -o ./out/
# Extract with language hint
mineru-open-api extract report.docx --language en -o ./out/
Authentication
No token needed for flash-extract on .docx. Token required for .doc and extract:
mineru-open-api auth # Interactive token setup
export MINERU_TOKEN="your-token" # Or via environment variable
Create token at: https://mineru.net/apiManage/token
Capabilities
- Supported input: .doc, .docx (local file or URL)
.docx: supportsflash-extract(no token, max 10 MB / 20 pages) andextract.doc: requiresextractwith token- Language hint with
--language(default:ch, useenfor English) - Page range with
--pages(e.g.1-10)
Notes
.docrequiresextractwith token;.docxworks withflash-extractfor quick extraction- Output goes to stdout by default; use
-o \x3Cdir>to save to a file or directory - All progress/status messages go to stderr; document content goes to stdout
- MinerU is open-source by OpenDataLab (Shanghai AI Lab): https://github.com/opendatalab/MinerU
- Make sure OpenClaw is installed (local or Docker)
- Run the install command in chat:
/install doc-extract - After installation, invoke the skill by name or use
/doc-extract - Provide required inputs per the skill's parameter spec and get structured output
What is Doc Extract?
Extract text and content from Word documents (.doc, .docx) to Markdown using MinerU. A straightforward tool for reading and extracting Word file content. Fea... It is an AI Agent Skill for Claude Code / OpenClaw, with 188 downloads so far.
How do I install Doc Extract?
Run "/install doc-extract" in the OpenClaw or Claude Code chat to install it in one step — no extra setup required.
Is Doc Extract free?
Yes, Doc Extract is completely free, licensed under MIT-0. You can download, install and use it at no cost.
Which platforms does Doc Extract support?
Doc Extract is cross-platform and runs anywhere OpenClaw / Claude Code is available (cross-platform).
Who created Doc Extract?
It is built and maintained by mzlzyCA (@mzlzyca); the current version is v0.4.0.