Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
aaaaqwq avatar

Mineru Extract

  • 11 installs
  • 82 repo stars
  • Updated August 2, 2026
  • aaaaqwq/claude-code-skills

mineru-extract is a Claude Code skill that uses the official MinerU API to convert URLs, PDFs, Office files, and images into clean Markdown with layout, table, formula, and OCR support.

About

mineru-extract is a Claude Code skill that uses the official MinerU parsing API to convert a URL, PDF, Office file, or image into clean Markdown plus structured outputs. It submits the source to MinerU, polls for completion, downloads the result zip, and extracts the main Markdown, handling layout, tables, formulas, and OCR. A developer uses it when web_fetch or a browser produces messy content and higher-fidelity parsing is needed. It requires a MINERU_TOKEN.

  • Uses the official MinerU API to convert URLs, PDFs, Office files, and images into clean Markdown
  • Higher-fidelity layout, table, formula, and OCR parsing than web_fetch or a browser
  • Ships scripts that submit a URL, poll for completion, download the zip, and extract Markdown

Mineru Extract by the numbers

  • 11 all-time installs (skills.sh)
  • Ranked #482 of 687 Office & Documents skills by installs in the Skillselion catalog
  • Data as of Aug 3, 2026 (Skillselion catalog sync)
At a glance

mineru-extract capabilities & compatibility

Requires a MINERU_TOKEN from mineru.net; usage billed by MinerU

Capabilities
pdf parsing · document extraction · ocr · markdown conversion
Use cases
pdf parsing · web scraping · documentation
Runs
Hosted SaaS
Pricing
Bring your own API key
From the docs

What mineru-extract says it does

Use MinerU as an upstream “content normalizer”: submit a URL to MinerU, poll for completion, download the result zip, and extract the main Markdown.
SKILL.md
URLs ending with `.pdf/.doc/.ppt/.png/.jpg` → `pipeline`
SKILL.md
npx skills add https://github.com/aaaaqwq/claude-code-skills --skill mineru-extract

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs11
repo stars82
Last updatedAugust 2, 2026
Repositoryaaaaqwq/claude-code-skills

What it does

Convert a URL, PDF, Office file, or image into clean Markdown using the MinerU parsing API.

Who is it for?

High-fidelity conversion of PDFs, Office files, images, or HTML pages into Markdown

Skip if: URLs MinerU cannot fetch due to anti-bot, geo, or login walls

When should I use this skill?

web_fetch or a browser extracts messy content and you need higher-fidelity parsing

What you get

Clean Markdown plus structured outputs extracted from a document or page, preserving layout, tables, and formulas.

  • Clean Markdown from the source document
  • Extracted result files (Markdown, JSON, assets) in the workspace
  • JSON contract with task_id and output paths

By the numbers

  • Supports 3 model modes: pipeline, vlm, MinerU-HTML

Files

SKILL.mdMarkdownGitHub ↗

MinerU Extract (official API)

Use MinerU as an upstream “content normalizer”: submit a URL to MinerU, poll for completion, download the result zip, and extract the main Markdown.

Quick start (MCP-aligned)

We align to the MinerU MCP mental model, but we do not run an MCP server.

  • Primary script (MCP-style): scripts/mineru_parse_documents.py
  • Input: --file-sources (comma/newline-separated)
  • Output: JSON contract on stdout: { ok, items, errors }
  • Low-level script (single URL): scripts/mineru_extract.py

Auth:

  • Set MINERU_TOKEN (Bearer token from mineru.net)

Default model heuristic:

  • URLs ending with .pdf/.doc/.ppt/.png/.jpgpipeline
  • Otherwise → MinerU-HTML (best for HTML pages like WeChat articles)

1) Configure token (skill-local)

Put secrets in skill root .env (do not paste into chat outputs):

# In the mineru-extract skill directory: .env
MINERU_TOKEN=your_token_here
MINERU_API_BASE=https://mineru.net

2) Parse URL(s) → Markdown (recommended)

MCP-style wrapper (returns JSON, optionally includes markdown text):

python3 mineru-extract/scripts/mineru_parse_documents.py \
  --file-sources "<URL1>\n<URL2>" \
  --language ch \
  --enable-ocr \
  --model-version MinerU-HTML

If you want the markdown content inline in the JSON (can be large):

python3 mineru-extract/scripts/mineru_parse_documents.py \
  --file-sources "<URL>" \
  --model-version MinerU-HTML \
  --emit-markdown --max-chars 20000

Low-level (single URL, print markdown to stdout):

python3 mineru-extract/scripts/mineru_extract.py "<URL>" --model MinerU-HTML --print > /tmp/out.md

Output

The script always downloads + extracts the MinerU result zip to:

~/.openclaw/workspace/mineru/<task_id>/

It writes:

  • result.zip
  • extracted files (Markdown + JSON + assets)

It prints a JSON summary to stderr with paths:

  • task_id, full_zip_url, out_dir, markdown_path

Parameters (common)

  • --model: pipeline | vlm | MinerU-HTML (HTML requires MinerU-HTML)
  • --ocr/--no-ocr: enable OCR (effective for pipeline/vlm)
  • --table/--no-table: table recognition
  • --formula/--no-formula: formula recognition
  • --language ch|en|...
  • --page-ranges "2,4-6" (non-HTML)
  • --timeout 600 / --poll-interval 2

Failure modes & fallbacks

  • MinerU may fail to fetch some URLs (anti-bot / geo / login).
  • Fallback: provide an HTML file or a PDF/long screenshot; then implement “upload + parse” flow with MinerU batch upload endpoints.
  • Always report the failing URL + MinerU err_msg and keep an original-source link in outputs.

References

  • MinerU API docs: https://mineru.net/apiManage/docs
  • MinerU output files: https://opendatalab.github.io/MinerU/reference/output_files/

Related skills

FAQ

What can mineru-extract parse?

URLs and HTML pages like WeChat articles, plus direct PDF, Office, and image links, into Markdown and structured outputs.

What credentials does it need?

A MINERU_TOKEN bearer token from mineru.net, set in the skill-local .env file.

Office & Documentspipelinesetl

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.