Agent♥︎Age
Catalog

MinerU

Official

by linxule · TypeScript

MinerU document parsing API — PDFs, images, DOCX, PPTX with OCR and batch processing.

io.github.linxule/mineru — MCP Server for MinerU Document Parsing API

This MCP server provides access to the MinerU document parsing API to extract text, tables, and formulas from PDFs, DOCs, and images. It supports both VLM and pipeline processing modes, including OCR, batch parsing, and extracting selected page ranges. It can download extracted markdown using original filenames.

🛠️ Key Features

  • VLM model (90%+ accuracy for complex documents)
  • Pipeline model (fast processing for simple documents)
  • Local file upload for batch parsing from disk
  • Batch processing for up to 200 documents at once
  • Download & rename extracted markdown with original filenames
  • Page ranges to extract specific pages
  • 109 language OCR support

🚀 Use Cases

  • Convert PDFs, DOCX/DOCs, and images into extracted markdown
  • Extract specific portions of documents via page ranges
  • Run high-volume parsing jobs using batch processing
  • Handle multilingual OCR (109 languages)

⚡ Developer Benefits

  • Supports optimized output for Claude Code (73% token reduction noted)
  • Works with local uploads and batch workflows
  • Produces markdown output tied to original filenames