mineru

Parse PDF, Office, and image files into clean Markdown.

94|3|Updated Feb 13, 2026
One-click install
npx skills add https://github.com/Nebutra/MinerU-Skill --skill mineru-nebutra
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: mineru
Source: https://github.com/Nebutra/MinerU-Skill/tree/main
Command: npx skills add https://github.com/Nebutra/MinerU-Skill --skill mineru-nebutra

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) and assets (resource) components.

What problem does it solve?

This Skill solves the friction of converting complex documents like PDFs, Office files, and images into clean, agent-ready Markdown without requiring API keys, complex installations, or manual configuration.

Core Features & Use Cases

  • Zero-Config Parsing: Automatically routes files to the appropriate backend, providing a free Agent API for quick tasks and a Standard API for large-scale batches.
  • Multi-Format Support: Converts PDF, Word, PPT, Excel, HTML, and images into structured Markdown with preserved LaTeX formulas and tables.
  • Agent-Native Delivery: Directly pipes parsed content into 17 different knowledge and content tools like Obsidian, Notion, and Slack.

Quick Start

Use the mineru skill to parse the document located at the provided URL and output the resulting Markdown directly to your terminal.

Frequently Asked Questions about mineru

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert PDF and Office files into Markdown for AI agents?

You can convert PDF and Office files into Markdown for AI agents by using a cloud-based parser that automatically routes files to the appropriate backend. It requires no local dependencies or API keys for quick tasks and preserves LaTeX formulas and tables.

What is AI-native document parsing and how does it work for scanned PDFs?

AI-native document parsing uses an intelligent cloud engine to extract content from files, supporting OCR for scanned documents. It automatically routes files to the appropriate backend to output clean Markdown with preserved structures like LaTeX formulas and tables.

Do I need an API key to parse documents into Markdown?

No, you do not need an API key to parse documents into Markdown for lightweight tasks. The system provides a free Agent API for quick parsing jobs and a Standard API for auto-scaling large document batches, requiring zero local dependencies.

Can I directly send parsed Markdown to knowledge management platforms like Notion?

Yes, you can directly send parsed Markdown to platforms like Notion. The agent-native delivery feature pipes parsed content directly into 17 different knowledge and content tools, including Obsidian, Notion, and Slack, streamlining your workflow.

What is the best way to batch process large PDF files into Markdown?

The best way to batch process large PDF files into Markdown is using a cloud-based engine with a Standard API that provides auto-scaling. This approach handles large document batches efficiently without requiring manual configuration or local installations.

Does this PDF to Markdown parser preserve complex tables and LaTeX formulas?

Yes, this PDF to Markdown parser preserves complex tables and LaTeX formulas. It converts PDF, Word, PPT, Excel, HTML, and images into structured Markdown while maintaining the original document's mathematical and tabular structures.