pdf-ocr

Convert PDF documents to editable Word files with Chinese OCR.

2|Updated Feb 27, 2026
One-click install
npx skills add https://github.com/jiyangnan/xiaonangua-openclaw-skills --skill pdf-ocr-jiyangnan
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: pdf-ocr
Source: https://github.com/jiyangnan/xiaonangua-openclaw-skills/tree/main/skills/tools/pdf-ocr
Command: npx skills add https://github.com/jiyangnan/xiaonangua-openclaw-skills --skill pdf-ocr-jiyangnan

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires pymupdf, python-docx, pillow, and includes scripts (resource) and assets (resource) components.

What problem does it solve?

This Skill automates the conversion of PDF documents into editable Word documents, providing accurate OCR text recognition and handling various page types like text pages, images, and colorful covers.

Core Features & Use Cases

  • PDF to Word Conversion: Accurately converts PDFs to Word format, preserving text, images, and formatting.
  • Chinese OCR: Supports Chinese OCR recognition for accurate text extraction.
  • Page Handling: Automatically handles different page types, including cropping headers and footers, and retaining images.
  • Use Case: Ideal for converting scanned documents or PDFs with complex layouts into editable text for further processing or archiving.

Quick Start

Convert the PDF file 'example.pdf' to a Word document using the pdf-ocr skill.

Frequently Asked Questions about pdf-ocr

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert a scanned PDF to an editable Word document?

The skill supports Chinese OCR recognition to accurately extract text from scanned PDFs. It automatically processes various page types including text pages, colorful covers, and images to convert them into editable Word documents.

Does PDF to Word conversion work with complex layouts and images?

Yes, PDF to Word conversion handles complex layouts by automatically cropping headers and footers while retaining images. It processes various page types to ensure accurate content extraction and digitization for your documents.

What Python libraries are required for PDF OCR text extraction?

PDF OCR text extraction requires PyMuPDF, python-docx, and Pillow. These Python libraries handle the processing of PDF documents and images to perform accurate text recognition and generate Word files.

Can I use this tool for Chinese OCR text recognition in PDFs?

Yes, you can use this tool for Chinese OCR text recognition in PDFs. It accurately extracts Chinese text from scanned documents and complex layouts, making it ideal for digitizing and archiving content.

What is the best way to extract text from scanned documents for archiving?

The best way to extract text from scanned documents for archiving is using automated OCR conversion. This approach accurately digitizes content by handling various page types and exporting directly to an editable Word format.