wikimedia-commons-pdf

Manage PDF and DjVu documents on Wikimedia Commons with OCR extraction.

15|6|Updated Feb 17, 2026
One-click install
npx skills add https://github.com/fuzheado/Wikipedia-AI-Skills --skill wikimedia-commons-pdf
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: wikimedia-commons-pdf
Source: https://github.com/fuzheado/Wikipedia-AI-Skills/tree/main/.claude/skills/wikimedia-commons-pdf
Command: npx skills add https://github.com/fuzheado/Wikipedia-AI-Skills --skill wikimedia-commons-pdf

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires wikimedia-api-access, wikimedia-commons, wikimedia-commons-thumbnails, and includes scripts (resource) and references (resource) and assets (resource) components.

What problem does it solve?

This Skill streamlines the process of working with PDF and DjVu documents on Wikimedia Commons, including document management, OCR, and integration with Wikisource.

Core Features & Use Cases

  • Document Management: Handle multi-page document models, page selection, and metadata editing.
  • OCR Text Extraction: Extract text from PDF and DjVu files for use in Wikisource.
  • Integration with Wikisource: Link documents to Wikisource pages for proofreading and transcription.
  • Use Case: A researcher wants to upload a scanned book to Commons and link it to a Wikisource project for collaborative proofreading.

Quick Start

Use the wikimedia-commons-pdf skill to upload the PDF file 'book.pdf' to Commons and link it to the corresponding Wikisource page.

Frequently Asked Questions about wikimedia-commons-pdf

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I extract OCR text from a DjVu file for Wikisource proofreading?

This Skill performs OCR text extraction from DjVu files, processing document pages and linking the extracted text to a Wikisource project for collaborative proofreading.

Can I manage multi-page PDF metadata on Wikimedia Commons?

Yes, this Skill supports multi-page PDF metadata management on Wikimedia Commons, handling page selection, document models, and various metadata editing operations.

What is the best way to upload a scanned PDF and link it to Wikisource?

The best way is using this Skill to upload the scanned PDF to Wikimedia Commons, extract its text via OCR, and connect it to the corresponding Wikisource page for transcription.

Does this approach work with both PDF and DjVu document formats?

Yes, this Skill works with both PDF and DjVu document formats, providing multi-page document handling, OCR text extraction, and Wikisource integration for both file types.

Do I need Wikimedia API access to extract text from Commons documents?

Yes, Wikimedia API access is required to extract text from Commons documents, as this Skill depends on wikimedia-api-access and wikimedia-commons dependencies to perform OCR and metadata operations.