pdf

Extract text, tables, images, and page structure from PDF files.

Updated Aug 27, 2026
One-click install
npx skills add https://github.com/meisijiya/ohMeisijiyaCode --skill pdf-meisijiya
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: pdf
Source: https://github.com/meisijiya/ohMeisijiyaCode/tree/main/skills/pdf
Command: npx skills add https://github.com/meisijiya/ohMeisijiyaCode --skill pdf-meisijiya

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill removes the friction of working with PDF files by helping you read, extract, convert, merge, split, and inspect documents without manual rework.

Core Features & Use Cases

  • Text and Table Extraction: Pull readable text and structured content from PDFs into usable output.
  • Visual Inspection: Convert pages to images when layout, figures, or scan quality need review.
  • File Operations: Create PDFs from HTML, images, or DOCX files, and merge or split documents as needed.
  • Use Case: If you receive invoices, reports, or scanned forms in PDF format, use this Skill to extract the content, verify it visually, and transform it into a format that is easier to analyze or share.

Quick Start

Use the pdf skill to extract the text and tables from the attached PDF and return the results in a clean, readable format.

Frequently Asked Questions about pdf

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I extract text and tables from a PDF file for document processing?

To extract text and tables from a PDF file, you can use command-line extraction tools to pull structured content and page structure into a clean, readable output format. This handles invoices and reports without manual rework.

What is the best way to convert PDF pages to images for visual inspection?

Converting PDF pages to images for visual inspection is best handled by applying document-processing workflows that render layouts and scanned forms into image files, allowing you to verify scan quality and formatting issues.

Can I merge or split PDF files without losing page structure?

Yes, you can merge or split PDF files while validating page structure. File operations allow you to combine multiple documents or separate pages as needed, checking for missing pages or formatting issues during the process.

How do I create a PDF from HTML, images, or DOCX files?

You can create a PDF from HTML, images, or DOCX files by applying file conversion operations within a document-processing workflow. This transforms your source documents into a standard PDF format for easier sharing.

Does this PDF extraction approach work with scanned forms and invoices?

Yes, this PDF extraction approach works with scanned forms and invoices. It applies visual inspection and command-line extraction to read content from scanned documents and verify layout or scan quality issues.

What are the limitations when extracting structured content from PDF files?

Limitations when extracting structured content from PDF files include potential formatting issues or missing pages. Visual inspection and validation are required to ensure the extracted text and tables accurately reflect the original document.