pdf-reader

Extract and analyze text, tables, and forms from PDF documents.

30|7|Updated Mar 1, 2026
One-click install
npx skills add https://github.com/rfdiosuao/openfang-cn --skill pdf-reader-rfdiosuao
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: pdf-reader
Source: https://github.com/rfdiosuao/openfang-cn/tree/main/crates/openfang-skills/bundled/pdf-reader
Command: npx skills add https://github.com/rfdiosuao/openfang-cn --skill pdf-reader-rfdiosuao

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill addresses the challenge of extracting and interpreting information locked within PDF documents, making content accessible and actionable.

Core Features & Use Cases

  • Content Extraction: Extracts text, tables, and form data from PDFs, preserving document structure.
  • Data Interpretation: Summarizes content, identifies key data points, and compares documents.
  • Use Case: A researcher needs to quickly understand the key findings from a dozen academic papers; this Skill can extract the abstract, methodology, and conclusions from each, providing a concise overview.

Quick Start

Use the pdf-reader skill to extract all text from the attached document 'report.pdf'.

Frequently Asked Questions about pdf-reader

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I extract text and tables from a PDF document?

To extract text and tables from a PDF document, the skill processes the file while preserving its logical structure. It accurately pulls content from text, tables, and forms, making the extracted data accessible for further analysis.

Can I summarize academic papers and financial reports directly from a PDF?

Yes, you can summarize academic papers and financial reports directly from a PDF. The skill identifies key data points and provides concise overviews of abstracts, methodologies, and conclusions from various document types.

Does PDF content extraction work on scanned documents?

PDF content extraction does work on scanned documents by utilizing OCR. This allows the skill to interpret and extract text from image-based PDFs, making scanned legal and financial forms searchable and actionable.

What is the best way to search and compare data across multiple PDF files?

The best way to search and compare data across multiple PDF files is using the skill's built-in search and comparison functionalities. It extracts and evaluates key data points from each document, enabling direct comparison of findings.

Are there limitations when extracting form data from complex PDF structures?

While the skill preserves logical structure and handles complex form data extraction, heavily corrupted or poorly scanned documents may affect OCR accuracy. It is designed to manage standard legal, financial, and academic structures effectively.