What problem does it solve?
Manual handling of PDFs—reading, extracting structured data, filling forms, merging, splitting, applying OCR, and annotating—is time consuming and error prone; this skill centralizes reliable, repeatable PDF workflows so you can process documents at scale with fewer mistakes.
Core Features & Use Cases
- Text and Table Extraction: Extract plain text, layout-preserving text, and structured tables from born-digital and scanned PDFs.
- Form Handling & Filling: Detect fillable fields, extract field metadata, or place annotations on non-fillable forms with coordinate conversion helpers.
- Manipulation and Utilities: Merge/split/rotate pages, add watermarks, extract images, perform OCR, and encrypt/decrypt documents. Use case: batch-extract invoice data and populate a CSV or automatically fill standardized application PDFs.
Quick Start
Extract all text and tables from invoice-q3.pdf, run OCR if needed, and output a merged CSV with detected line items and totals.