What problem does it solve?
PDF documents are often siloed in workflows where text extraction, data extraction from forms, page rotation or merging for reports, and secure handling are performed manually. This skill provides a cohesive, programmable approach to reading, analyzing, and transforming PDFs, including reading content, extracting text and tables, merging or splitting documents, rotating pages, adding watermarks, creating new PDFs, filling forms, encrypting/decrypting, extracting images, and applying OCR to scanned files.
Core Features & Use Cases
- Automated PDF processing workflow: extract text and tables, merge or split PDFs, rotate pages, and apply watermarks to batches of documents.
- Form handling and security: fill PDF forms and manage encryption/decryption, enabling compliant document handling.
- Create and manipulate PDFs for reporting or archival tasks, including extraction of images and OCR-enabled searchability.
- Use Case: Digitize invoices by extracting numbers, merging related documents, and storing searchable PDFs in a records system.
Quick Start
Run the pdf skill to extract all text from a sample document and save it to a new file.