What problem does it solve?
Working with PDFs often leads to frustrating, unreliable results: extracted text misses layout context, scanned documents are completely unreadable with standard tools, and generated PDFs have broken fonts, clipped content, or misaligned elements. This Skill eliminates those issues by providing vetted tools and clear workflows for every common PDF task.
Core Features & Use Cases
- Layout-Aware Text & Data Extraction: Pull text and structured table data from text-based PDFs while preserving coordinate context for accurate retrieval.
- Scanned Document OCR: Convert image-based scanned PDFs, including those with equations or multi-column layouts, into readable, editable content with correct reading order.
- Polished PDF Generation & Editing: Split, merge, rotate, or generate PDFs from structured content, with mandatory visual verification to ensure layout fidelity before delivery.
Use case: For example, process a batch of scanned vendor invoices by running OCR to extract line items, then compile the data into a structured spreadsheet for accounting.
Quick Start
Use this Skill to extract all text and table data from your attached scanned invoice PDF and export the results to a CSV file.