What problem does it solve?
This Skill removes the manual burden of reading, copying, and reformatting text from PDFs, scanned pages, images, and other document formats by turning them into searchable, editable output.
Core Features & Use Cases
- Remote-first extraction: Uses web extraction first for document URLs, which is ideal for arXiv papers and other public PDFs.
- Local document parsing: Extracts text from text-based PDFs with lightweight tools and handles complex layout or OCR-heavy documents with higher-quality OCR workflows.
- Scanned document recovery: Supports scanned PDFs, equations, tables, forms, code blocks, image extraction, and metadata retrieval for research, archival, and document-processing tasks.
- Fallback workflows: Provides guidance for large or difficult scans, including targeted page extraction and Google Drive OCR when local OCR is impractical.
- Use case: A researcher can upload an arXiv PDF, a scanned report, or a folder of mixed documents and quickly convert them into text or markdown for analysis.
Quick Start
Ask the assistant to extract the text or markdown from your PDF or scanned document, and if the file is a URL it should try web extraction first before falling back to local OCR tools.