What problem does it solve?
Removes repetitive manual work around PDFs by automating text and table extraction, programmatic form filling, and PDF manipulation so teams can process documents at scale without manual data entry.
Core Features & Use Cases
- Programmatic Form Filling: Fill native PDF form fields or add annotation-based text for non-fillable forms with validation and bounding-box workflows.
- Extraction & OCR: Extract plain text, structured tables, and images from born-digital and scanned PDFs with an OCR fallback.
- PDF Manipulation: Merge, split, rotate, watermark, encrypt, and generate PDFs for batch processing and reporting pipelines.
- Use Case: Automate invoice ingestion: extract invoice fields and tables, fill standardized forms, merge documents, and produce CSVs for accounting systems.
Quick Start
Use ck:pdf to extract all text and tables from input.pdf, perform OCR on scanned pages if needed, and save extracted data to output.csv while producing a merged PDF.