What problem does it solve? Working with PDF files programmatically is fragmented across many libraries and tools, making tasks like text extraction, form filling, merging, and OCR error-prone and time-consuming without clear guidance. ## Core Features & Use Cases - PDF Manipulation: Merge, split, rotate, encrypt, decrypt, watermark, and crop PDFs using pypdf, qpdf, and pdftk. - Content Extraction: Extract text, tables, metadata, and embedded images with pdfplumber, pypdfium2, and poppler-utils, including OCR for scanned documents via pytesseract. - PDF Creation & Form Filling: Generate new PDFs with reportlab or pdf-lib, and fill both fillable and non-fillable forms using dedicated scripts with bounding-box validation. - Use Case: Given a stack of scanned vendor invoices, convert them to images, run OCR to extract text, pull out table data with pdfplumber, and export the results to an Excel spreadsheet. ## Quick Start Ask the assistant to extract all text and tables from your PDF file, or to merge several PDF documents into one.