What problem does it solve?
Getting usable content out of messy documents — scanned PDFs, slide decks, spreadsheets, emails, images — usually requires building and maintaining custom parsing and OCR pipelines. This Skill connects the hosted Unstructured Transform MCP server so the assistant can partition, enrich, and structure documents into clean, AI-ready data for search, Q&A, and summarization.
Core Features & Use Cases
- Document parsing across 60+ formats: Partitions PDFs, Word/Excel/PowerPoint, images, scanned files, and emails into structured elements like titles, paragraphs, and tables.
- Enrichment: Adds metadata, table and image descriptions, and entity recognition to extracted content.
- Grounded Q&A and search: Makes document sets searchable and answerable for RAG, knowledge bases, and summarization workflows.
- Use Case: Hand the assistant a folder of contracts and scanned reports, then ask it to extract the tables from each PDF and answer questions grounded in the source content.
Quick Start
Ask the assistant to connect Unstructured Transform, sign in during the OAuth step, then say "Use Unstructured Transform to parse these files and extract the tables from this PDF."