What problem does it solve?
This Skill helps you make sense of complex document pages by identifying layout regions, reading order, and structural elements that are hard to inspect manually. It is useful when a PDF or image contains mixed content such as headers, tables, figures, captions, and multi-column text.
Core Features & Use Cases
- Layout Detection: Finds text blocks, titles, section headers, lists, tables, figures, captions, footnotes, formulas, headers, and footers.
- Reading Order Analysis: Determines the correct sequence of elements so you can reconstruct document flow accurately.
- PDF and Image Analysis: Works on page images and PDF conversions for research papers, forms, articles, and scanned documents.
- Use Case: If you upload a dense report, this Skill can identify each content region and help you map the page into a clean structured representation.
Quick Start
Ask the Skill to analyze the layout of your document page and return the detected regions, reading order, and any tables or figures it finds.