doc-reader

Extract text from PDF, Word, Excel, PowerPoint, and image documents.

Updated May 25, 2026
One-click install
npx skills add https://github.com/NigarumOvum/AutoTrading --skill doc-reader-nigarumovum
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: doc-reader
Source: https://github.com/NigarumOvum/AutoTrading/tree/main/Vibe-Trading/agent/src/skills/doc-reader
Command: npx skills add https://github.com/NigarumOvum/AutoTrading --skill doc-reader-nigarumovum

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill solves the problem of manually extracting text from various document formats, allowing users to quickly convert documents into editable text.

Core Features & Use Cases

  • Universal Document Reading: Extracts text from PDF, Word, Excel, PowerPoint, images, CSV/TSV, plain text, JSON/YAML/TOML, HTML/XML, and source-code files.
  • PDF Text Extraction: Offers OCR for image-based PDFs and allows for selective page extraction.
  • Excel Preview: Provides a quick preview of the first 100 rows of each sheet in Excel files.
  • Use Case: Imagine you need to extract text from a complex PDF report. Use this Skill to quickly convert the entire document into a single, editable text file.

Quick Start

Use the doc-reader skill to extract text from the attached file 'report.pdf'.

Frequently Asked Questions about doc-reader

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I extract text from a scanned PDF using OCR?

To extract text from a scanned PDF, you need OCR processing. This Skill supports OCR for image-based PDFs, converting scanned document content into editable text while allowing selective page extraction for specific sections.

Can I extract text from Word, Excel, and PowerPoint documents?

Yes, you can extract text from Word, Excel, and PowerPoint files. This universal document reader processes multiple formats, providing a quick preview of the first 100 rows per Excel sheet and converting presentation content into editable text.

What is the best way to convert JSON, YAML, and HTML files into plain text?

The best way to convert JSON, YAML, and HTML files into plain text is using a universal extraction tool. This Skill reads structured data and markup languages alongside source-code files, transforming them into a single editable text format.

Does this document extraction tool support selective page extraction for large reports?

Yes, this document extraction tool supports selective page extraction for PDF files. You can target specific pages within large reports, avoiding the need to process the entire document and streamlining your administrative workflows.

How do I extract data from CSV and TSV files for data entry tasks?

To extract data from CSV and TSV files for data entry tasks, use a universal document reader. This Skill processes delimited text formats alongside other document types, converting tabular data into editable text for administrative workflows.