PaddleOCR-Skills
Extract text and structure from images, scans, and PDFs
All Skills in This Repository (2)
Pure Emerald Level IndicatorsFrequently Asked Questions
FAQPage SchemaHow to install PaddleOCR-Skills?โผ
Run `npx skills add Aidenwu0209/PaddleOCR-Skills --all -g -y` in your terminal to install both skills globally for your agent.
How to extract text from images and PDFs?โผ
The text recognition skill pulls exact text from screenshots, photos, scans, and PDFs, including Chinese and handwritten content. Just give your agent the file path or URL and it returns the full text.
Can it convert PDFs to Markdown with tables?โผ
Yes. The document parsing skill rebuilds tables, formulas as LaTeX, figures, and multi-column layouts into structured Markdown or JSON with correct reading order.
Does PaddleOCR-Skills work with Claude Code and Cursor?โผ
Yes. Both skills follow the universal SKILL.md standard and run in Claude Code, Codex, Cursor, GitHub Copilot, OpenCode, and OpenClaw.
Do I need an API key to use PaddleOCR-Skills?โผ
Yes. You need a free access token and endpoint URL from paddleocr.com, set as environment variables. The skills guide you through configuration on first run.
Related Repositories in Content & Communication
View All in Content & Communicationโskills
Official Anthropic skills for documents, design, and developer workflows
humanizer
Rewrite AI-sounding text so it reads like a person wrote it
UniGetUI
Automate software translation reviews, diffs, and localization quality checks