nutrient-document-processing

Automate document conversion, OCR, extraction, redaction, watermarking, signing, and form filling via the Nutrient DWS API.

1|Updated Mar 18, 2026
One-click install
npx skills add https://github.com/xxih/ai-harness-zh --skill nutrient-document-processing-xxih
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: nutrient-document-processing
Source: https://github.com/xxih/ai-harness-zh/tree/main/references/translations/everything-claude-code/docs/zh-CN/skills/nutrient-document-processing
Command: npx skills add https://github.com/xxih/ai-harness-zh --skill nutrient-document-processing-xxih

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

该技能解决了对文档进行格式转换、文本与表格提取、对扫描件进行 OCR、编辑敏感信息、添加水印、数字签名以及填写 PDF 表单的繁琐工作。

Core Features & Use Cases

  • 文档格式转换、文本与表格提取、OCR、PII 编辑、添加水印、数字签名以及填写表单等场景的端到端处理。
  • 支持 PDF、DOCX、XLSX、PPTX、HTML 等多种文档格式,适用于合同、发票、报告等文档的批量处理。

Quick Start

Provide a document and request conversion, OCR, redaction, and signing using the Nutrient DWS API.

Frequently Asked Questions about nutrient-document-processing

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate PDF form filling and document conversion for invoices?

Automating PDF form filling and document conversion is handled end-to-end using the Nutrient DWS API. It processes invoices and contracts by converting formats, extracting data, filling forms, and returning completed outputs.

Can I perform OCR on scanned PDF documents and extract text from tables?

Yes, OCR on scanned PDF documents extracts text and tables efficiently. The processing workflow applies optical character recognition to scanned images, converting them into searchable and extractable text and table data.

What is the best way to redact sensitive PII data across multiple document formats?

Redacting sensitive PII data across multiple document formats is best achieved through API-driven document processing workflows. It supports PDF, DOCX, and XLSX files, automatically identifying and redacting sensitive information.

Does document processing support digital signing and watermarking for PDF files?

Document processing fully supports digital signing and watermarking for PDF files. By leveraging the Nutrient DWS API, you can apply secure digital signatures and custom watermarks to enforce document authenticity and protect intellectual property.

Do I need an API key to extract text from DOCX and PPTX files?

Yes, an API key is a required input to extract text from DOCX and PPTX files. The API-driven processing workflow mandates valid authentication credentials to execute text extraction, OCR, and format conversion tasks.

What are the limitations of using OCR for document processing on scanned images?

Limitations of OCR for document processing on scanned images depend on input quality and API constraints. While it effectively extracts text from various document types, severely degraded scans may result in lower accuracy and require manual verification.