New to Claude Skills? Learn how to install them →

affaan-m on GitHub

Nutrient Document Processing

Free

Efficient document processing and conversion with ease.

Get this skill

Free · Opens the source repo

What Nutrient Document Processing does

Nutrient Document Processing utilizes the Nutrient DWS API to handle a variety of document-related tasks, including format conversion, text extraction, optical character recognition (OCR), and sensitive information redaction. This skill is designed for developers and designers who need to automate document workflows, convert files between different formats, or extract data from scanned documents. With support for multiple file types such as PDF, DOCX, XLSX, PPTX, and various image formats, it offers a comprehensive solution for document management.

The API allows users to convert documents from one format to another seamlessly. For example, you can convert a DOCX file to PDF or extract plain text and tables from a PDF into Excel format. This functionality is particularly useful for those who regularly work with reports, presentations, or any documents that require frequent format changes or data extraction. The OCR capabilities enable users to make scanned documents searchable, supporting over 100 languages, which is essential for organizations dealing with international documents.

Additionally, the skill provides features for editing sensitive information, such as redacting personal identifiable information (PII) like social security numbers or email addresses. This makes it a valuable tool for compliance and privacy-focused tasks. Users can also add watermarks to documents and apply digital signatures, enhancing document security and integrity. Overall, Nutrient Document Processing is a versatile tool that streamlines document handling processes, making it an ideal choice for businesses and individuals looking to improve their document workflows.

When to use it

Use this skill when you need to convert documents between formats, extract data from PDFs, or process scanned documents for OCR.

When not to use it

This skill may not be suitable for basic document viewing or editing tasks that do not require automation or batch processing.

What you can build with it

Convert DOCX to PDF

Easily convert a DOCX file to PDF format for sharing or printing, ensuring compatibility across platforms.

Extract Tables from PDF

Extract structured data from PDF documents into Excel format for analysis or reporting.

Redact Sensitive Information

Automatically redact sensitive information from documents before sharing, ensuring compliance with privacy regulations.

How to install Nutrient Document Processing

View source

1. Install with the skills CLI

npx skills add affaan-m/ecc/nutrient-document-processing --agent claude-code

2. Or install it manually

Download the skill folder and drop it into ~/.claude/skills/ for all projects, or .claude/skills/ to scope it to one repo. Restart Claude Code so it picks up the new skill.

Anthropic's agentic coding CLI, and the reference implementation of Agent Skills. Drop a skill folder into ~/.claude/skills and Claude Code loads it automatically whenever a task matches the skill's description. Claude Code docs

Inside SKILL.md

Written by affaan-m

Nutrient Document Processing

Nutrient DWS Processor API でドキュメントを処理します。フォーマット変換、テキストとテーブルの抽出、スキャンされたドキュメントの OCR、PII の編集、ウォーターマークの追加、デジタル署名、PDF フォームの入力が可能です。

セットアップ

nutrient.io で無料の API キーを取得してください

export NUTRIENT_API_KEY="pdf_live_..."

すべてのリクエストは https://api.nutrient.io/buildinstructions JSON フィールドを含むマルチパート POST として送信されます。

操作

ドキュメントの変換

# DOCX から PDF へ
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.docx=@document.docx" \
  -F 'instructions={"parts":[{"file":"document.docx"}]}' \
  -o output.pdf

# PDF から DOCX へ
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.pdf=@document.pdf" \
  -F 'instructions={"parts":[{"file":"document.pdf"}],"output":{"type":"docx"}}' \
  -o output.docx

# HTML から PDF へ
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "index.html=@index.html" \
  -F 'instructions={"parts":[{"html":"index.html"}]}' \
  -o output.pdf

サポートされている入力形式: PDF、DOCX、XLSX、PPTX、DOC、XLS、PPT、PPS、PPSX、ODT、RTF、HTML、JPG、PNG、TIFF、HEIC、GIF、WebP、SVG、TGA、EPS。

テキストとデータの抽出

# プレーンテキストの抽出
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.pdf=@document.pdf" \
  -F 'instructions={"parts":[{"file":"document.pdf"}],"output":{"type":"text"}}' \
  -o output.txt

# テーブルを Excel として抽出
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.pdf=@document.pdf" \
  -F 'instructions={"parts":[{"file":"document.pdf"}],"output":{"type":"xlsx"}}' \
  -o tables.xlsx

スキャンされたドキュメントの OCR

# 検索可能な PDF への OCR(100以上の言語をサポート)
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "scanned.pdf=@scanned.pdf" \
  -F 'instructions={"parts":[{"file":"scanned.pdf"}],"actions":[{"type":"ocr","language":"english"}]}' \
  -o searchable.pdf

言語: ISO 639-2 コード(例: engdeufraspajpnkorchi_simchi_traarahinrus)を介して100以上の言語をサポートしています。englishgerman などの完全な言語名も機能します。サポートされているすべてのコードについては、完全な OCR 言語表を参照してください。

機密情報の編集

# パターンベース(SSN、メール)
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.pdf=@document.pdf" \
  -F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"redaction","strategy":"preset","strategyOptions":{"preset":"social-security-number"}},{"type":"redaction","strategy":"preset","strategyOptions":{"preset":"email-address"}}]}' \
  -o redacted.pdf

# 正規表現ベース
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.pdf=@document.pdf" \
  -F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"redaction","strategy":"regex","strategyOptions":{"regex":"\\b[A-Z]{2}\\d{6}\\b"}}]}' \
  -o redacted.pdf

プリセット: social-security-numberemail-addresscredit-card-numberinternational-phone-numbernorth-american-phone-numberdatetimeurlipv4ipv6mac-addressus-zip-codevin

ウォーターマークの追加

curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.pdf=@document.pdf" \
  -F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"watermark","text":"CONFIDENTIAL","fontSize":72,"opacity":0.3,"rotation":-45}]}' \
  -o watermarked.pdf

デジタル署名

# 自己署名 CMS 署名
curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "document.pdf=@document.pdf" \
  -F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"sign","signatureType":"cms"}]}' \
  -o signed.pdf

PDF フォームの入力

curl -X POST https://api.nutrient.io/build \
  -H "Authorization: Bearer $NUTRIENT_API_KEY" \
  -F "form.pdf=@form.pdf" \
  -F 'instructions={"parts":[{"file":"form.pdf"}],"actions":[{"type":"fillForm","formFields":{"name":"Jane Smith","email":"jane@example.com","date":"2026-02-06"}}]}' \
  -o filled.pdf

MCP サーバー(代替)

ネイティブツール統合には、curl の代わりに MCP サーバーを使用します:

{
  "mcpServers": {
    "nutrient-dws": {
      "command": "npx",
      "args": ["-y", "@nutrient-sdk/dws-mcp-server"],
      "env": {
        "NUTRIENT_DWS_API_KEY": "YOUR_API_KEY",
        "SANDBOX_PATH": "/path/to/working/directory"
      }
    }
  }
}

使用タイミング

  • フォーマット間でのドキュメント変換(PDF、DOCX、XLSX、PPTX、HTML、画像)
  • PDF からテキスト、テーブル、キー値ペアの抽出
  • スキャンされたドキュメントまたは画像の OCR
  • ドキュメントを共有する前の PII の編集
  • ドラフトまたは機密文書へのウォーターマークの追加
  • 契約または合意書へのデジタル署名
  • プログラムによる PDF フォームの入力

リンク

Frequently asked questions about Nutrient Document Processing

Similar skills