mirror of
https://github.com/affaan-m/ECC.git
synced 2026-08-21 23:12:24 +02:00
* fix(skills): move version into metadata and normalize to semver 29 skills declared `version` at the top level of their frontmatter. The schema reads it from `metadata`, so tooling that follows the schema either misses it or has to special-case the top level. Three motion skills also declared `version: 1.0`, which is not a valid semantic version; normalized to `1.0.0`. No behavioral change — frontmatter metadata only. * fix(skills): state activation triggers in skill descriptions 148 skills described what they cover but never named the situation that should trigger them. Since the description is what Claude matches against to decide whether to load a skill, a description without a trigger makes activation guesswork — the skill is either missed or loaded at the wrong time. Added a "Use when ..." clause to each, derived from the skill's own body (most already stated the trigger under "## When to Use" or in the opening line; that intent is now reflected in the frontmatter where it is actually read from). Descriptions were only appended to; no existing wording was removed. * fix(skills): sync activation triggers into the Codex skill mirror 10 of the skills whose descriptions changed are also mirrored under `.agents/skills/`, where the description was previously a verbatim copy. Left alone, the two surfaces would disagree about when the skill applies. Only the description line is synced; the Codex copies keep their reduced frontmatter, since that validator accepts only name, description, metadata, license, and allowed-tools. * fix(skills): correct three activation clauses from review - autonomous-loops: the clause pulled new loop work into a skill that its own body marks as a compatibility shim retained for one release. It now points at the canonical continuous-agent-loop instead. - continuous-learning: the description carried the v1 routing directive twice; collapsed to one. - homelab-pihole-dns: the clause fired on any broken home DNS. Narrowed to tasks that actually involve Pi-hole. * chore: retain current main lockfile --------- Co-authored-by: Çağrı Solakoğlu <cagri.solakoglu@vtcenerji.com> Co-authored-by: haelyra <49814733+haelyra@users.noreply.github.com>
169 lines
5.9 KiB
Markdown
169 lines
5.9 KiB
Markdown
---
|
|
name: nutrient-document-processing
|
|
description: Process, convert, OCR, extract, redact, sign, and fill documents using the Nutrient DWS API. Works with PDFs, DOCX, XLSX, PPTX, HTML, and images. Use when converting, OCRing, extracting from, redacting, signing, or filling documents via the Nutrient DWS API.
|
|
metadata:
|
|
origin: ECC
|
|
---
|
|
|
|
# Nutrient Document Processing
|
|
|
|
> **Note:** This skill integrates with the Nutrient commercial API. Review their terms before use.
|
|
|
|
Process documents with the [Nutrient DWS Processor API](https://www.nutrient.io/api/). Convert formats, extract text and tables, OCR scanned documents, redact PII, add watermarks, digitally sign, and fill PDF forms.
|
|
|
|
## Setup
|
|
|
|
Get a free API key at **[nutrient.io](https://dashboard.nutrient.io/sign_up/?product=processor)**
|
|
|
|
```bash
|
|
export NUTRIENT_API_KEY="pdf_live_..."
|
|
```
|
|
|
|
All requests go to `https://api.nutrient.io/build` as multipart POST with an `instructions` JSON field.
|
|
|
|
## Operations
|
|
|
|
### Convert Documents
|
|
|
|
```bash
|
|
# DOCX to PDF
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.docx=@document.docx" \
|
|
-F 'instructions={"parts":[{"file":"document.docx"}]}' \
|
|
-o output.pdf
|
|
|
|
# PDF to DOCX
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.pdf=@document.pdf" \
|
|
-F 'instructions={"parts":[{"file":"document.pdf"}],"output":{"type":"docx"}}' \
|
|
-o output.docx
|
|
|
|
# HTML to PDF
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "index.html=@index.html" \
|
|
-F 'instructions={"parts":[{"html":"index.html"}]}' \
|
|
-o output.pdf
|
|
```
|
|
|
|
Supported inputs: PDF, DOCX, XLSX, PPTX, DOC, XLS, PPT, PPS, PPSX, ODT, RTF, HTML, JPG, PNG, TIFF, HEIC, GIF, WebP, SVG, TGA, EPS.
|
|
|
|
### Extract Text and Data
|
|
|
|
```bash
|
|
# Extract plain text
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.pdf=@document.pdf" \
|
|
-F 'instructions={"parts":[{"file":"document.pdf"}],"output":{"type":"text"}}' \
|
|
-o output.txt
|
|
|
|
# Extract tables as Excel
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.pdf=@document.pdf" \
|
|
-F 'instructions={"parts":[{"file":"document.pdf"}],"output":{"type":"xlsx"}}' \
|
|
-o tables.xlsx
|
|
```
|
|
|
|
### OCR Scanned Documents
|
|
|
|
```bash
|
|
# OCR to searchable PDF (supports 100+ languages)
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "scanned.pdf=@scanned.pdf" \
|
|
-F 'instructions={"parts":[{"file":"scanned.pdf"}],"actions":[{"type":"ocr","language":"english"}]}' \
|
|
-o searchable.pdf
|
|
```
|
|
|
|
Languages: Supports 100+ languages via ISO 639-2 codes (e.g., `eng`, `deu`, `fra`, `spa`, `jpn`, `kor`, `chi_sim`, `chi_tra`, `ara`, `hin`, `rus`). Full language names like `english` or `german` also work. See the [complete OCR language table](https://www.nutrient.io/guides/document-engine/ocr/language-support/) for all supported codes.
|
|
|
|
### Redact Sensitive Information
|
|
|
|
```bash
|
|
# Pattern-based (SSN, email)
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.pdf=@document.pdf" \
|
|
-F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"redaction","strategy":"preset","strategyOptions":{"preset":"social-security-number"}},{"type":"redaction","strategy":"preset","strategyOptions":{"preset":"email-address"}}]}' \
|
|
-o redacted.pdf
|
|
|
|
# Regex-based
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.pdf=@document.pdf" \
|
|
-F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"redaction","strategy":"regex","strategyOptions":{"regex":"\\b[A-Z]{2}\\d{6}\\b"}}]}' \
|
|
-o redacted.pdf
|
|
```
|
|
|
|
Presets: `social-security-number`, `email-address`, `credit-card-number`, `international-phone-number`, `north-american-phone-number`, `date`, `time`, `url`, `ipv4`, `ipv6`, `mac-address`, `us-zip-code`, `vin`.
|
|
|
|
### Add Watermarks
|
|
|
|
```bash
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.pdf=@document.pdf" \
|
|
-F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"watermark","text":"CONFIDENTIAL","fontSize":72,"opacity":0.3,"rotation":-45}]}' \
|
|
-o watermarked.pdf
|
|
```
|
|
|
|
### Digital Signatures
|
|
|
|
```bash
|
|
# Self-signed CMS signature
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "document.pdf=@document.pdf" \
|
|
-F 'instructions={"parts":[{"file":"document.pdf"}],"actions":[{"type":"sign","signatureType":"cms"}]}' \
|
|
-o signed.pdf
|
|
```
|
|
|
|
### Fill PDF Forms
|
|
|
|
```bash
|
|
curl -X POST https://api.nutrient.io/build \
|
|
-H "Authorization: Bearer $NUTRIENT_API_KEY" \
|
|
-F "form.pdf=@form.pdf" \
|
|
-F 'instructions={"parts":[{"file":"form.pdf"}],"actions":[{"type":"fillForm","formFields":{"name":"Jane Smith","email":"jane@example.com","date":"2026-02-06"}}]}' \
|
|
-o filled.pdf
|
|
```
|
|
|
|
## MCP Server (Alternative)
|
|
|
|
For native tool integration, use the MCP server instead of curl:
|
|
|
|
```json
|
|
{
|
|
"mcpServers": {
|
|
"nutrient-dws": {
|
|
"command": "npx",
|
|
"args": ["-y", "@nutrient-sdk/dws-mcp-server"],
|
|
"env": {
|
|
"NUTRIENT_DWS_API_KEY": "YOUR_API_KEY",
|
|
"SANDBOX_PATH": "/path/to/working/directory"
|
|
}
|
|
}
|
|
}
|
|
}
|
|
```
|
|
|
|
## When to Use
|
|
|
|
- Converting documents between formats (PDF, DOCX, XLSX, PPTX, HTML, images)
|
|
- Extracting text, tables, or key-value pairs from PDFs
|
|
- OCR on scanned documents or images
|
|
- Redacting PII before sharing documents
|
|
- Adding watermarks to drafts or confidential documents
|
|
- Digitally signing contracts or agreements
|
|
- Filling PDF forms programmatically
|
|
|
|
## Links
|
|
|
|
- [API Playground](https://dashboard.nutrient.io/processor-api/playground/)
|
|
- [Full API Docs](https://www.nutrient.io/guides/dws-processor/)
|
|
- [npm MCP Server](https://www.npmjs.com/package/@nutrient-sdk/dws-mcp-server)
|