agentsclimarketplace

Pdf processing

Skill ComeOnOliver/skillshub/skills/aiskillstore/marketplace/0xkynz/pdf-processing

🧠 The right skill, one API call. AI agent skills registry with token-efficient skill resolution. 5,000+ skills from 500+ top repos.

Install
npx -y skills add ComeOnOliver/skillshub --skill pdf-processing

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

What its author says it does

Copied from the file, not written here

Extract text and tables from PDF files, fill forms, merge documents. Use when working with PDF files or when the user mentions PDFs, forms, or document extraction.

SKILL.md

1.2 KB, as published. Nobody here has run it

PDF Processing Skill

This skill provides capabilities for working with PDF documents.

Quick Start

Use pdfplumber to extract text from PDFs:

import pdfplumber

with pdfplumber.open("document.pdf") as pdf:
    text = pdf.pages[0].extract_text()

Capabilities

Text Extraction

  • Extract text from single or multiple pages
  • Preserve layout and formatting
  • Handle multi-column documents

Table Extraction

  • Identify and extract tables
  • Convert to structured data (CSV, JSON)
  • Handle complex table layouts

Form Operations

  • Fill PDF forms programmatically
  • Extract form field values
  • Create fillable forms

Document Operations

  • Merge multiple PDFs
  • Split PDFs by page
  • Rotate pages
  • Add watermarks

Best Practices

  1. Always check if the PDF is encrypted before processing
  2. Handle OCR cases for scanned documents
  3. Validate extracted data for accuracy
  4. Use appropriate libraries (pdfplumber for extraction, PyPDF2 for manipulation)

Keep looking

Skills are one crate of 328,083. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.