pdf - PDF Processing and Manipulation Toolkit
Extract text and tables, create PDFs, merge/split files, fill forms for document processing
Tags
Updated: 2026-02-11Capabilities
Typical Inputs
Typical Outputs
What this skill does
- read PDF files
- extract text content
- extract table data
- create PDF documents
- merge multiple PDFs
- split PDF into pages
- fill PDF forms
- add watermark to PDF
- encrypt PDF files
- rotate PDF pages
- extract PDF metadata
Inputs
- PDF files
- text content
- table data
- PDF forms
- watermark files
- passwords
Outputs
- modified PDF files
- extracted text files
- extracted table files
- new PDF documents
- merged PDF files
- split PDF pages
- filled PDF forms
- encrypted PDF files
Requirements
- Python environment
- pypdf library
- pdfplumber library
- reportlab library
- poppler-utils for command-line
- pytesseract for OCR
- pdf2image for OCR
