pdf - PDF Processing Toolkit for Document Manipulation and Analysis
Extract text and tables, create PDFs, merge/split files, and fill forms for programmatic document processing
Tags
Updated: 2026-02-11Capabilities
Typical Inputs
Typical Outputs
What this skill does
- extract text from PDF
- extract tables from PDF
- create PDF documents
- merge PDF files
- split PDF files
- fill PDF forms
- rotate PDF pages
- add watermark to PDF
- extract PDF metadata
Inputs
- PDF files
- text content
- table data
- form templates
- watermark images
Outputs
- processed PDF files
- extracted text files
- extracted table data
- form-filled PDFs
- merged PDF documents
- split PDF pages
Requirements
- Python environment
- pypdf library
- pdfplumber library
- reportlab library
- pytesseract for OCR
