pdf-tools
View, extract, edit, and manipulate PDF files. Supports text extraction, text editing (overlay and replacement), merging, splitting, rotating pages, and getting PDF metadata. Use when working with PDF documents for reading content, adding/editing text, reorganizing pages, combining files, or extracting information.
What this skill does
# PDF Tools Tools for viewing, extracting, and editing PDF files using Python libraries (pdfplumber and PyPDF2). ## Quick Start All scripts require dependencies: ```bash pip3 install pdfplumber PyPDF2 ``` ## Core Operations ### Extract Text Extract text from PDF (all pages or specific pages): ```bash scripts/extract_text.py document.pdf scripts/extract_text.py document.pdf -p 1 3 5 scripts/extract_text.py document.pdf -o output.txt ``` ### Get PDF Info View metadata and structure: ```bash scripts/pdf_info.py document.pdf scripts/pdf_info.py document.pdf -f json ``` ### Merge PDFs Combine multiple PDFs into one: ```bash scripts/merge_pdfs.py file1.pdf file2.pdf file3.pdf -o merged.pdf ``` ### Split PDF Split into individual pages: ```bash scripts/split_pdf.py document.pdf -o output_dir/ ``` Split by page ranges: ```bash scripts/split_pdf.py document.pdf -o output_dir/ -m ranges -r "1-3,5-7,10-12" ``` ### Rotate Pages Rotate all pages or specific pages: ```bash scripts/rotate_pdf.py document.pdf -o rotated.pdf -r 90 scripts/rotate_pdf.py document.pdf -o rotated.pdf -r 180 -p 1 3 5 ``` ### Edit Text Add text overlay on a page: ```bash scripts/edit_text.py document.pdf -o edited.pdf --overlay "New Text" --page 1 --x 100 --y 700 scripts/edit_text.py document.pdf -o edited.pdf --overlay "Watermark" --page 1 --x 200 --y 400 --font-size 20 ``` Replace text (limited, works best for simple cases): ```bash scripts/edit_text.py document.pdf -o edited.pdf --replace "Old Text" "New Text" ``` **Note:** PDF text editing is complex due to the format. The overlay method is more reliable than replacement. ## Workflow Patterns ### Viewing PDF Content 1. Get basic info: `scripts/pdf_info.py file.pdf` 2. Extract text to preview: `scripts/extract_text.py file.pdf -p 1` 3. Extract full text if needed: `scripts/extract_text.py file.pdf -o content.txt` ### Reorganizing PDFs 1. Split into pages: `scripts/split_pdf.py input.pdf -o pages/` 2. Merge selected pages: `scripts/merge_pdfs.py pages/page_1.pdf pages/page_3.pdf -o reordered.pdf` ### Extracting Sections 1. Get page count: `scripts/pdf_info.py document.pdf` 2. Split by ranges: `scripts/split_pdf.py document.pdf -o sections/ -m ranges -r "1-5,10-15"` ## Advanced Usage For detailed library documentation and advanced patterns, see [references/libraries.md](references/libraries.md). ## Notes - Page numbers are **1-indexed** in all scripts (page 1 = first page) - Text extraction works best with text-based PDFs (not scanned images) - Rotation angles: 90, 180, 270, or -90 (counterclockwise) - All scripts validate file existence before processing
Related in Writing & Docs
jax-development
IncludedUse this skill when the user is writing, debugging, profiling, refactoring, reviewing, benchmarking, parallelising, exporting, or explaining JAX code, or when they mention JAX, jax.numpy, jit, grad, value_and_grad, vmap, scan, lax, random keys, pytrees, jax.Array, sharding, Mesh, PartitionSpec, NamedSharding, pmap, shard_map, Pallas, XLA, StableHLO, checkify, profiler, or the JAX repo. It helps turn NumPy or PyTorch-style code into pure functional JAX, fix tracer/control-flow/shape/PRNG bugs, remove recompiles and host-device syncs, choose transforms and sharding strategies, inspect jaxpr/lowering/IR, and benchmark compiled code correctly.
nature-article-writer
IncludedDrafts, rewrites, diagnostically critiques, and style-calibrates primary research manuscripts for Nature and Nature Portfolio journals. Use when the user wants a Nature-style title, summary paragraph or abstract, introduction, results, discussion, methods, figure legends, presubmission enquiry, cover letter, reviewer response, or when a scientific draft sounds generic, jargon-heavy, structurally weak, or AI-ish and needs precise, broad-reader-friendly prose without inventing data, analyses, or references. Best for primary research articles and letters rather than reviews or press releases unless explicitly adapting one.
deckrd
IncludedDocument-driven framework that derives requirements, specifications, implementation plans, and executable tasks from goals through structured AI dialogue. Use when user says "write requirements", "create spec", "plan implementation", "derive tasks", "structure this feature", "break down into tasks", or "document this module". Also use for reverse engineering existing code into docs (/deckrd rev). Do NOT use for direct code writing — use /deckrd-coder after tasks are generated. Do NOT use when the user only wants to run or fix existing code without planning.
clinical-decision-support
IncludedGenerate professional clinical decision support (CDS) documents for pharmaceutical and clinical research settings, including patient cohort analyses (biomarker-stratified with outcomes) and treatment recommendation reports (evidence-based guidelines with decision algorithms). Supports GRADE evidence grading, statistical analysis (hazard ratios, survival curves, waterfall plots), biomarker integration, and regulatory compliance. Outputs publication-ready LaTeX/PDF format optimized for drug development, clinical research, and evidence synthesis.
handling-sf-data
IncludedSalesforce data operations with 130-point scoring. Use this skill to create, update, delete, bulk import/export, generate test data, and clean up org records using sf CLI and anonymous Apex. TRIGGER when: user creates test data, performs bulk import/export, uses sf data CLI commands, needs data factory patterns for Apex tests, or needs to seed/clean records in a Salesforce org. DO NOT TRIGGER when: SOQL query writing only (use querying-soql), Apex test execution (use running-apex-tests), or metadata deployment (use deploying-metadata).
accelint-ac-to-playwright
IncludedConvert and validate acceptance criteria for Playwright test automation. Use when user asks to (1) review/evaluate/check if AC are ready for automation, (2) assess if AC can be converted as-is, (3) validate AC quality for Playwright, (4) turn AC into tests, (5) generate tests from acceptance criteria, (6) convert .md bullets or .feature Gherkin files to Playwright specs, (7) create test automation from requirements. Handles both bullet-style markdown and Gherkin syntax with JSON test plan generation and validation.