AgentStack
Browse Sign in
Browse Why AgentStack Sell Docs
Sign in
SKILL verified MIT Self-run

Document Skills Pdf

skill-ihatesea69-kiro-kit-pdf · by ihatesea69

Extract and process PDF documents for data pipelines. Use when parsing research papers, extracting tables from reports, or converting PDFs to structured data.

No reviews yet
0 installs
21 views
0.0% view→install

Install

$ agentstack add skill-ihatesea69-kiro-kit-pdf

✓ scanned · ✓ verified, works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access No
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

View the full security report →

Verified badge

Passed review? Show it. Paste this badge into your README, it links to the public security report.

AgentStack Verified badge Links to your public security report.
[![AgentStack Verified](https://agentstack.voostack.com/badges/verified.svg)](https://agentstack.voostack.com/security/report/skill-ihatesea69-kiro-kit-pdf)

Reliability & compatibility

Security review passed
0 installs to date
no reviews yet
2mo ago

Declared compatibility

Claude CodeClaude Desktop

Compatibility is declared by the source manifest. End-to-end runtime verification is coming, see below.

Preview Execution monitoring

We're building live execution health for every listing: tool-call success rate, median latency, uptime, and last-checked timestamps, measured, not self-reported. It isn't live yet, so we don't show numbers we can't stand behind.

How agent discovery & health will work →
Are you the author of Document Skills Pdf? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

Document Skills - PDF

Activate this skill when working with PDF documents in data/AI workflows.

When to Use

  • Extracting tables from financial reports
  • Parsing research papers for literature review
  • Converting scanned PDFs to text (OCR)
  • Extracting metadata from document collections
  • Building document processing pipelines

Libraries

  • pypdf: Read/write PDF, extract text and metadata
  • pdfplumber: Table extraction with spatial awareness
  • PyMuPDF (fitz): Fast rendering and text extraction
  • camelot-py: Table extraction from PDFs
  • pytesseract: OCR for scanned documents

Usage

import pdfplumber

with pdfplumber.open("report.pdf") as pdf:
    for page in pdf.pages:
        tables = page.extract_tables()
        text = page.extract_text()

# OCR for scanned documents
import pytesseract
from pdf2image import convert_from_path

images = convert_from_path("scanned.pdf")
text = pytesseract.image_to_string(images[0])

Rules

  • Check if PDF is text-based or scanned before processing
  • Use pdfplumber for table extraction over regex parsing
  • Handle multi-column layouts carefully
  • Validate extracted numbers against visual inspection
  • Process large PDFs page-by-page to manage memory

Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet, be the first.

Versions

  • v0.1.0 Imported from the upstream source.