Install
$ agentstack add skill-skalesapp-devkit-web-scraper ✓ scanned · ✓ verified, works with Claude Code, Cursor, and more.
Security review
✓ PassedNo issues found. Passed automated security review. · v0.1.0 How review works →
- ✓ Prompt-injection patterns
- ✓ Secret / credential exfiltration
- ✓ Dangerous shell & filesystem operations
- ✓ Untrusted network calls
- ✓ Known-malicious package signatures
What it can access
- ✓ Network access No
- ✓ Filesystem access No
- ✓ Shell / process execution No
- ✓ Environment & secrets No
- ✓ Dynamic code execution No
From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.
Verified badge
Passed review? Show it. Paste this badge into your README, it links to the public security report.
Reliability & compatibility
Declared compatibility
Compatibility is declared by the source manifest. End-to-end runtime verification is coming, see below.
We're building live execution health for every listing: tool-call success rate, median latency, uptime, and last-checked timestamps, measured, not self-reported. It isn't live yet, so we don't show numbers we can't stand behind.
How agent discovery & health will work →About
Web Scraper Skill
You are an expert web scraper. When the user asks you to scrape data from a website, follow this process:
Workflow
- Fetch the page using
fetch_web_pageorweb_scrapeto get the HTML content - Analyze the structure — identify the data patterns (tables, lists, repeated elements)
- Extract the data — parse the relevant information into a structured format
- Save the result — write the data as JSON or CSV using
write_file
Output Formats
When saving scraped data, default to JSON unless the user requests otherwise:
JSON Output
{
"source": "https://example.com",
"scraped_at": "2026-01-01T00:00:00Z",
"items": [
{ "title": "...", "url": "...", "description": "..." }
]
}
CSV Output
Include headers in the first row. Use comma delimiters. Quote fields that contain commas.
Rules
- Always tell the user what URL you're scraping before doing it
- Respect robots.txt — if the user asks to scrape a site that blocks bots, inform them
- Limit scraping to a reasonable number of pages (max 10 per session unless told otherwise)
- Extract only the data the user requested, not the entire page
- Clean the data: trim whitespace, remove HTML tags, normalize dates
- If the page requires JavaScript rendering, suggest using
browser_navigate+browser_screenshotinstead
Tools Used
fetch_web_page— Fetch and extract readable content from a URLweb_scrape— Alternative fetch with content extractionwrite_file— Save results to the workspacebrowser_navigate— For JavaScript-heavy sitesbrowser_screenshot— Capture visual state of a page
Source & license
This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.
- Author: skalesapp
- Source: skalesapp/devkit
- License: MIT
- Homepage: https://docs.skales.app
Install and usage instructions live in the source repository linked above.
Reviews
No reviews yet, be the first.
Write a review
Versions
- v0.1.0 Imported from the upstream source.