Install
$ agentstack add skill-liu-zhangzhu-opencode-vision-skill-opencode-vision-skill ✓ scanned · ✓ verified, works with Claude Code, Cursor, and more.
Security review
✓ PassedNo issues found. Passed automated security review. · v0.1.0 How review works →
- ✓ Prompt-injection patterns
- ✓ Secret / credential exfiltration
- ✓ Dangerous shell & filesystem operations
- ✓ Untrusted network calls
- ✓ Known-malicious package signatures
What it can access
- ✓ Network access No
- ✓ Filesystem access No
- ✓ Shell / process execution No
- ✓ Environment & secrets No
- ✓ Dynamic code execution No
From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.
Verified badge
Passed review? Show it. Paste this badge into your README, it links to the public security report.
Reliability & compatibility
Declared compatibility
Compatibility is declared by the source manifest. End-to-end runtime verification is coming, see below.
We're building live execution health for every listing: tool-call success rate, median latency, uptime, and last-checked timestamps, measured, not self-reported. It isn't live yet, so we don't show numbers we can't stand behind.
How agent discovery & health will work →About
Vision Reader Skill
Enables pure-text LLMs (DeepSeek, etc.) to read binary files by calling Python tools.
Install Path
All tools are located at ~/opencode-vision-skill/tools/. The ~ means your home directory:
- Windows:
%USERPROFILE%\opencode-vision-skill\tools\ - macOS/Linux:
~/opencode-vision-skill/tools/
If you installed elsewhere, adjust the paths below accordingly.
Available Tools
| Tool | Command | |------|---------| | pdf-reader | python ~/opencode-vision-skill/tools/pdf-reader.py | | image-reader | python ~/opencode-vision-skill/tools/image-reader.py | | ppt-reader | python ~/opencode-vision-skill/tools/ppt-reader.py | | screenshot | python ~/opencode-vision-skill/tools/screenshot.py [x,y,w,h] |
Always use absolute file paths for the filepath argument.
Tool Details
pdf-reader — Read PDF files
python ~/opencode-vision-skill/tools/pdf-reader.py "/absolute/path/to/file.pdf"
Returns Markdown with page headings, text content, and tables.
image-reader — Read and describe images
python ~/opencode-vision-skill/tools/image-reader.py "/absolute/path/to/image.png"
Returns three sections:
- Metadata — Format, mode, size, DPI
- Visual Description — AI-generated caption describing objects, layout, arrows, diagrams
- OCR Text — Recognized text from the image
ppt-reader — Read PowerPoint files
python ~/opencode-vision-skill/tools/ppt-reader.py "/absolute/path/to/slides.pptx"
Returns slide-by-slide text with titles, bullet points, and tables.
screenshot — Capture the screen
python ~/opencode-vision-skill/tools/screenshot.py
Captures full screen. For a region: python ~/opencode-vision-skill/tools/screenshot.py 0,0,800,600. Returns the saved PNG file path.
Notes
- The first
image-readercall downloads Florence-2 model (~400MB) if not pre-downloaded - Tesseract OCR is optional — image-reader works without it using easyocr
- Screenshot requires a display server (X11/Wayland on Linux, native on Windows/macOS)
Source & license
This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.
- Author: Liu-Zhangzhu
- Source: Liu-Zhangzhu/opencode-vision-skill
- License: MIT
Install and usage instructions live in the source repository linked above.
Reviews
No reviews yet, be the first.
Write a review
Versions
- v0.1.0 Imported from the upstream source.