Install
$ agentstack add skill-shen-shanshan-vllm-dev-skills-vllm-feature-tutorial ✓ scanned · ✓ verified, works with Claude Code, Cursor, and more.
Security review
✓ PassedNo issues found. Passed automated security review. · v0.1.0 How review works →
- ✓ Prompt-injection patterns
- ✓ Secret / credential exfiltration
- ✓ Dangerous shell & filesystem operations
- ✓ Untrusted network calls
- ✓ Known-malicious package signatures
What it can access
- ✓ Network access No
- ✓ Filesystem access No
- ✓ Shell / process execution No
- ✓ Environment & secrets No
- ✓ Dynamic code execution No
From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.
Verified badge
Passed review? Show it. Paste this badge into your README, it links to the public security report.
Reliability & compatibility
Declared compatibility
Compatibility is declared by the source manifest. End-to-end runtime verification is coming, see below.
We're building live execution health for every listing: tool-call success rate, median latency, uptime, and last-checked timestamps, measured, not self-reported. It isn't live yet, so we don't show numbers we can't stand behind.
How agent discovery & health will work →About
vLLM Feature Tutorial Generator
Generate a comprehensive code walkthrough tutorial document for a given vLLM feature or module.
Workflow
Step 1: Identify the Feature
Extract the feature/module name from the user's request. Examples:
- "spec decode" / "speculative decoding"
- "chunked prefill"
- "automatic prefix caching"
- "tensor parallelism"
- "KV cache management"
- "continuous batching"
- "LoRA"
- "multimodal"
Step 2: Research the Feature
Gather information from multiple sources. This is the most important step — thorough research determines document quality.
2a. vLLM Official Documentation
- Fetch relevant pages from
https://docs.vllm.ai/en/latest/using WebFetch - Look for design docs, API references, usage guides
2b. vLLM Source Code
- Use
gh apior WebFetch to browse the vLLM GitHub repo (vllm-project/vllm) - Identify core source files for the feature (use GitHub code search or browse directory structure)
- Read key implementation files to understand:
- Core classes and their responsibilities
- Key interfaces and method signatures
- Data flow and control flow
- Important algorithms and data structures
2c. Related Resources
- Search for relevant blog posts, papers, or design documents
- Check vLLM GitHub discussions/issues for design rationale
Step 3: Generate the Tutorial Document
Read [references/style-guide.md](references/style-guide.md) for the complete document structure and formatting conventions.
Key requirements:
- Write in Chinese (简体中文), keeping English for technical terms
- Follow the multi-part structure defined in the style guide
- Include rich visual elements:
- Mermaid flowcharts for workflows and data flow
- Mermaid architecture diagrams (flowchart TB with subgraph) for system overview
- Mermaid sequence diagrams for component interactions
- Mermaid class diagrams for class hierarchies
- Tables for parameter references, method comparisons, performance metrics
- Code snippets with file path annotations from actual vLLM source
- LaTeX formulas for algorithm analysis (where applicable)
- Include a document header with version info and date
- Include a 文档概述 section with target audience and reading guide
- End with appendices: code location index and glossary
Step 4: Save Output
Save the generated markdown file to outputs/ directory relative to this skill's location:
- File path:
outputs/{feature_name}.md(e.g.,outputs/spec_decode.md,outputs/chunked_prefill.md) - Create the
outputs/directory if it doesn't exist - Use snake_case for file names
The skill directory is located at the same directory as this SKILL.md file.
Content Depth Guidelines
- Part 1 (Basics & Architecture): Start from first principles. Explain the "what" and "why". Include theoretical analysis with formulas if the feature involves algorithms. Provide architecture overview with Mermaid diagrams.
- Part 2 (Core Interfaces): Analyze base classes, interfaces, factory methods. Include class diagrams and method signatures.
- Part 3 (Deep Implementation): Walk through actual vLLM code. Include code snippets with annotations. Trace execution flow with concrete examples.
- Part 4+ (Variants/Comparisons): If the feature has multiple implementations or methods, compare them in tables.
- Final Part (Configuration): Provide practical usage guidance, key parameters, tuning tips.
- Appendix: Code location index table mapping components to file paths.
Quality Checklist
Before saving the document, verify:
- [ ] Document has 3+ Mermaid diagrams (architecture, flow, sequence/class)
- [ ] Document has 3+ comparison/reference tables
- [ ] Document includes actual code snippets from vLLM source with file path annotations
- [ ] Document follows the Chinese writing convention with English technical terms
- [ ] All sections have substantive content (no placeholder text)
- [ ] Document header includes version and date metadata
References
- [references/style-guide.md](references/style-guide.md) — Document structure, formatting conventions, and content patterns
Source & license
This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.
- Author: shen-shanshan
- Source: shen-shanshan/vllm-dev-skills
- License: Apache-2.0
- Homepage: https://zhuanlan.zhihu.com/p/2031696581678866733
Install and usage instructions live in the source repository linked above.
Reviews
No reviews yet, be the first.
Write a review
Versions
- v0.1.0 Imported from the upstream source.