PluginBench
Skill
Pass
Audit score 90

skill-vetter

skills.volces.com

Security-first vetting protocol for AI agent skills before installation.

What is skill-vetter?

Skill Vetter is a security checklist and review framework for evaluating AI agent skills before installing them from any source. Use it to identify red flags, assess permission scope, and classify risk level before running unknown code.

  • Provides a structured source-check protocol to verify skill origin and author reputation
  • Identifies critical red flags in skill code (credential access, obfuscation, external data exfiltration, eval/exec)
  • Evaluates permission scope: what files, network access, and commands a skill requires
  • Classifies risk level from LOW to EXTREME based on capability and scope
  • Generates a standardized vetting report with verdict and notes
  • Includes quick commands to fetch and review GitHub-hosted skills

How to install skill-vetter

npx skills add null --skill skill-vetter
Claude Code
Cursor
Windsurf
Cline

How to use skill-vetter

  1. 1.Obtain the skill source (GitHub URL, ClawdHub link, or local files)
  2. 2.Run Step 1: Source Check—verify author, download count, last update date, and reviews
  3. 3.Run Step 2: Code Review—read all skill files and check against the RED FLAGS list; reject immediately if any flag is found
  4. 4.Run Step 3: Permission Scope—document what files, network, and commands the skill needs
  5. 5.Run Step 4: Risk Classification—assign a risk level (LOW/MEDIUM/HIGH/EXTREME) based on capabilities
  6. 6.Generate the vetting report with verdict (SAFE TO INSTALL / INSTALL WITH CAUTION / DO NOT INSTALL)
  7. 7.Document your vetting decision for future reference

Use cases

Good for
  • Before installing any skill from ClawdHub or GitHub repositories
  • Evaluating skills shared by other agents or team members
  • Assessing third-party skills that request credentials or API keys
  • Reviewing skills with file system or network access before deployment
  • Documenting security decisions for high-risk skill installations
Who it's for
  • AI agents managing their own skill installations
  • Security-conscious developers vetting third-party code
  • Teams with shared skill repositories requiring approval workflows
  • Anyone installing skills from untrusted or unknown sources

skill-vetter FAQ

What should I do if a skill has a red flag?

Reject it immediately. Do not install. Red flags indicate potential security risks like credential theft, data exfiltration, or code injection. No skill is worth compromising security.

How do I vet a GitHub-hosted skill?

Use the Quick Vet Commands provided: fetch repo stats via GitHub API, list skill files, and review the SKILL.md file directly from the raw GitHub URL.

What's the difference between MEDIUM and HIGH risk?

MEDIUM risk skills (file ops, browser access, APIs) require full code review. HIGH risk skills (credentials, trading, system access) require human approval before installation.

Can I install skills from official OpenClaw with lower scrutiny?

Official OpenClaw skills warrant lower scrutiny than unknown sources, but you should still review them. Always maintain some level of verification.

What should I do if I'm unsure about a skill's safety?

When in doubt, don't install. Ask your human for high-risk decisions and document your concerns in the vetting report.

Full instructions (SKILL.md)

Source of truth, from skills.volces.com.


name: skill-vetter version: 1.0.0 description: Security-first skill vetting for AI agents. Use before installing any skill from ClawdHub, GitHub, or other sources. Checks for red flags, permission scope, and suspicious patterns.

Skill Vetter 🔒

Security-first vetting protocol for AI agent skills. Never install a skill without vetting it first.

When to Use

  • Before installing any skill from ClawdHub
  • Before running skills from GitHub repos
  • When evaluating skills shared by other agents
  • Anytime you're asked to install unknown code

Vetting Protocol

Step 1: Source Check

Questions to answer:
- [ ] Where did this skill come from?
- [ ] Is the author known/reputable?
- [ ] How many downloads/stars does it have?
- [ ] When was it last updated?
- [ ] Are there reviews from other agents?

Step 2: Code Review (MANDATORY)

Read ALL files in the skill. Check for these RED FLAGS:

🚨 REJECT IMMEDIATELY IF YOU SEE:
─────────────────────────────────────────
• curl/wget to unknown URLs
• Sends data to external servers
• Requests credentials/tokens/API keys
• Reads ~/.ssh, ~/.aws, ~/.config without clear reason
• Accesses MEMORY.md, USER.md, SOUL.md, IDENTITY.md
• Uses base64 decode on anything
• Uses eval() or exec() with external input
• Modifies system files outside workspace
• Installs packages without listing them
• Network calls to IPs instead of domains
• Obfuscated code (compressed, encoded, minified)
• Requests elevated/sudo permissions
• Accesses browser cookies/sessions
• Touches credential files
─────────────────────────────────────────

Step 3: Permission Scope

Evaluate:
- [ ] What files does it need to read?
- [ ] What files does it need to write?
- [ ] What commands does it run?
- [ ] Does it need network access? To where?
- [ ] Is the scope minimal for its stated purpose?

Step 4: Risk Classification

Risk LevelExamplesAction
🟢 LOWNotes, weather, formattingBasic review, install OK
🟡 MEDIUMFile ops, browser, APIsFull code review required
🔴 HIGHCredentials, trading, systemHuman approval required
⛔ EXTREMESecurity configs, root accessDo NOT install

Output Format

After vetting, produce this report:

SKILL VETTING REPORT
═══════════════════════════════════════
Skill: [name]
Source: [ClawdHub / GitHub / other]
Author: [username]
Version: [version]
───────────────────────────────────────
METRICS:
• Downloads/Stars: [count]
• Last Updated: [date]
• Files Reviewed: [count]
───────────────────────────────────────
RED FLAGS: [None / List them]

PERMISSIONS NEEDED:
• Files: [list or "None"]
• Network: [list or "None"]  
• Commands: [list or "None"]
───────────────────────────────────────
RISK LEVEL: [🟢 LOW / 🟡 MEDIUM / 🔴 HIGH / ⛔ EXTREME]

VERDICT: [✅ SAFE TO INSTALL / ⚠️ INSTALL WITH CAUTION / ❌ DO NOT INSTALL]

NOTES: [Any observations]
═══════════════════════════════════════

Quick Vet Commands

For GitHub-hosted skills:

# Check repo stats
curl -s "https://api.github.com/repos/OWNER/REPO" | jq '{stars: .stargazers_count, forks: .forks_count, updated: .updated_at}'

# List skill files
curl -s "https://api.github.com/repos/OWNER/REPO/contents/skills/SKILL_NAME" | jq '.[].name'

# Fetch and review SKILL.md
curl -s "https://raw.githubusercontent.com/OWNER/REPO/main/skills/SKILL_NAME/SKILL.md"

Trust Hierarchy

  1. Official OpenClaw skills → Lower scrutiny (still review)
  2. High-star repos (1000+) → Moderate scrutiny
  3. Known authors → Moderate scrutiny
  4. New/unknown sources → Maximum scrutiny
  5. Skills requesting credentials → Human approval always

Remember

  • No skill is worth compromising security
  • When in doubt, don't install
  • Ask your human for high-risk decisions
  • Document what you vet for future reference

Paranoia is a feature. 🔒🦀