Skill Market

safety scan

Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence

GitHub
githubcommunityclaudemcp
0.0
0 installs73.4K GitHub starsby ruvnet

Skill Introduction

Overview
Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence

Core value

Turns reusable Other know-how into an installable skill, helping users complete github, community, claude, mcp work faster.

Target users

  • Developers, testers, and maintainers who handle Other tasks in Focus Code.
  • Teams that already trust workflows or content from ruvnet.
  • Users who want standardized prompts, steps, or conventions instead of repeating setup work.

Best practices

  • Read the skill content first to confirm required inputs, expected outputs, and dependencies.
  • Try it on a small task before relying on it for critical work.
  • Add project-specific constraints such as coding style, target platform, test expectations, and delivery format.
  • For external sources, verify the source link, version, and recent maintenance activity.

Best use cases

  • Tasks related to github, community, claude, mcp that need a reusable execution flow.
  • Converting a community repo, team convention, or personal workflow into day-to-day assistance.
  • Starting from a proven skill instead of writing prompts or procedures from scratch.

Limits and boundaries

  • Results depend on the quality of the original skill content and may need human correction.
  • It does not replace code review, tests, security review, or professional judgment.
  • External tools, APIs, account permissions, and local dependencies still need separate setup.

Differentiation

  • Structured around Other, making it easier to discover and reuse than loose prompt snippets.
  • Marked as GitHub, which helps users judge trust and maintenance expectations.
  • Keeps the original source link available for repository, documentation, or discussion follow-up.
  • Tagged with github, community, claude, mcp, so it can be filtered by concrete task intent.

Install and use

Install
Copy Install Command
focus install safety-scan-81bfdf
View source

Detail Preview

SKILL.md

Primary filemarkdown2 KB

name: safety-scan description: Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence argument-hint: "<input-text>" allowed-tools: mcp__claude-flow__aidefence_scan mcp__claude-flow__aidefence_analyze mcp__claude-flow__aidefence_is_safe mcp__claude-flow__aidefence_learn mcp__claude-flow__aidefence_stats Bash

Safety Scan

Scan content for prompt injection, jailbreak attempts, and unsafe patterns.

When to use

Before processing untrusted input (user submissions, API payloads, webhook data), scan it to detect prompt injection, adversarial content, or policy violations.

Steps

  1. Quick safety check — call mcp__claude-flow__aidefence_is_safe with the input text for a boolean safe/unsafe result
  2. Deep analysis — call mcp__claude-flow__aidefence_analyze for detailed threat classification and confidence scores
  3. Full scan — call mcp__claude-flow__aidefence_scan for comprehensive multi-layer scanning
  4. Train defenses — call mcp__claude-flow__aidefence_learn with confirmed threats to improve detection
  5. View stats — call mcp__claude-flow__aidefence_stats for detection rates and false positive metrics

Threat categories

  • Prompt injection (direct and indirect)
  • Jailbreak attempts
  • Data exfiltration patterns
  • Instruction override attacks
  • Social engineering prompts

Reviews

Overall rating

0.0
0.0

0 comments

No reviews yet