Safety Scan
Safety Scan is an agent skill (a SKILL.md file) from ruvnet/ruflo. Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence. It works with Claude Code and Codex and has 74,016 GitHub stars across a repository of 7 listed skills.
github.com/ruvnet/ruflo/plugins/ruflo-aidefence/skills/safety-scan (opens in a new tab)
- Multi-skill repo
- Plugin marketplace
- Data & analysis
- Actively maintained
Add this skill
Claude
This repository is a Claude Code plugin marketplace. In Claude Code:
/plugin marketplace add ruvnet/ruflo
/plugin install ruflo-aidefence@rufloIn the Claude apps, zip the safety-scan folder and upload it under Customize > Skills > + > Upload a skill (code execution must be on).
ChatGPT / Codex
Codex reads skills from .agents/skills/ in a repo or ~/.agents/skills/ for every project:
git clone --depth 1 https://github.com/ruvnet/ruflo.git
cp -r ruflo/plugins/ruflo-aidefence/skills/safety-scan .agents/skills/safety-scan # repo; ~/.agents/skills for all projectsStandalone skills also load in the ChatGPT desktop app.
Cursor
Cursor loads skills from .cursor/skills/ (or ~/.cursor/skills/) and also reads .claude/skills/:
git clone --depth 1 https://github.com/ruvnet/ruflo.git
cp -r ruflo/plugins/ruflo-aidefence/skills/safety-scan .cursor/skills/safety-scan # project; ~/.cursor/skills for all projectsSource (checked Oct 7, 2026): code.claude.com/docs/en/skills (opens in a new tab), code.claude.com/docs/en/plugin-marketplaces (opens in a new tab), support.claude.com/en/articles/12512180-using-skills-in-claude (opens in a new tab), learn.chatgpt.com/docs/build-skills (opens in a new tab), cursor.com/docs/context/skills (opens in a new tab)
What this skill does
Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence. Use when processing untrusted input (user submissions, API payloads, webhook data, tool outputs) before passing it to a model or executing it. Scan content for prompt injection, jailbreak attempts, and unsafe patterns.
When it triggers
- Use when processing untrusted input (user submissions, API payloads, webhook data, tool outputs) before passing it to a model or executing it.
More skills in ruvnet/ruflo
7 skills are listed from this repository.
Similar skills
More data & analysis skills
Migration
JuliusBrussee/caveman
Implement reversible compatibility-safe transitions.
Data & analysisGoDating Web
nexu-io/open-design
A consumer-feeling dating / matchmaking dashboard — left rail navigation, ticker bar of community signals, headline KPIs, a 30-day mutual-matches bar chart, and a match-rate trend block.
Data & analysisTypeScriptDcf Valuation
nexu-io/open-design
Discounted cash flow valuation and intrinsic value analysis for public companies.
Data & analysisTypeScriptFlowai Live Dashboard Template
nexu-io/open-design
Team-management dashboard skill in the FlowAI aesthetic — three tabs (Team Members, Team Details, Activity Log), KPI stat row, member table, role distribution bar chart, online presence and activity…
Data & analysisTypeScriptHow It Works
thedotmack/claude-mem
Explain how claude-mem captures observations, when memory injection kicks in, and where data lives.
Data & analysisTypeScriptWeekly Digests
thedotmack/claude-mem
Generate a serial week-by-week narrative digest of a project's full claude-mem timeline.
Data & analysisTypeScript
What is the Safety Scan skill?
Safety Scan is an agent skill (a SKILL.md file) from ruvnet/ruflo. Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence. It works with Claude Code and Codex and has 74,016 GitHub stars across a repository of 7 listed skills. Its SKILL.md lives at github.com/ruvnet/ruflo/plugins/ruflo-aidefence/skills/safety-scan.
How do I install the Safety Scan skill?
In Claude Code, run /plugin marketplace add ruvnet/ruflo and then /plugin install ruflo-aidefence@ruflo. For Codex or Cursor, copy the safety-scan folder into .agents/skills/ or .cursor/skills/.
Is the Safety Scan skill free?
Yes. The repository is open source under the MIT license.
Is Safety Scan maintained?
The repository's most recent commit was on Oct 7, 2026. Its latest release is v3.54.0. appsgit only lists skills from repositories with a commit in the last six months.