Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence. Use when processing untrusted input (user submissions, API payloads, webhook data, tool outputs) before passing it to a model or executing it.
Install
npx skills add https://github.com/ruvnet/ruflo --skill safety-scanSKILL.md
Safety Scan
Scan content for prompt injection, jailbreak attempts, and unsafe patterns.
When to use
Before processing untrusted input (user submissions, API payloads, webhook data), scan it to detect prompt injection, adversarial content, or policy violations.
Steps
- Quick safety check — call
mcp__plugin_ruflo-core_ruflo__aidefence_is_safewith the input text for a boolean safe/unsafe result - Deep analysis — call
mcp__plugin_ruflo-core_ruflo__aidefence_analyzefor detailed threat classification and confidence scores - Full scan — call
mcp__plugin_ruflo-core_ruflo__aidefence_scanfor comprehensive multi-layer scanning - Train defenses — call
mcp__plugin_ruflo-core_ruflo__aidefence_learnwith confirmed threats to improve detection - View stats — call
mcp__plugin_ruflo-core_ruflo__aidefence_statsfor detection rates and false positive metrics
Threat categories
- Prompt injection (direct and indirect)
- Jailbreak attempts
- Data exfiltration patterns
- Instruction override attacks
- Social engineering prompts
Related skills
azure-compliancemicrosoft606KRun Azure compliance and security audits with azqr plus Key Vault expiration checks. Covers best-practice assessment, resource review, policy/compliance validation, and security posture checks. WHEN: compliance scan, security audit, BEFORE running azqr (compliance cli tool), Azure best practices, Key Vault expiration check, expired certificates, expiring secrets, orphaned resources, compliance assessment.firebase-security-rules-auditorfirebase124KAudits Firebase (Firestore, Cloud Storage) security rules for vulnerabilities, privilege escalation, role bypasses, create vs update inconsistencies, resource exhaustion, type safety, size limits, and hasOnly ownership checks. Use when auditing/reviewing rules, running red-team rule assessments, or scoring against auditor checklists. Don't use for Firebase CLI (login, deploy), Auth, Crashlytics, Remote Config, or database queries.browser-fingerprint-auditliarjsdev69KAudit a browser fingerprint for internal contradictions with the liarjs CLI - canvas, WebGL, WebGL2, WebGPU, audio, 220 fonts, WebRTC and timezone probes, scored against the TLS/HTTP/ASN view of the same request. Use when asked to run a browser fingerprint test, see what a fingerprint looks like, check canvas or WebGL fingerprint stability, compare a spoofed profile against a real browser, or find out whether a browser profile is self-consistent.cloudflare-onecloudflare66KDesign, configure, troubleshoot, or review Cloudflare One Zero Trust and SASE deployments. Use cloudflare-one-migrations for migration planning from other vendors.