Topic · Security

Best prompt injection and agent security skills, page 3

Skills #97–144 of 156, ranked by score.

Prompt injection and agent security skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Prompt injection and agent security skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
97

Audit raytsystem changes for prompt injection, provenance bypass, path/symlink/hardlink escape, secret leakage, stale fencing, partial promotion, unsafe parsing, and unapproved side effects.

romarayt/raytsystem-public-os149—~572Automated safety check: PassApache-2.02 days ago
98

Extract structured data via stored browser-templates or one-shot DOM queries, with mandatory AIDefence PII + prompt-injection gates before content reaches the model

ruvnet/ruflo74k—~888Automated safety check: NotesMITtoday
99

Threat-model and harden AI agents, RAG systems, assistants, and tool-using workflows against direct, indirect, stored, cross-agent, and multimodal prompt injection.

seb1n/awesome-ai-agent-skills206—~2.6kAutomated safety check: PassMIT2 mo ago
100

Verify exact SillyTavern, Tavern Helper / JS-Slash-Runner, STScript, macro, prompt-injection, worldbook, EJS, MVU, and runtime-library capabilities before implementing or reviewing rolecard…

LiarMTTT/TavernWeave154—~2.4kAutomated safety check: PassUnknown6 days ago
101

Use before committing changes to auth, workspace scoping, channel webhooks, AI tools/MCP, permission settings, or anything handling untrusted channel content in ChatbotX.

ChatbotXIO/ChatbotX885—~1.8kAutomated safety check: NotesUnknownyesterday
102

Defend AI systems against prompt injection and indirect prompt attacks using input controls, tool permissions, output validation, and isolation boundaries.

sickn33/agentic-awesome-skills47k2 repos~4.2kAutomated safety check: WarnMITyesterday
103

Scan inputs for prompt injection, unsafe content, and adversarial attacks using AIDefence.

ruvnet/ruflo74k—~409Automated safety check: NotesMITtoday
104

Screen fetched web pages, tool results, emails and file contents for instructions aimed at the agent before they enter context, using a keyless multi-label classifier over chunks.

mrmps/classifier-dev424—~1.5kAutomated safety check: PassMIT3 days ago
105

[omh] Agent or automation safety risks: review prompt, tool, secret, dependency, destructive-action, and explicit local plugin risks before agent or code execution.

rlaope/oh-my-hermes3.2k—~2.3kAutomated safety check: PassMITtoday
106

AI/LLM defensive security reference: prompt-injection defense, OWASP LLM Top 10 defensive mapping, MCP and agentic tool-call hardening, training-data poisoning detection, model-output validation and…

modu-ai/moai-adk1.2k—~4.5kAutomated safety check: PassApache-2.0today
107

DevSecOps, container, and API operational defensive security reference: CI/CD pipeline hardening, secret scanning, IaC misconfiguration detection, SAST/DAST integration, container image scanning…

modu-ai/moai-adk1.2k—~2.6kAutomated safety check: PassApache-2.0today
108
108.Stack

Run and inspect the local pieces of Guardana — the throwaway PostgreSQL for the collector, the collector itself, a fake OpenAI-compatible endpoint to probe, the documentation site served locally…

guardana/guardana131—~568Automated safety check: PassApache-2.0today
109

Authorized security assessment of LLM applications and AI agents: prompt injection, tool abuse, RAG exposure, memory poisoning, system-prompt extraction, and agent-compliance engineering per OWASP…

sickn33/agentic-awesome-skills47k1 repo~1.2kAutomated safety check: WarnMITyesterday
110

Detects prompt injection using regex signature matching, heuristic scoring for structural anomalies, and DeBERTa-based transformer classification, flagging direct injections (system-prompt…

mukul975/Anthropic-Cybersecurity-Skills34k—~1.9kAutomated safety check: WarnApache-2.01 mo ago
111

Detect and defend against indirect prompt injection hidden in web pages, documents, and images consumed by an agent, via content extraction (HTML/PDF/OCR), normalization, and scanning with LLM…

mukul975/Anthropic-Cybersecurity-Skills34k—~2.8kAutomated safety check: WarnApache-2.01 mo ago
112

Runs NVIDIA garak probe suites (jailbreak, prompt injection, data leakage, toxicity, and more) against an LLM endpoint - Hugging Face models, OpenAI-compatible APIs, or Bedrock - then interprets the…

mukul975/Anthropic-Cybersecurity-Skills34k—~2.9kAutomated safety check: WarnApache-2.01 mo ago
113

Deploys Llama Guard 3 safety classification, NeMo Guardrails programmable dialogue rails, and LLM Guard input/output scanner pipelines as complementary runtime defenses that inspect and constrain…

mukul975/Anthropic-Cybersecurity-Skills34k—~3.1kAutomated safety check: WarnApache-2.01 mo ago
114

Runtime security guardian for OpenClaw agents. An agent skill from LeoYeAI/openclaw-master-skills.

LeoYeAI/openclaw-master-skills2.2k—~1.7kAutomated safety check: PassMIT2 mo ago
115

This skill should be used when the user asks to "scan AI systems for security threats", "check for prompt injection vulnerabilities", "assess model security posture", "detect data poisoning risks"…

borghei/Claude-Skills891—~1.1kAutomated safety check: PassMIT3 days ago
116

A skill your agent uses when assessing AI/ML systems for prompt injection, jailbreak vulnerabilities, model inversion risk, data poisoning exposure, or agent tool abuse.

alirezarezvani/claude-skills28k—~4.5kAutomated safety check: WarnMIT1 mo ago
117

Consolidate brain memory and mine user sessions since the last consolidation checkpoint (sleep-cycle style).

mikeyobrien/rho373—~1.3kAutomated safety check: PassMIT9 days ago
118

Binance Web3 official skill — security audit for token contracts, detecting honeypots, rug pulls, and malicious functions across BSC, Base, Solana, and Ethereum.

TermiX-official/cryptoclaw100—~695Automated safety check: PassMIT4 mo ago
119

Pull AWS Security Agent findings (penetration tests and code reviews) and drive remediation.

aws/agent-toolkit-for-aws2.8k—~2.9kAutomated safety check: PassApache-2.0today
120

Behavioral trust grades (A–F) for MCP servers. An agent skill from BankrBot/skills.

BankrBot/skills1.2k—~3.4kAutomated safety check: PassNo licencetoday
121

AI Agent 安全开发与防护最佳实践,包含prompt注入防护、代码执行安全、敏感信息保护、合规审计全流程规范. An agent skill from ProgrammerAnthony/Expert-Coding-Harness.

ProgrammerAnthony/Expert-Coding-Harness235—~3kAutomated safety check: PassMIT5 mo ago
122

LLM / AI application attack hunting - prompt injection (direct + indirect), excessive agency, insecure output handling, system-prompt + data leakage.

Encod3d-Sec/TORCH329—~1.7kAutomated safety check: PassMIT1 mo ago
123

MCP server attack hunting - tool poisoning, indirect prompt injection via tool output, rug-pull updates, cross-tool shadowing, over-permissioned/excessive-agency tools, lethal trifecta.

Encod3d-Sec/TORCH329—~1.4kAutomated safety check: PassMIT1 mo ago
124

Test LLM-integrated applications against known prompt injection techniques, evasion methods, and attack intents using the Arcanum PI Taxonomy.

OWASP/secure-agent-playbook188—~758Automated safety check: PassCC-BY-4.015 days ago
125

Audit an LLM application for indirect prompt injection - compose objective x technique payloads, deliver them through the channels the agent actually reads, and prove impact with an out-of-band…

forefy/.context152—~2.3kAutomated safety check: PassMIT5 days ago
126

Defensive security engineering judgment, distilled from a stronger model - invoke when THREAT MODELING a system or feature; making security-relevant design decisions (auth, crypto, trust boundaries…

telagod/code-abyss244—~907Automated safety check: PassMIT2 mo ago
127

Audit MCP servers for tool poisoning, tool shadowing, rug pulls, SSRF, and unauthenticated exposure using Invariant Labs' mcp-scan for static/runtime scanning plus manual SSRF/auth checks and…

mukul975/Anthropic-Cybersecurity-Skills34k—~2.7kAutomated safety check: WarnApache-2.01 mo ago
128

Test whether authorized direct or indirect untrusted content can alter an AI system's protected behavior, context use, memory, retrieval, output handling, or downstream capability.

cyberful/cyberful135—~538Automated safety check: PassAGPL-3.01 mo ago
129

Reconstruct how instructions, retrieved content, memory, identities, approvals, tool schemas, arguments, outputs, delegation, and fallbacks propagate through an AI system.

cyberful/cyberful135—~517Automated safety check: PassAGPL-3.01 mo ago
130
130.Debug

Systematic diagnosis of a red test, a rule that fires or stays silent wrongly, a scan or probe whose artifact looks wrong, a collector error, a red CI run or a gate that is green for the wrong reason.

guardana/guardana131—~933Automated safety check: PassApache-2.0today
131

Implement content safety guardrails for Claude — input filtering, Use when working with policy-guardrails patterns.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: PassMITtoday
132

Secure your Anthropic integration — API key management, input validation, Use when working with security-basics patterns.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: NotesMITtoday
133

Security and compliance review framework for Kling AI integrations.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.8kAutomated safety check: PassMITtoday
134

Audit a Claude/Agent SKILL.md (or any AI skill / system prompt) for safety before installing or merging it.

mohitagw15856/pm-claude-skills1.4k—~1.5kAutomated safety check: PassMITyesterday
135

AI/LLM 间接 Prompt 注入攻击。当目标 AI 系统会处理外部数据源(网页、文档、邮件、数据库、API 返回)时使用。覆盖间接注入、工具链劫持、RAG 投毒、数据外泄等技术。OWASP LLM Top 10 1 漏洞类别

wgpsec/AboutSecurity1.8k—~704Automated safety check: NotesNo licencetoday
136

Adversarial defense layer for the mortgage plugin — protects against prompt injection, system prompt extraction, PII leakage, workflow bypass, and social engineering attacks.

davepoon/buildwithclaude3.6k—~455Automated safety check: PassMITyesterday
137

服务上线前的 AI 安全审查,聚焦需要理解代码语义和业务逻辑的安全问题. An agent skill from 312362115/claude.

312362115/claude107—~1.3kAutomated safety check: PassMIT4 mo ago
138

A skill your agent uses when auditing an ICML main-track submission for OpenReview, LaTeX formatting, 8-page body, anonymity, supplementary material, impact statement, dual submission, concurrent…

franklee16/academic-research-skills2231 repo~655Automated safety check: PassNo licence22 days ago
139

Detects prompt injection hidden in documents from the other side (pleadings, skeletons, bundles, served evidence, opponents' emails and attachments) before an AI reads them, so the model is not…

lawve-ai/awesome-legal-skills847—~4.6kAutomated safety check: PassMIT7 days ago
140

Route broad AI-agent security reviews to focused Cyberful skills across risk, model supply chain, context and capabilities, prompt injection, tool authorization, and RAG isolation.

cyberful/cyberful135—~723Automated safety check: PassAGPL-3.01 mo ago
141

Guide an OpenART agent or contributor through planning, running, extending, and debugging the framework.

AI45Lab/OpenART235—~918Automated safety check: NotesAGPL-3.07 days ago
142
142.Bagman

Secure key management for AI agents. An agent skill from LeoYeAI/openclaw-master-skills.

LeoYeAI/openclaw-master-skills2.2k—~2.9kAutomated safety check: NotesMIT2 mo ago
143

Build AI agents with console.agent() - the jQuery of AI Agents.

LeoYeAI/openclaw-master-skills2.2k—~4.1kAutomated safety check: PassMIT2 mo ago
144

A skill your agent uses when enforcing spend limits on AI agent wallets, validating transactions before signing, configuring allowlists or approval workflows, detecting prompt injection in agent…

LeoYeAI/openclaw-master-skills2.2k—~5kAutomated safety check: NotesMIT2 mo ago