Search

LLM guardrails

217 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Set up Claude Code hooks to block dangerous git commands (push, reset --hard, clean, branch -D, etc.) before they execute.

fossasia/eventyay-interpretation1.6k12 repos~578Automated safety check: PassApache-2.0today
2

Query AI Safety HOT news, research papers, incidents, hot topics, and daily/weekly/monthly reports through its public read-only MCP service.

wuyoscar/AISafetyHot-Hub827—~1.4kAutomated safety check: PassUnknowntoday
3

设置 Claude Code hooks,在危险 git commands(push、reset --hard、clean、branch -D 等)执行前阻止它们。适用于用户想防止破坏性 git 操作、添加 git safety hooks,或在 Claude Code 中阻止 git push/reset 时。

vinvcn/mattpocock-skills-zh-CN4.7k—~461Automated safety check: PassMIT2 days ago
4

Remove refusal behaviors from open-weight LLMs using OBLITERATUS — mechanistic interpretability techniques (diff-in-means, SVD, whitened SVD, LEACE, SAE decomposition, etc.) to excise guardrails…

RedWoodOG/Hermes-Desktop1775 repos~3.8kAutomated safety check: PassMIT4 mo ago
5

Behavioral guardrails for LLM-assisted coding. An agent skill from alirezarezvani/ClaudeForge.

alirezarezvani/ClaudeForge429—~1.2kAutomated safety check: PassMIT4 mo ago
6

Systematically reduce the shipped bundle size of a JS/TS library without sacrificing code readability or breaking consumer APIs.

kcsujeet/ilamy-calendar351—~2.9kAutomated safety check: PassMITtoday
7

Interactive scaffold generator for Orloj multi-agent systems.

OrlojHQ/orloj123—~2.6kAutomated safety check: PassApache-2.022 days ago
8

Installs, tunes and enforces Sponsio contracts that block unsafe tool calls in LLM agents, covering setup, auditing, observe mode and flipping to enforce.

SponsioLabs/Sponsio454—~12kAutomated safety check: PassApache-2.0yesterday
9

Turns a natural-language description of routing intent into a valid Lemonade collection.router policy JSON.

amd/skills408—~4kAutomated safety check: PassMITyesterday
10

按中国法规(网安法 / PIPL / 等保2.0 / 数据出境 / AI生成内容标识)审计一个 AI 项目的代码仓库,产出每条都带 文件:行 取证、经独立复核、经脚本校验的合规报告。当用户问「这个项目上线合不合规」「调用了 OpenAI/Claude 算不算数据出境」「要不要做 AI 标识」「帮我做合规自查/等保/PIPL 检查」时使用。Audit an AI project's…

jnMetaCode/shellward140—~1.1kAutomated safety check: PassApache-2.012 days ago
11

Expert ISO 42001 AI Management System (AIMS) compliance advisor.

Sushegaad/Claude-Skills-Governance-Risk-and-Compliance9461 repo~3.7kAutomated safety check: PassMITyesterday
12

Installs hooks that check each agent action against security policies before it runs, blocking destructive commands and logging every decision.

pegasi-ai/reins392—~1.4kAutomated safety check: WarnApache-2.0yesterday
13

Guide for writing eval conversation JSONs and running them through policy engines

open-bias/open-bias143—~1.5kAutomated safety check: PassApache-2.03 days ago
14

A skill your agent uses when you need a deterministic inspection of a WordPress repository (plugin/theme/block theme/WP core/Gutenberg/full site) including tooling/tests/version hints, and a…

gambitph/Stackable3513 repos~371Automated safety check: PassGPL-3.04 days ago
15
15.Wa GuardrailsOfficial

Generate preventive Well-Architected guardrails — AWS Config rules, Service Control Policies, permission boundaries, CloudWatch alarms, and IaC policy checks (CDK Aspects, cfn-guard, OPA/Sentinel) —…

aws-samples/sample-well-architected-skills-and-steering275—~2.8kAutomated safety check: PassMIT-05 days ago
16

Route Celestia requests to the correct repo and apply canonical blob submit/retrieve guidance (Go, Rust, and Node RPC) with docs guardrails.

celestiaorg/docs183—~2.8kAutomated safety check: PassNo licenceyesterday
17

Pick the right SDAF BOM for a target SAP product / release / DB platform / version / kernel / topology.

Azure/sap-automation146—~1.6kAutomated safety check: PassMIT2 days ago
18

设置 Claude Code 钩子,在危险 Git 命令(push、reset --hard、clean、branch -D 等)执行前将其拦截。当用户想要防止破坏性 Git 操作、添加 Git 安全钩子或在 Claude Code 中阻止 git push/reset 时使用。

devcxl/mattpocock-skills-zh449—~420Automated safety check: PassMITyesterday
19

A skill your agent uses when we want to turn a just-finished Formax workflow (e.g.

yusifeng/formax194—~530Automated safety check: PassMIT2 mo ago
20

Ship and spec AI features, LLM products, agents, copilots, and generative UX — including when to use a model vs.

andreaskelm/pm-brain234—~1.8kAutomated safety check: PassUnknownyesterday
21

A skill your agent uses for any question or action about the user's AI/GenAI applications or agents — their behavior, prompts/responses, quality, hallucinations, guardrails, security, cost/tokens…

coralogix/cx-cli121—~2.5kAutomated safety check: PassApache-2.04 days ago
22

Guardrail tester that checks whether the permission rules and PreToolUse hooks already set up in Claude Code, Codex, Gemini CLI, OpenCode, or Cursor stop a battery of dangerous commands, including…

RyanAlberts/best-of-Agent-Harnesses1.1k—~2.9kAutomated safety check: PassMIT2 days ago
23

Check user-generated text against a written policy before it is published.

mrmps/classifier-dev424—~1.5kAutomated safety check: PassMIT3 days ago
24

Guide for creating a new policy engine under openbias/policy/engines/

open-bias/open-bias143—~1.3kAutomated safety check: PassApache-2.03 days ago
25

A skill your agent uses when designing or reviewing a backend MVP with tight budget, evolving schema, and reliance on third-party backends where idempotency, replay, and responsibility attribution…

victorGPT/vibeusage130—~1.3kAutomated safety check: PassMIT2 mo ago
26

A skill your agent uses when the user wants to turn an application, product, startup idea, SaaS, mobile app, web app, API, AI product, or internal tool into a production-ready Markdown specification…

instructa/agent-skills139—~1.5kAutomated safety check: PassNo licence12 days ago
27

Select and run Guardrails verification for code, dependency, test, tooling, documentation, or repository-workflow changes.

openai/openai-guardrails-js105—~619Automated safety check: PassMIT3 days ago
28

How to dispatch well — choosing flash vs pro, writing self-contained briefs, parallelism, verifying results, and guardrails

ZSeven-W/dsh-crew158—~1.1kAutomated safety check: PassMIT2 days ago
29

TypeScript anti-slop guardrails. An agent skill from web-infra-dev/rstest.

web-infra-dev/rstest505—~1.3kAutomated safety check: PassMITyesterday
30

Always-on operational guardrails, model-independent. An agent skill from mrtooher/fable-mode.

mrtooher/fable-mode873—~918Automated safety check: PassNo licencetoday
31

当需要为 AI 产品定义核心身份、角色声明和能力边界时调用此 skill。典型场景包括:设计新 AI 产品的 system prompt 首段、为不同场景创建差异化角色(如教学助手 vs 编程代理)、重新定义 AI 与用户的关系框架。

kangarooking/system-prompt-skills208—~956Automated safety check: PassMIT5 mo ago
32

Step-by-step guide for adding a new guardrail provider to Agent Kernel.

yaalalabs/agent-kernel192—~3.5kAutomated safety check: PassApache-2.0yesterday
33

Evaluate, rank, and communicate work priorities using AI as a structured thinking partner.

ILoveDotNet/ilovedotnet155—~3.5kAutomated safety check: PassCC0-1.04 days ago
34

Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2kAutomated safety check: PassMIT3 mo ago
35

Migrate Python OpenAI Agents SDK applications to Pydantic AI and, when warranted, Pydantic AI Harness.

pydantic/pydantic-ai21k—~1.8kAutomated safety check: PassMITtoday
36

Prepare independent review of a complete Guardrails change before pushing or updating a PR, using the repository adversarial-review procedure.

openai/openai-guardrails-js105—~825Automated safety check: PassMIT3 days ago
37

Build AI applications with OpenAI Agents SDK - text agents, voice agents, multi-agent handoffs, tools with Zod schemas, guardrails, and streaming.

coco-research/coco531—~3.3kAutomated safety check: PassMITtoday
38

A skill your agent uses when authoring or repairing Kilroy Attractor DOT graphs from requirements, with template-first topology, routing guardrails, and validator-clean output.

danshapiro/kilroy222—~6.4kAutomated safety check: PassMIT5 mo ago
39

Diagnose and draft Amazon Canada apparel advertising plans with lifecycle and seasonal timing, English/French search coverage, account evidence, profitability guardrails, and approval-ready…

xjli360/sealeap-amazon-skills251—~1.2kAutomated safety check: PassMIT13 days ago
40

Uses Meta's LlamaGuard moderation model to screen prompts and model replies against six safety categories, with vLLM, FastAPI and NeMo Guardrails setups.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.3kAutomated safety check: PassMIT3 mo ago
41

A skill your agent uses when an instructor wants to create, interview for, configure, install, update, or review a course AI-use policy for MATLAB AI tutoring.

matlab/agent-skills-playground184—~1.2kAutomated safety check: PassUnknownyesterday
42

Report/investigate RUNTIME ACTIVITY of AI agents (Agent 365 / Copilot Studio / M365 Copilot / Work IQ) — agents used, tools/connectors, channels, tokens, prompt/reply content, and Prompt Shield…

SCStelz/security-investigator249—~17kAutomated safety check: PassMIT2 days ago
43

当需要为 AI 系统设计多层安全防线、内容过滤策略和伦理边界时调用此 skill。典型场景包括:设计拒绝策略与升级机制、防御 prompt 注入攻击、实现领域特定安全规则(教育、医疗、金融等)、定义 AI 的价值观锚点。

kangarooking/system-prompt-skills208—~1.2kAutomated safety check: PassMIT5 mo ago
44

Design or review specx core scope boundaries in Python services.

maksimzayats/specx202—~1.8kAutomated safety check: PassMIT2 mo ago
45

OpenClaw 安全部署指南 / Security deployment guide — help users secure their OpenClaw installation

jnMetaCode/shellward140—~644Automated safety check: WarnApache-2.012 days ago
46

AI governance, EU AI Act compliance, OWASP LLM security, responsible AI practices for GitHub Copilot agents

Hack23/cia239—~1.4kAutomated safety check: PassApache-2.0today
47

Run or install repo security leak checks with BetterLeaks and Trivy.

instructa/agent-skills139—~557Automated safety check: PassNo licence12 days ago
48

Generate, edit, and compose images using Gemini Nano Banana models via portable Python scripts.

cnemri/google-genai-skills127—~587Automated safety check: PassMIT8 mo ago