Topic · AI & LLM Engineering

Best LLM guardrails skills for Claude Code, Codex and other agents.

Skills that add input and output guardrails to language-model applications.
skills
208
official
22

LLM guardrails skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

LLM guardrails skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Set up Claude Code hooks to block dangerous git commands (push, reset --hard, clean, branch -D, etc.) before they execute.

fossasia/eventyay-interpretation1.6k12 repos~578Automated safety check: PassApache-2.02 days ago
2

Remove refusal behaviors from open-weight LLMs using OBLITERATUS — mechanistic interpretability techniques (diff-in-means, SVD, whitened SVD, LEACE, SAE decomposition, etc.) to excise guardrails…

RedWoodOG/Hermes-Desktop1776 repos~3.8kAutomated safety check: PassMIT4 mo ago
3

设置 Claude Code hooks,在危险 git commands(push、reset --hard、clean、branch -D 等)执行前阻止它们。适用于用户想防止破坏性 git 操作、添加 git safety hooks,或在 Claude Code 中阻止 git push/reset 时。

vinvcn/mattpocock-skills-zh-CN4.6k—~474Automated safety check: PassMIT9 days ago
4

Behavioral guardrails for LLM-assisted coding. An agent skill from alirezarezvani/ClaudeForge.

alirezarezvani/ClaudeForge429—~1.2kAutomated safety check: PassMIT4 mo ago
5

Systematically reduce the shipped bundle size of a JS/TS library without sacrificing code readability or breaking consumer APIs.

kcsujeet/ilamy-calendar351—~2.9kAutomated safety check: PassMIT3 days ago
6

Interactive scaffold generator for Orloj multi-agent systems.

OrlojHQ/orloj123—~2.6kAutomated safety check: PassApache-2.019 days ago
7

Read AI Safety HOT daily digests, search recent AI safety research and incidents, and follow current hot topics.

wuyoscar/AISafetyHot-Hub175—~1.2kAutomated safety check: PassUnknowntoday
8

Installs, tunes and enforces Sponsio contracts that block unsafe tool calls in LLM agents, covering setup, auditing, observe mode and flipping to enforce.

SponsioLabs/Sponsio440—~12kAutomated safety check: PassApache-2.0yesterday
9

Checks each shell command for destructive patterns such as recursive deletes, force pushes and dropped tables, and asks before letting them run.

garrytan/gstack136k—~931Automated safety check: NotesMITtoday
10

Turns a natural-language description of routing intent into a valid Lemonade collection.router policy JSON.

amd/skills395—~4kAutomated safety check: PassMITtoday
11

当需要为 AI 产品定义核心身份、角色声明和能力边界时调用此 skill。典型场景包括:设计新 AI 产品的 system prompt 首段、为不同场景创建差异化角色(如教学助手 vs 编程代理)、重新定义 AI 与用户的关系框架。

kangarooking/system-prompt-skills2051 repo~956Automated safety check: PassMIT5 mo ago
12

按中国法规(网安法 / PIPL / 等保2.0 / 数据出境 / AI生成内容标识)审计一个 AI 项目的代码仓库,产出每条都带 文件:行 取证、经独立复核、经脚本校验的合规报告。当用户问「这个项目上线合不合规」「调用了 OpenAI/Claude 算不算数据出境」「要不要做 AI 标识」「帮我做合规自查/等保/PIPL 检查」时使用。Audit an AI project's…

jnMetaCode/shellward140—~1.1kAutomated safety check: PassApache-2.08 days ago
13

Expert ISO 42001 AI Management System (AIMS) compliance advisor.

Sushegaad/Claude-Skills-Governance-Risk-and-Compliance9391 repo~3.7kAutomated safety check: PassMIT2 days ago
14

Designs, tests and refines LLM prompts: zero-shot, few-shot and chain-of-thought patterns, system prompts, structured output schemas and evaluation test suites.

Jeffallan/claude-skills12k1 repo~1.5kAutomated safety check: PassMIT4 days ago
15

Always-on operational guardrails, model-independent. An agent skill from mrtooher/fable-mode.

mrtooher/fable-mode870—~1kAutomated safety check: PassNo licence1 mo ago
16

Installs hooks that check each agent action against security policies before it runs, blocking destructive commands and logging every decision.

pegasi-ai/reins392—~1.4kAutomated safety check: WarnApache-2.04 mo ago
17

Guide for writing eval conversation JSONs and running them through policy engines

open-bias/open-bias143—~1.5kAutomated safety check: PassApache-2.04 mo ago
18

A skill your agent uses when you need a deterministic inspection of a WordPress repository (plugin/theme/block theme/WP core/Gutenberg/full site) including tooling/tests/version hints, and a…

gambitph/Stackable3503 repos~371Automated safety check: PassGPL-3.0today
19
19.Wa GuardrailsOfficial

Generate preventive Well-Architected guardrails — AWS Config rules, Service Control Policies, permission boundaries, CloudWatch alarms, and IaC policy checks (CDK Aspects, cfn-guard, OPA/Sentinel) —…

aws-samples/sample-well-architected-skills-and-steering273—~2.8kAutomated safety check: PassMIT-0yesterday
20

Route Celestia requests to the correct repo and apply canonical blob submit/retrieve guidance (Go, Rust, and Node RPC) with docs guardrails.

celestiaorg/docs183—~2.7kAutomated safety check: PassNo licencetoday
21

Pick the right SDAF BOM for a target SAP product / release / DB platform / version / kernel / topology.

Azure/sap-automation145—~1.6kAutomated safety check: PassMITtoday
22

A skill your agent uses when we want to turn a just-finished Formax workflow (e.g.

yusifeng/formax195—~530Automated safety check: PassMIT2 mo ago
23

设置 Claude Code 钩子,在危险 Git 命令(push、reset --hard、clean、branch -D 等)执行前将其拦截。当用户想要防止破坏性 Git 操作、添加 Git 安全钩子或在 Claude Code 中阻止 git push/reset 时使用。

devcxl/mattpocock-skills-zh433—~420Automated safety check: PassMITyesterday
24

A skill your agent uses for any question or action about the user's AI/GenAI applications or agents — their behavior, prompts/responses, quality, hallucinations, guardrails, security, cost/tokens…

coralogix/cx-cli121—~2.5kAutomated safety check: PassApache-2.02 days ago
25

Check user-generated text against a written policy before it is published.

mrmps/classifier-dev424—~1.5kAutomated safety check: PassMITyesterday
26

Enforces staged execution discipline on large tasks: a written stage plan, delegation to named fable agents where the runtime supports it, a failable verification check at each stage, and a…

mrtooher/fable-mode870—~1kAutomated safety check: PassNo licence1 mo ago
27

Guide for creating a new policy engine under openbias/policy/engines/

open-bias/open-bias143—~1.3kAutomated safety check: PassApache-2.04 mo ago
28

A skill your agent uses when designing or reviewing a backend MVP with tight budget, evolving schema, and reliance on third-party backends where idempotency, replay, and responsibility attribution…

victorGPT/vibeusage131—~1.3kAutomated safety check: PassMIT2 mo ago
29

A skill your agent uses when the user wants to turn an application, product, startup idea, SaaS, mobile app, web app, API, AI product, or internal tool into a production-ready Markdown specification…

instructa/agent-skills139—~1.5kAutomated safety check: PassNo licence8 days ago
30

Select and run Guardrails verification for code, dependency, test, tooling, documentation, or repository-workflow changes.

openai/openai-guardrails-js104—~619Automated safety check: PassMIT9 days ago
31

How to dispatch well — choosing flash vs pro, writing self-contained briefs, parallelism, verifying results, and guardrails

ZSeven-W/dsh-crew156—~1.1kAutomated safety check: PassMITtoday
32

Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2kAutomated safety check: PassMIT3 mo ago
33

TypeScript anti-slop guardrails. An agent skill from web-infra-dev/rstest.

web-infra-dev/rstest505—~1.3kAutomated safety check: PassMIT7 days ago
34

Step-by-step guide for adding a new guardrail provider to Agent Kernel.

yaalalabs/agent-kernel191—~3.5kAutomated safety check: PassApache-2.0today
35

Evaluate, rank, and communicate work priorities using AI as a structured thinking partner.

ILoveDotNet/ilovedotnet155—~3.5kAutomated safety check: PassCC0-1.0yesterday
36

Migrate Python OpenAI Agents SDK applications to Pydantic AI and, when warranted, Pydantic AI Harness.

pydantic/pydantic-ai20k—~1.8kAutomated safety check: PassMITtoday
37

Uses Meta's LlamaGuard moderation model to screen prompts and model replies against six safety categories, with vLLM, FastAPI and NeMo Guardrails setups.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.3kAutomated safety check: PassMIT3 mo ago
38

Prepare independent review of a complete Guardrails change before pushing or updating a PR, using the repository adversarial-review procedure.

openai/openai-guardrails-js104—~825Automated safety check: PassMIT9 days ago
39

A skill your agent uses when an instructor wants to create, interview for, configure, install, update, or review a course AI-use policy for MATLAB AI tutoring.

matlab/agent-skills-playground181—~1.2kAutomated safety check: PassUnknown26 days ago
40

A skill your agent uses when authoring or repairing Kilroy Attractor DOT graphs from requirements, with template-first topology, routing guardrails, and validator-clean output.

danshapiro/kilroy221—~6.4kAutomated safety check: PassMIT5 mo ago
41

Diagnose and draft Amazon Canada apparel advertising plans with lifecycle and seasonal timing, English/French search coverage, account evidence, profitability guardrails, and approval-ready…

xjli360/sealeap-amazon-skills237—~1.2kAutomated safety check: PassMIT9 days ago
42

Report/investigate RUNTIME ACTIVITY of AI agents (Agent 365 / Copilot Studio / M365 Copilot / Work IQ) — agents used, tools/connectors, channels, tokens, prompt/reply content, and Prompt Shield…

SCStelz/security-investigator249—~17kAutomated safety check: PassMITyesterday
43

Design or review specx core scope boundaries in Python services.

maksimzayats/specx202—~1.8kAutomated safety check: PassMIT2 mo ago
44

OpenClaw 安全部署指南 / Security deployment guide — help users secure their OpenClaw installation

jnMetaCode/shellward140—~644Automated safety check: WarnApache-2.08 days ago
45

AI governance, EU AI Act compliance, OWASP LLM security, responsible AI practices for GitHub Copilot agents

Hack23/cia239—~1.4kAutomated safety check: PassApache-2.0today
46

Run or install repo security leak checks with BetterLeaks and Trivy.

instructa/agent-skills139—~557Automated safety check: PassNo licence8 days ago
47

Generate, edit, and compose images using Gemini Nano Banana models via portable Python scripts.

cnemri/google-genai-skills127—~587Automated safety check: PassMIT8 mo ago
48

Start or resume a requested Guardrails implementation or PR takeover in the selected linked worktree with bounded scope and verification.

openai/openai-guardrails-js104—~1kAutomated safety check: PassMIT9 days ago

Questions, answered from the data.

What is the best LLM guardrails skill?

Git Guardrails Claude Code from fossasia/eventyay-interpretation ranks first of the 208 LLM guardrails skills listed here, with the highest score: its repository has 1.6k GitHub stars, 12 other GitHub owners carry a copy, its SKILL.md loads about 578 tokens and it passes the automated safety check with no findings. Next come Obliteratus and Git Guardrails Claude Code.

Which LLM guardrails skills are official?

22 of the 208 LLM guardrails skills are official, published by the vendor's own GitHub organization: Wa Guardrails, Sdaf Bom Selection, Code Change Verification, Migrating Openai Agents SDK To Pydantic AI, Implementation Final Review and 17 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.