Topic · AI & LLM Engineering
Best LLM guardrails skills for Claude Code, Codex and other agents.
- skills
- 208
- official
- 22
LLM guardrails skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Set up Claude Code hooks to block dangerous git commands (push, reset --hard, clean, branch -D, etc.) before they execute. | fossasia/ | 1.6k | 12 repos | ~578 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 2 | Remove refusal behaviors from open-weight LLMs using OBLITERATUS — mechanistic interpretability techniques (diff-in-means, SVD, whitened SVD, LEACE, SAE decomposition, etc.) to excise guardrails… | RedWoodOG/ | 177 | 6 repos | ~3.8k | Automated safety check: Pass | MIT | 4 mo ago |
| 3 | 设置 Claude Code hooks,在危险 git commands(push、reset --hard、clean、branch -D 等)执行前阻止它们。适用于用户想防止破坏性 git 操作、添加 git safety hooks,或在 Claude Code 中阻止 git push/reset 时。 | vinvcn/ | 4.6k | — | ~474 | Automated safety check: Pass | MIT | 9 days ago |
| 4 | Behavioral guardrails for LLM-assisted coding. An agent skill from alirezarezvani/ClaudeForge. | alirezarezvani/ | 429 | — | ~1.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 5 | Systematically reduce the shipped bundle size of a JS/TS library without sacrificing code readability or breaking consumer APIs. | kcsujeet/ | 351 | — | ~2.9k | Automated safety check: Pass | MIT | 3 days ago |
| 6 | Interactive scaffold generator for Orloj multi-agent systems. | OrlojHQ/ | 123 | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | 19 days ago |
| 7 | Read AI Safety HOT daily digests, search recent AI safety research and incidents, and follow current hot topics. | wuyoscar/ | 175 | — | ~1.2k | Automated safety check: Pass | Unknown | today |
| 8 | Installs, tunes and enforces Sponsio contracts that block unsafe tool calls in LLM agents, covering setup, auditing, observe mode and flipping to enforce. | SponsioLabs/ | 440 | — | ~12k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 9 | Checks each shell command for destructive patterns such as recursive deletes, force pushes and dropped tables, and asks before letting them run. | garrytan/ | 136k | — | ~931 | Automated safety check: Notes | MIT | today |
| 10 | Turns a natural-language description of routing intent into a valid Lemonade collection.router policy JSON. | amd/ | 395 | — | ~4k | Automated safety check: Pass | MIT | today |
| 11 | 当需要为 AI 产品定义核心身份、角色声明和能力边界时调用此 skill。典型场景包括:设计新 AI 产品的 system prompt 首段、为不同场景创建差异化角色(如教学助手 vs 编程代理)、重新定义 AI 与用户的关系框架。 | kangarooking/ | 205 | 1 repo | ~956 | Automated safety check: Pass | MIT | 5 mo ago |
| 12 | 按中国法规(网安法 / PIPL / 等保2.0 / 数据出境 / AI生成内容标识)审计一个 AI 项目的代码仓库,产出每条都带 文件:行 取证、经独立复核、经脚本校验的合规报告。当用户问「这个项目上线合不合规」「调用了 OpenAI/Claude 算不算数据出境」「要不要做 AI 标识」「帮我做合规自查/等保/PIPL 检查」时使用。Audit an AI project's… | jnMetaCode/ | 140 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 13 | 13.Iso42001 Expert ISO 42001 AI Management System (AIMS) compliance advisor. | Sushegaad/ | 939 | 1 repo | ~3.7k | Automated safety check: Pass | MIT | 2 days ago |
| 14 | Designs, tests and refines LLM prompts: zero-shot, few-shot and chain-of-thought patterns, system prompts, structured output schemas and evaluation test suites. | Jeffallan/ | 12k | 1 repo | ~1.5k | Automated safety check: Pass | MIT | 4 days ago |
| 15 | Always-on operational guardrails, model-independent. An agent skill from mrtooher/fable-mode. | mrtooher/ | 870 | — | ~1k | Automated safety check: Pass | No licence | 1 mo ago |
| 16 | Installs hooks that check each agent action against security policies before it runs, blocking destructive commands and logging every decision. | pegasi-ai/ | 392 | — | ~1.4k | Automated safety check: Warn | Apache-2.0 | 4 mo ago |
| 17 | Guide for writing eval conversation JSONs and running them through policy engines | open-bias/ | 143 | — | ~1.5k | Automated safety check: Pass | Apache-2.0 | 4 mo ago |
| 18 | A skill your agent uses when you need a deterministic inspection of a WordPress repository (plugin/theme/block theme/WP core/Gutenberg/full site) including tooling/tests/version hints, and a… | gambitph/ | 350 | 3 repos | ~371 | Automated safety check: Pass | GPL-3.0 | today |
| 19 | Generate preventive Well-Architected guardrails — AWS Config rules, Service Control Policies, permission boundaries, CloudWatch alarms, and IaC policy checks (CDK Aspects, cfn-guard, OPA/Sentinel) —… | aws-samples/ | 273 | — | ~2.8k | Automated safety check: Pass | MIT-0 | yesterday |
| 20 | 20.Celestia Route Celestia requests to the correct repo and apply canonical blob submit/retrieve guidance (Go, Rust, and Node RPC) with docs guardrails. | celestiaorg/ | 183 | — | ~2.7k | Automated safety check: Pass | No licence | today |
| 21 | Pick the right SDAF BOM for a target SAP product / release / DB platform / version / kernel / topology. | Azure/ | 145 | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 22 | A skill your agent uses when we want to turn a just-finished Formax workflow (e.g. | yusifeng/ | 195 | — | ~530 | Automated safety check: Pass | MIT | 2 mo ago |
| 23 | 设置 Claude Code 钩子,在危险 Git 命令(push、reset --hard、clean、branch -D 等)执行前将其拦截。当用户想要防止破坏性 Git 操作、添加 Git 安全钩子或在 Claude Code 中阻止 git push/reset 时使用。 | devcxl/ | 433 | — | ~420 | Automated safety check: Pass | MIT | yesterday |
| 24 | 24.Cx AI Center A skill your agent uses for any question or action about the user's AI/GenAI applications or agents — their behavior, prompts/responses, quality, hallucinations, guardrails, security, cost/tokens… | coralogix/ | 121 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 25 | Check user-generated text against a written policy before it is published. | mrmps/ | 424 | — | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 26 | 26.Fable Mode Enforces staged execution discipline on large tasks: a written stage plan, delegation to named fable agents where the runtime supports it, a failable verification check at each stage, and a… | mrtooher/ | 870 | — | ~1k | Automated safety check: Pass | No licence | 1 mo ago |
| 27 | Guide for creating a new policy engine under openbias/policy/engines/ | open-bias/ | 143 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 4 mo ago |
| 28 | A skill your agent uses when designing or reviewing a backend MVP with tight budget, evolving schema, and reliance on third-party backends where idempotency, replay, and responsibility attribution… | victorGPT/ | 131 | — | ~1.3k | Automated safety check: Pass | MIT | 2 mo ago |
| 29 | A skill your agent uses when the user wants to turn an application, product, startup idea, SaaS, mobile app, web app, API, AI product, or internal tool into a production-ready Markdown specification… | instructa/ | 139 | — | ~1.5k | Automated safety check: Pass | No licence | 8 days ago |
| 30 | Select and run Guardrails verification for code, dependency, test, tooling, documentation, or repository-workflow changes. | openai/ | 104 | — | ~619 | Automated safety check: Pass | MIT | 9 days ago |
| 31 | 31.Dsh Playbook How to dispatch well — choosing flash vs pro, writing self-contained briefs, parallelism, verifying results, and guardrails | ZSeven-W/ | 156 | — | ~1.1k | Automated safety check: Pass | MIT | today |
| 32 | Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage. | Orchestra-Research/ | 13k | 3 repos | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 33 | 33.Typescript TypeScript anti-slop guardrails. An agent skill from web-infra-dev/rstest. | web-infra-dev/ | 505 | — | ~1.3k | Automated safety check: Pass | MIT | 7 days ago |
| 34 | Step-by-step guide for adding a new guardrail provider to Agent Kernel. | yaalalabs/ | 191 | — | ~3.5k | Automated safety check: Pass | Apache-2.0 | today |
| 35 | Evaluate, rank, and communicate work priorities using AI as a structured thinking partner. | ILoveDotNet/ | 155 | — | ~3.5k | Automated safety check: Pass | CC0-1.0 | yesterday |
| 36 | Migrate Python OpenAI Agents SDK applications to Pydantic AI and, when warranted, Pydantic AI Harness. | pydantic/ | 20k | — | ~1.8k | Automated safety check: Pass | MIT | today |
| 37 | Uses Meta's LlamaGuard moderation model to screen prompts and model replies against six safety categories, with vLLM, FastAPI and NeMo Guardrails setups. | Orchestra-Research/ | 13k | 3 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 38 | Prepare independent review of a complete Guardrails change before pushing or updating a PR, using the repository adversarial-review procedure. | openai/ | 104 | — | ~825 | Automated safety check: Pass | MIT | 9 days ago |
| 39 | A skill your agent uses when an instructor wants to create, interview for, configure, install, update, or review a course AI-use policy for MATLAB AI tutoring. | matlab/ | 181 | — | ~1.2k | Automated safety check: Pass | Unknown | 26 days ago |
| 40 | A skill your agent uses when authoring or repairing Kilroy Attractor DOT graphs from requirements, with template-first topology, routing guardrails, and validator-clean output. | danshapiro/ | 221 | — | ~6.4k | Automated safety check: Pass | MIT | 5 mo ago |
| 41 | Diagnose and draft Amazon Canada apparel advertising plans with lifecycle and seasonal timing, English/French search coverage, account evidence, profitability guardrails, and approval-ready… | xjli360/ | 237 | — | ~1.2k | Automated safety check: Pass | MIT | 9 days ago |
| 42 | Report/investigate RUNTIME ACTIVITY of AI agents (Agent 365 / Copilot Studio / M365 Copilot / Work IQ) — agents used, tools/connectors, channels, tokens, prompt/reply content, and Prompt Shield… | SCStelz/ | 249 | — | ~17k | Automated safety check: Pass | MIT | yesterday |
| 43 | Design or review specx core scope boundaries in Python services. | maksimzayats/ | 202 | — | ~1.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 44 | OpenClaw 安全部署指南 / Security deployment guide — help users secure their OpenClaw installation | jnMetaCode/ | 140 | — | ~644 | Automated safety check: Warn | Apache-2.0 | 8 days ago |
| 45 | AI governance, EU AI Act compliance, OWASP LLM security, responsible AI practices for GitHub Copilot agents | Hack23/ | 239 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 46 | Run or install repo security leak checks with BetterLeaks and Trivy. | instructa/ | 139 | — | ~557 | Automated safety check: Pass | No licence | 8 days ago |
| 47 | Generate, edit, and compose images using Gemini Nano Banana models via portable Python scripts. | cnemri/ | 127 | — | ~587 | Automated safety check: Pass | MIT | 8 mo ago |
| 48 | Start or resume a requested Guardrails implementation or PR takeover in the selected linked worktree with bounded scope and verification. | openai/ | 104 | — | ~1k | Automated safety check: Pass | MIT | 9 days ago |
Questions, answered from the data.
What is the best LLM guardrails skill?
Git Guardrails Claude Code from fossasia/eventyay-interpretation ranks first of the 208 LLM guardrails skills listed here, with the highest score: its repository has 1.6k GitHub stars, 12 other GitHub owners carry a copy, its SKILL.md loads about 578 tokens and it passes the automated safety check with no findings. Next come Obliteratus and Git Guardrails Claude Code.
Which LLM guardrails skills are official?
22 of the 208 LLM guardrails skills are official, published by the vendor's own GitHub organization: Wa Guardrails, Sdaf Bom Selection, Code Change Verification, Migrating Openai Agents SDK To Pydantic AI, Implementation Final Review and 17 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Retrieval-augmented generation360
- Prompt engineering350
- Fine-tuning313
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM cost and token optimization256
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- Computer vision206
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23