Topic · AI & LLM Engineering

Best LLM guardrails skills, page 2

Skills #49–96 of 221, ranked by score.

LLM guardrails skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

LLM guardrails skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Guides designing a layered permission pipeline for agent tools that decides which calls are allowed, need confirmation or are denied, with scopes and hooks.

simbajigege/book2skills184—~2.1kAutomated safety check: PassApache-2.01 mo ago
50

当需要为 AI 系统设计多层安全防线、内容过滤策略和伦理边界时调用此 skill。典型场景包括:设计拒绝策略与升级机制、防御 prompt 注入攻击、实现领域特定安全规则(教育、医疗、金融等)、定义 AI 的价值观锚点。

kangarooking/system-prompt-skills205—~1.2kAutomated safety check: PassMIT5 mo ago
51

Diagnose and strengthen a repository's harness layer: AGENTS.md rules, knowledge layout, architecture boundaries, lint and type gates, API and generated-client contracts, test scaffolding…

PacificStudio/openase268—~1.2kAutomated safety check: PassApache-2.02 mo ago
52

Medium-native long-form research, author-assistance, writing, editing, review, topic discovery, publication matching, and packaging workflow.

flaqai/backlink_skills752—~5.6kAutomated safety check: PassMIT14 days ago
53

Screen LLM input and output with TypeSafe Jev Noul hazard batteries plus a harm Score; policy in code returns pass, review, or block.

cobusgreyling/Jev134—~544Automated safety check: WarnMIT18 days ago
54

WandB-specific PerforatedAI integration guardrail skill. An agent skill from PerforatedAI/PerforatedAI.

PerforatedAI/PerforatedAI237—~2.8kAutomated safety check: PassApache-2.0today
55

A skill your agent uses when building production LLM applications — designing RAG pipelines, choosing vector databases, implementing agent orchestration, optimizing cost, or adding AI safety…

kid-sid/claude-spellbook189—~3.7kAutomated safety check: PassMIT2 mo ago
56

Choose bounded implementation scope and existing owning modules before Guardrails runtime, configuration, client, resource, streaming, Agents, or SDK compatibility changes and feedback fixes.

openai/openai-guardrails-js105—~951Automated safety check: PassMITtoday
57

Build content moderation applications with Azure AI Content Safety SDK for Java.

microsoft/skills3.1k6 repos~2.1kAutomated safety check: PassMIT2 days ago
58

Azure AI Content Safety SDK for Python. An agent skill from microsoft/skills.

microsoft/skills3.1k6 repos~2.2kAutomated safety check: PassMIT2 days ago
59

Implement and review trust-platform HMI schema/value/write contracts with safety guardrails.

johannesPettersson80/trust-platform221—~611Automated safety check: PassApache-2.0yesterday
60

A skill your agent uses when reviewing, auditing, scoring, or improving a real or synthetic MATLAB AI tutor transcript, tutoring prompt, generated lesson, exercise, feedback sequence, or skill…

matlab/agent-skills-playground181—~1.4kAutomated safety check: PassUnknown27 days ago
61
61.PR Draft SummaryOfficial

Draft a Guardrails PR from its complete final diff and carry out authorized CI and review follow-up, including stacked PRs and takeovers.

openai/openai-guardrails-js105—~1kAutomated safety check: PassMITtoday
62

NVIDIA's runtime safety framework for LLM applications. An agent skill from Orchestra-Research/AI-Research-SKILLs.

Orchestra-Research/AI-Research-SKILLs13k2 repos~1.9kAutomated safety check: WarnMIT3 mo ago
63

Applies the reasoning style of Geoffrey Hinton, deep learning pioneer and 2018 Turing Award winner.

K-Dense-AI/mimeo282—~1.8kAutomated safety check: PassMIT1 mo ago
64

Add strict Python project tooling for a specx service. An agent skill from maksimzayats/specx.

maksimzayats/specx202—~965Automated safety check: PassMIT2 mo ago
65

A skill your agent uses when the user wants to tailor a workflow for a specific industry, domain, or vertical with specialized expertise, terminology, and guardrails.

sharpdeveye/maestro592—~614Automated safety check: PassMIT5 mo ago
66

A skill your agent uses for all frontend and web UI tasks by default to apply a business-agnostic desktop web visual style system with medium-strength guardrails; if an existing design system is…

yusifeng/formax195—~902Automated safety check: PassMIT2 mo ago
67

Configure the shared guard against catastrophic shell commands in local AI agents.

davidondrej/skills4.1k—~1.6kAutomated safety check: PassMITyesterday
68

Secure AI coding agents (Claude Code, Cursor, Codex, Copilot) with permission boundaries, secret protection, code review gates, and safe sandbox configurations for team environments.

sickn33/agentic-awesome-skills47k2 repos~3.3kAutomated safety check: NotesMITyesterday
69

A skill your agent uses when reasoning about generative AI, adversarial machine learning, neural network security, algorithmic fairness, or deep learning fundamentals.

K-Dense-AI/mimeo282—~1.7kAutomated safety check: PassMIT1 mo ago
70

LLM-powered quality verification using prompt hooks. An agent skill from rohitg00/pro-workflow.

rohitg00/pro-workflow2.9k—~757Automated safety check: PassNo licence9 days ago
71

Meta's 86M prompt injection and jailbreak detector. An agent skill from Orchestra-Research/AI-Research-SKILLs.

Orchestra-Research/AI-Research-SKILLs13k1 repo~2.4kAutomated safety check: WarnMIT3 mo ago
72

Creates actionable alignment frameworks that give teams a shared North Star (direction), values (guardrails), and decision tenets (behavioral standards).

lyndonkl/claude164—~1.6kAutomated safety check: PassNo licence1 mo ago
73

Generate documents with writefile when retrieval tools fail, with explicit guardrails against task context drift

HKUDS/OpenSpace7.7k—~2.6kAutomated safety check: PassMIT1 mo ago
74

Apex code quality guardrails for Salesforce development. An agent skill from github/awesome-copilot.

github/awesome-copilot40k1 repo~1.8kAutomated safety check: PassMITtoday
75

AI agent and LLM system engineering reference covering single-agent dev (ReAct, tool calling, plan-execute), multi-agent coordination (swarm, role decomposition, file locking), LLM security (prompt…

telagod/code-abyss243—~691Automated safety check: PassMIT2 mo ago
76

Design spec with 98 rules for building CLI tools that AI agents can safely use.

sickn33/agentic-awesome-skills47k2 repos~3.3kAutomated safety check: PassMITyesterday
77

Builds your north star metric, a metric tree with owned input metrics and guardrails, and a review cadence, as a one-page metrics spec you can paste into a doc.

menkesu/awesome-pm-skills429—~5kAutomated safety check: PassUnknown2 days ago
78

Find, compare, adapt, and design bounded AI-agent feedback loops with explicit checks, stop rules, guardrails, and handoffs.

sickn33/agentic-awesome-skills47k1 repo~2.2kAutomated safety check: PassMITyesterday
79

Authorized Android/iOS application reverse engineering and security testing: APK/IPA analysis, runtime instrumentation (Frida/Objection), SSL-pinning and jailbreak/root-detection bypass, per OWASP…

sickn33/agentic-awesome-skills47k1 repo~1.5kAutomated safety check: PassMITyesterday
80

Wires Promptfoo and DeepTeam into CI/CD for automated, repeatable red-teaming of LLM apps against OWASP LLM Top 10, OWASP Agentic, and MITRE ATLAS presets, failing the build when jailbreak or…

mukul975/Anthropic-Cybersecurity-Skills34k—~2.5kAutomated safety check: PassApache-2.01 mo ago
81

Implements input/output validation guardrails for LLM applications using NVIDIA NeMo Guardrails (Colang), custom Python validators for PII detection, and the Guardrails AI framework, intercepting…

mukul975/Anthropic-Cybersecurity-Skills34k—~2.2kAutomated safety check: PassApache-2.01 mo ago
82

Senior engineer CLI expertise for AI agents — workflows, safety guardrails, gotchas, and anti-patterns across cloud, IaC, containers, databases, dev tools, and platforms

SylphAI-Inc/skills111—~1.7kAutomated safety check: PassMIT15 days ago
83

Add or refine tests for specx Python services. An agent skill from maksimzayats/specx.

maksimzayats/specx202—~1.9kAutomated safety check: PassMIT2 mo ago
84

Autonomous AI code generation safety guardrail register: static AST analysis, forbidden import filters, and zero-day vulnerability checks.

sickn33/agentic-awesome-skills47k1 repo~1.4kAutomated safety check: PassMITyesterday
85

Deploy frontend and full-stack apps on Vercel with previews, edge functions, environment promotion, and production guardrails.

sickn33/agentic-awesome-skills47k1 repo~1.9kAutomated safety check: NotesMITyesterday
86

Deploys a baseline landing zone foundation for a Google Cloud Organization, establishing security guardrails using Organization Policies, resource hierarchy folders and projects, billing…

google/skills21k—~4.8kAutomated safety check: PassApache-2.0today
87

A skill your agent uses when managing Alibaba Cloud Content Moderation (Green) via OpenAPI/SDK, including the user needs content moderation resource and policy operations, including…

cinience/alicloud-skills397—~728Automated safety check: PassMIT1 mo ago
88

A skill your agent uses when writing, reviewing, or refactoring Manor code to avoid overcomplication, make surgical changes, surface assumptions, and define verifiable success criteria.

manor-os/manor-ai162—~816Automated safety check: PassMIT1 mo ago
89

A skill your agent uses when assessing AI/ML systems for prompt injection, jailbreak vulnerabilities, model inversion risk, data poisoning exposure, or agent tool abuse.

alirezarezvani/claude-skills28k1 repo~4.5kAutomated safety check: WarnMIT1 mo ago
90

Use before creating a worktree, script, or ad hoc orchestration, handling generated Bazel artifacts, changing native add-ons, or adding serviceradarcore tests.

carverauto/serviceradar921—~1.3kAutomated safety check: PassApache-2.0today
91

Application security defense knowledge for builders. An agent skill from telagod/code-abyss.

telagod/code-abyss243—~777Automated safety check: PassMIT2 mo ago
92

Control bkit automation level (L0-L4), view trust score, and manage guardrails.

ww-w-ai/bkit-claude-code601—~1.6kAutomated safety check: NotesApache-2.011 days ago
93

Generates BYO custom safety policies for NVIDIA Nemotron content-safety guardrails — Nemotron-Content-Safety-Reasoning-4B (text) and multimodal Nemotron-3-Content-Safety.

NVIDIA/skills3.5k1 repo~4.9kAutomated safety check: PassApache-2.0yesterday
94

Run an improve-my-MCP campaign: an autoresearch-style loop that measures the MCP agent experience with the eval harness, picks the highest-impact tool problem from production data, makes one bounded…

PostHog/posthog40k—~1.5kAutomated safety check: PassUnknowntoday
95

AI text humanization and 윤문 (post-editing) specialist that detects and removes AI tells while preserving meaning, facts, and figures.

modu-ai/moai-adk1.2k—~4.7kAutomated safety check: PassApache-2.0today
96

Creates a reusable use case specification file that defines the business problem, stakeholders, and measurable success criteria for model customization, as recommended by the AWS Responsible AI Lens.

awslabs/agent-plugins915—~1kAutomated safety check: PassApache-2.02 days ago