Search

LLM guardrails

217 skills found, page 2.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Start or resume a requested Guardrails implementation or PR takeover in the selected linked worktree with bounded scope and verification.

openai/openai-guardrails-js105—~1kAutomated safety check: PassMIT3 days ago
50

Guides designing a layered permission pipeline for agent tools that decides which calls are allowed, need confirmation or are denied, with scopes and hooks.

simbajigege/book2skills183—~2.1kAutomated safety check: PassApache-2.01 mo ago
51

Diagnose and strengthen a repository's harness layer: AGENTS.md rules, knowledge layout, architecture boundaries, lint and type gates, API and generated-client contracts, test scaffolding…

PacificStudio/openase268—~1.2kAutomated safety check: PassApache-2.02 mo ago
52

Medium-native long-form research, author-assistance, writing, editing, review, topic discovery, publication matching, and packaging workflow.

flaqai/backlink_skills756—~5.6kAutomated safety check: PassMIT17 days ago
53

Screen LLM input and output with TypeSafe Jev Noul hazard batteries plus a harm Score; policy in code returns pass, review, or block.

cobusgreyling/Jev136—~544Automated safety check: WarnMITyesterday
54

Designs, tests and refines LLM prompts: zero-shot, few-shot and chain-of-thought patterns, system prompts, structured output schemas and evaluation test suites.

Jeffallan/claude-skills12k—~1.5kAutomated safety check: PassMIT8 days ago
55

WandB-specific PerforatedAI integration guardrail skill. An agent skill from PerforatedAI/PerforatedAI.

PerforatedAI/PerforatedAI237—~2.8kAutomated safety check: PassApache-2.03 days ago
56

A skill your agent uses when building production LLM applications — designing RAG pipelines, choosing vector databases, implementing agent orchestration, optimizing cost, or adding AI safety…

kid-sid/claude-spellbook190—~3.7kAutomated safety check: PassMIT2 mo ago
57

Choose bounded implementation scope and existing owning modules before Guardrails runtime, configuration, client, resource, streaming, Agents, or SDK compatibility changes and feedback fixes.

openai/openai-guardrails-js105—~951Automated safety check: PassMIT3 days ago
58

Build content moderation applications with Azure AI Content Safety SDK for Java.

microsoft/skills3.1k5 repos~2.1kAutomated safety check: PassMIT2 days ago
59

Azure AI Content Safety SDK for Python. An agent skill from microsoft/skills.

microsoft/skills3.1k5 repos~2.2kAutomated safety check: PassMIT2 days ago
60
60.PR Draft SummaryOfficial

Draft a Guardrails PR from its complete final diff and carry out authorized CI and review follow-up, including stacked PRs and takeovers.

openai/openai-guardrails-js105—~1kAutomated safety check: PassMIT3 days ago
61

A skill your agent uses when reviewing, auditing, scoring, or improving a real or synthetic MATLAB AI tutor transcript, tutoring prompt, generated lesson, exercise, feedback sequence, or skill…

matlab/agent-skills-playground184—~1.4kAutomated safety check: PassUnknownyesterday
62

NVIDIA's runtime safety framework for LLM applications. An agent skill from Orchestra-Research/AI-Research-SKILLs.

Orchestra-Research/AI-Research-SKILLs13k2 repos~1.9kAutomated safety check: WarnMIT3 mo ago
63

Applies the reasoning style of Geoffrey Hinton, deep learning pioneer and 2018 Turing Award winner.

K-Dense-AI/mimeo282—~1.8kAutomated safety check: PassMIT1 mo ago
64

Add strict Python project tooling for a specx service. An agent skill from maksimzayats/specx.

maksimzayats/specx202—~965Automated safety check: PassMIT2 mo ago
65

A skill your agent uses when the user wants to tailor a workflow for a specific industry, domain, or vertical with specialized expertise, terminology, and guardrails.

sharpdeveye/maestro591—~614Automated safety check: PassMIT5 mo ago
66

A skill your agent uses for all frontend and web UI tasks by default to apply a business-agnostic desktop web visual style system with medium-strength guardrails; if an existing design system is…

yusifeng/formax194—~902Automated safety check: PassMIT2 mo ago
67

Configure the shared guard against catastrophic shell commands in local AI agents.

davidondrej/skills4.1k—~1.6kAutomated safety check: PassMITtoday
68

Secure AI coding agents (Claude Code, Cursor, Codex, Copilot) with permission boundaries, secret protection, code review gates, and safe sandbox configurations for team environments.

sickn33/agentic-awesome-skills47k2 repos~3.3kAutomated safety check: NotesMITyesterday
69

A skill your agent uses when reasoning about generative AI, adversarial machine learning, neural network security, algorithmic fairness, or deep learning fundamentals.

K-Dense-AI/mimeo282—~1.7kAutomated safety check: PassMIT1 mo ago
70

LLM-powered quality verification using prompt hooks. An agent skill from rohitg00/pro-workflow.

rohitg00/pro-workflow2.9k—~757Automated safety check: PassNo licence12 days ago
71

Meta's 86M prompt injection and jailbreak detector. An agent skill from Orchestra-Research/AI-Research-SKILLs.

Orchestra-Research/AI-Research-SKILLs13k1 repo~2.4kAutomated safety check: WarnMIT3 mo ago
72

Creates actionable alignment frameworks that give teams a shared North Star (direction), values (guardrails), and decision tenets (behavioral standards).

lyndonkl/claude164—~1.6kAutomated safety check: PassNo licence1 mo ago
73

Generate documents with writefile when retrieval tools fail, with explicit guardrails against task context drift

HKUDS/OpenSpace7.8k—~2.6kAutomated safety check: PassMIT2 mo ago
74

Apex code quality guardrails for Salesforce development. An agent skill from github/awesome-copilot.

github/awesome-copilot40k1 repo~1.8kAutomated safety check: PassMIT2 days ago
75

AI agent and LLM system engineering reference covering single-agent dev (ReAct, tool calling, plan-execute), multi-agent coordination (swarm, role decomposition, file locking), LLM security (prompt…

telagod/code-abyss244—~691Automated safety check: PassMIT2 mo ago
76

Design spec with 98 rules for building CLI tools that AI agents can safely use.

sickn33/agentic-awesome-skills47k2 repos~3.3kAutomated safety check: PassMIT2 days ago
77

Builds your north star metric, a metric tree with owned input metrics and guardrails, and a review cadence, as a one-page metrics spec you can paste into a doc.

menkesu/awesome-pm-skills434—~5kAutomated safety check: PassUnknown5 days ago
78

Find, compare, adapt, and design bounded AI-agent feedback loops with explicit checks, stop rules, guardrails, and handoffs.

sickn33/agentic-awesome-skills47k1 repo~2.2kAutomated safety check: PassMIT2 days ago
79

Authorized Android/iOS application reverse engineering and security testing: APK/IPA analysis, runtime instrumentation (Frida/Objection), SSL-pinning and jailbreak/root-detection bypass, per OWASP…

sickn33/agentic-awesome-skills47k1 repo~1.5kAutomated safety check: PassMIT2 days ago
80

Wires Promptfoo and DeepTeam into CI/CD for automated, repeatable red-teaming of LLM apps against OWASP LLM Top 10, OWASP Agentic, and MITRE ATLAS presets, failing the build when jailbreak or…

mukul975/Anthropic-Cybersecurity-Skills34k—~2.5kAutomated safety check: PassApache-2.01 mo ago
81

Implements input/output validation guardrails for LLM applications using NVIDIA NeMo Guardrails (Colang), custom Python validators for PII detection, and the Guardrails AI framework, intercepting…

mukul975/Anthropic-Cybersecurity-Skills34k—~2.2kAutomated safety check: PassApache-2.01 mo ago
82

Senior engineer CLI expertise for AI agents — workflows, safety guardrails, gotchas, and anti-patterns across cloud, IaC, containers, databases, dev tools, and platforms

SylphAI-Inc/skills112—~1.7kAutomated safety check: PassMIT18 days ago
83

Add or refine tests for specx Python services. An agent skill from maksimzayats/specx.

maksimzayats/specx202—~1.9kAutomated safety check: PassMIT2 mo ago
84

Autonomous AI code generation safety guardrail register: static AST analysis, forbidden import filters, and zero-day vulnerability checks.

sickn33/agentic-awesome-skills47k1 repo~1.4kAutomated safety check: PassMIT2 days ago
85

Deploy frontend and full-stack apps on Vercel with previews, edge functions, environment promotion, and production guardrails.

sickn33/agentic-awesome-skills47k1 repo~1.9kAutomated safety check: NotesMIT2 days ago
86

Deploys a baseline landing zone foundation for a Google Cloud Organization, establishing security guardrails using Organization Policies, resource hierarchy folders and projects, billing…

google/skills21k—~4.8kAutomated safety check: PassApache-2.0yesterday
87

A skill your agent uses when managing Alibaba Cloud Content Moderation (Green) via OpenAPI/SDK, including the user needs content moderation resource and policy operations, including…

cinience/alicloud-skills397—~728Automated safety check: PassMIT2 mo ago
88

A skill your agent uses when writing, reviewing, or refactoring Manor code to avoid overcomplication, make surgical changes, surface assumptions, and define verifiable success criteria.

manor-os/manor-ai162—~816Automated safety check: PassMIT1 mo ago
89

Use before creating a worktree, script, or ad hoc orchestration, handling generated Bazel artifacts, changing native add-ons, or adding serviceradarcore tests.

carverauto/serviceradar921—~1.3kAutomated safety check: PassApache-2.0yesterday
90

Application security defense knowledge for builders. An agent skill from telagod/code-abyss.

telagod/code-abyss244—~777Automated safety check: PassMIT2 mo ago
91

Control bkit automation level (L0-L4), view trust score, and manage guardrails.

ww-w-ai/bkit-claude-code601—~1.6kAutomated safety check: NotesApache-2.014 days ago
92

Generates BYO custom safety policies for NVIDIA Nemotron content-safety guardrails — Nemotron-Content-Safety-Reasoning-4B (text) and multimodal Nemotron-3-Content-Safety.

NVIDIA/skills3.6k1 repo~4.9kAutomated safety check: PassApache-2.02 days ago
93

AI text humanization and 윤문 (post-editing) specialist that detects and removes AI tells while preserving meaning, facts, and figures.

modu-ai/moai-adk1.2k—~4.7kAutomated safety check: PassApache-2.02 days ago
94

Run an improve-my-MCP campaign: an autoresearch-style loop that measures the MCP agent experience with the eval harness, picks the highest-impact tool problem from production data, makes one bounded…

PostHog/posthog40k—~1.5kAutomated safety check: PassUnknownyesterday
95

Creates a reusable use case specification file that defines the business problem, stakeholders, and measurable success criteria for model customization, as recommended by the AWS Responsible AI Lens.

awslabs/agent-plugins916—~1kAutomated safety check: PassApache-2.0yesterday
96

Safety guardrails — blocks destructive commands (rm -rf, DROP TABLE, force-push, git reset --hard) and optionally restricts file edits to a specific directory.

Houseofmvps/ultraship123—~790Automated safety check: NotesMIT3 mo ago