Search
LLM guardrails
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Start or resume a requested Guardrails implementation or PR takeover in the selected linked worktree with bounded scope and verification. | openai/ | 105 | — | ~1k | Automated safety check: Pass | MIT | 3 days ago |
| 50 | Guides designing a layered permission pipeline for agent tools that decides which calls are allowed, need confirmation or are denied, with scopes and hooks. | simbajigege/ | 183 | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 51 | 51.Auto Harness Diagnose and strengthen a repository's harness layer: AGENTS.md rules, knowledge layout, architecture boundaries, lint and type gates, API and generated-client contracts, test scaffolding… | PacificStudio/ | 268 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 52 | Medium-native long-form research, author-assistance, writing, editing, review, topic discovery, publication matching, and packaging workflow. | flaqai/ | 756 | — | ~5.6k | Automated safety check: Pass | MIT | 17 days ago |
| 53 | Screen LLM input and output with TypeSafe Jev Noul hazard batteries plus a harm Score; policy in code returns pass, review, or block. | cobusgreyling/ | 136 | — | ~544 | Automated safety check: Warn | MIT | yesterday |
| 54 | Designs, tests and refines LLM prompts: zero-shot, few-shot and chain-of-thought patterns, system prompts, structured output schemas and evaluation test suites. | Jeffallan/ | 12k | — | ~1.5k | Automated safety check: Pass | MIT | 8 days ago |
| 55 | WandB-specific PerforatedAI integration guardrail skill. An agent skill from PerforatedAI/PerforatedAI. | PerforatedAI/ | 237 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 56 | 56.AI Engineer A skill your agent uses when building production LLM applications — designing RAG pipelines, choosing vector databases, implementing agent orchestration, optimizing cost, or adding AI safety… | kid-sid/ | 190 | — | ~3.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 57 | Choose bounded implementation scope and existing owning modules before Guardrails runtime, configuration, client, resource, streaming, Agents, or SDK compatibility changes and feedback fixes. | openai/ | 105 | — | ~951 | Automated safety check: Pass | MIT | 3 days ago |
| 58 | Build content moderation applications with Azure AI Content Safety SDK for Java. | microsoft/ | 3.1k | 5 repos | ~2.1k | Automated safety check: Pass | MIT | 2 days ago |
| 59 | Azure AI Content Safety SDK for Python. An agent skill from microsoft/skills. | microsoft/ | 3.1k | 5 repos | ~2.2k | Automated safety check: Pass | MIT | 2 days ago |
| 60 | Draft a Guardrails PR from its complete final diff and carry out authorized CI and review follow-up, including stacked PRs and takeovers. | openai/ | 105 | — | ~1k | Automated safety check: Pass | MIT | 3 days ago |
| 61 | A skill your agent uses when reviewing, auditing, scoring, or improving a real or synthetic MATLAB AI tutor transcript, tutoring prompt, generated lesson, exercise, feedback sequence, or skill… | matlab/ | 184 | — | ~1.4k | Automated safety check: Pass | Unknown | yesterday |
| 62 | NVIDIA's runtime safety framework for LLM applications. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 2 repos | ~1.9k | Automated safety check: Warn | MIT | 3 mo ago |
| 63 | Applies the reasoning style of Geoffrey Hinton, deep learning pioneer and 2018 Turing Award winner. | K-Dense-AI/ | 282 | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 64 | Add strict Python project tooling for a specx service. An agent skill from maksimzayats/specx. | maksimzayats/ | 202 | — | ~965 | Automated safety check: Pass | MIT | 2 mo ago |
| 65 | 65.Specialize A skill your agent uses when the user wants to tailor a workflow for a specific industry, domain, or vertical with specialized expertise, terminology, and guardrails. | sharpdeveye/ | 591 | — | ~614 | Automated safety check: Pass | MIT | 5 mo ago |
| 66 | A skill your agent uses for all frontend and web UI tasks by default to apply a business-agnostic desktop web visual style system with medium-strength guardrails; if an existing design system is… | yusifeng/ | 194 | — | ~902 | Automated safety check: Pass | MIT | 2 mo ago |
| 67 | Configure the shared guard against catastrophic shell commands in local AI agents. | davidondrej/ | 4.1k | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 68 | Secure AI coding agents (Claude Code, Cursor, Codex, Copilot) with permission boundaries, secret protection, code review gates, and safe sandbox configurations for team environments. | sickn33/ | 47k | 2 repos | ~3.3k | Automated safety check: Notes | MIT | yesterday |
| 69 | A skill your agent uses when reasoning about generative AI, adversarial machine learning, neural network security, algorithmic fairness, or deep learning fundamentals. | K-Dense-AI/ | 282 | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 70 | 70.LLM Gate LLM-powered quality verification using prompt hooks. An agent skill from rohitg00/pro-workflow. | rohitg00/ | 2.9k | — | ~757 | Automated safety check: Pass | No licence | 12 days ago |
| 71 | 71.Prompt Guard Meta's 86M prompt injection and jailbreak detector. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 1 repo | ~2.4k | Automated safety check: Warn | MIT | 3 mo ago |
| 72 | Creates actionable alignment frameworks that give teams a shared North Star (direction), values (guardrails), and decision tenets (behavioral standards). | lyndonkl/ | 164 | — | ~1.6k | Automated safety check: Pass | No licence | 1 mo ago |
| 73 | Generate documents with writefile when retrieval tools fail, with explicit guardrails against task context drift | HKUDS/ | 7.8k | — | ~2.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 74 | Apex code quality guardrails for Salesforce development. An agent skill from github/awesome-copilot. | github/ | 40k | 1 repo | ~1.8k | Automated safety check: Pass | MIT | 2 days ago |
| 75 | AI agent and LLM system engineering reference covering single-agent dev (ReAct, tool calling, plan-execute), multi-agent coordination (swarm, role decomposition, file locking), LLM security (prompt… | telagod/ | 244 | — | ~691 | Automated safety check: Pass | MIT | 2 mo ago |
| 76 | Design spec with 98 rules for building CLI tools that AI agents can safely use. | sickn33/ | 47k | 2 repos | ~3.3k | Automated safety check: Pass | MIT | 2 days ago |
| 77 | 77.Metrics Builds your north star metric, a metric tree with owned input metrics and guardrails, and a review cadence, as a one-page metrics spec you can paste into a doc. | menkesu/ | 434 | — | ~5k | Automated safety check: Pass | Unknown | 5 days ago |
| 78 | 78.Loop Library Find, compare, adapt, and design bounded AI-agent feedback loops with explicit checks, stop rules, guardrails, and handoffs. | sickn33/ | 47k | 1 repo | ~2.2k | Automated safety check: Pass | MIT | 2 days ago |
| 79 | Authorized Android/iOS application reverse engineering and security testing: APK/IPA analysis, runtime instrumentation (Frida/Objection), SSL-pinning and jailbreak/root-detection bypass, per OWASP… | sickn33/ | 47k | 1 repo | ~1.5k | Automated safety check: Pass | MIT | 2 days ago |
| 80 | Wires Promptfoo and DeepTeam into CI/CD for automated, repeatable red-teaming of LLM apps against OWASP LLM Top 10, OWASP Agentic, and MITRE ATLAS presets, failing the build when jailbreak or… | mukul975/ | 34k | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 81 | Implements input/output validation guardrails for LLM applications using NVIDIA NeMo Guardrails (Colang), custom Python validators for PII detection, and the Guardrails AI framework, intercepting… | mukul975/ | 34k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 82 | Senior engineer CLI expertise for AI agents — workflows, safety guardrails, gotchas, and anti-patterns across cloud, IaC, containers, databases, dev tools, and platforms | SylphAI-Inc/ | 112 | — | ~1.7k | Automated safety check: Pass | MIT | 18 days ago |
| 83 | 83.Specx Tests Add or refine tests for specx Python services. An agent skill from maksimzayats/specx. | maksimzayats/ | 202 | — | ~1.9k | Automated safety check: Pass | MIT | 2 mo ago |
| 84 | Autonomous AI code generation safety guardrail register: static AST analysis, forbidden import filters, and zero-day vulnerability checks. | sickn33/ | 47k | 1 repo | ~1.4k | Automated safety check: Pass | MIT | 2 days ago |
| 85 | Deploy frontend and full-stack apps on Vercel with previews, edge functions, environment promotion, and production guardrails. | sickn33/ | 47k | 1 repo | ~1.9k | Automated safety check: Notes | MIT | 2 days ago |
| 86 | Deploys a baseline landing zone foundation for a Google Cloud Organization, establishing security guardrails using Organization Policies, resource hierarchy folders and projects, billing… | google/ | 21k | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 87 | A skill your agent uses when managing Alibaba Cloud Content Moderation (Green) via OpenAPI/SDK, including the user needs content moderation resource and policy operations, including… | cinience/ | 397 | — | ~728 | Automated safety check: Pass | MIT | 2 mo ago |
| 88 | A skill your agent uses when writing, reviewing, or refactoring Manor code to avoid overcomplication, make surgical changes, surface assumptions, and define verifiable success criteria. | manor-os/ | 162 | — | ~816 | Automated safety check: Pass | MIT | 1 mo ago |
| 89 | Use before creating a worktree, script, or ad hoc orchestration, handling generated Bazel artifacts, changing native add-ons, or adding serviceradarcore tests. | carverauto/ | 921 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 90 | Application security defense knowledge for builders. An agent skill from telagod/code-abyss. | telagod/ | 244 | — | ~777 | Automated safety check: Pass | MIT | 2 mo ago |
| 91 | 91.Control Control bkit automation level (L0-L4), view trust score, and manage guardrails. | ww-w-ai/ | 601 | — | ~1.6k | Automated safety check: Notes | Apache-2.0 | 14 days ago |
| 92 | Generates BYO custom safety policies for NVIDIA Nemotron content-safety guardrails — Nemotron-Content-Safety-Reasoning-4B (text) and multimodal Nemotron-3-Content-Safety. | NVIDIA/ | 3.6k | 1 repo | ~4.9k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 93 | AI text humanization and 윤문 (post-editing) specialist that detects and removes AI tells while preserving meaning, facts, and figures. | modu-ai/ | 1.2k | — | ~4.7k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 94 | Run an improve-my-MCP campaign: an autoresearch-style loop that measures the MCP agent experience with the eval harness, picks the highest-impact tool problem from production data, makes one bounded… | PostHog/ | 40k | — | ~1.5k | Automated safety check: Pass | Unknown | yesterday |
| 95 | Creates a reusable use case specification file that defines the business problem, stakeholders, and measurable success criteria for model customization, as recommended by the AWS Responsible AI Lens. | awslabs/ | 916 | — | ~1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 96 | 96.Guard Safety guardrails — blocks destructive commands (rm -rf, DROP TABLE, force-push, git reset --hard) and optionally restricts file edits to a specific directory. | Houseofmvps/ | 123 | — | ~790 | Automated safety check: Notes | MIT | 3 mo ago |