Topic · AI & LLM Engineering
Best LLM guardrails skills, page 2
LLM guardrails skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Guides designing a layered permission pipeline for agent tools that decides which calls are allowed, need confirmation or are denied, with scopes and hooks. | simbajigege/ | 184 | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 50 | 当需要为 AI 系统设计多层安全防线、内容过滤策略和伦理边界时调用此 skill。典型场景包括:设计拒绝策略与升级机制、防御 prompt 注入攻击、实现领域特定安全规则(教育、医疗、金融等)、定义 AI 的价值观锚点。 | kangarooking/ | 205 | — | ~1.2k | Automated safety check: Pass | MIT | 5 mo ago |
| 51 | 51.Auto Harness Diagnose and strengthen a repository's harness layer: AGENTS.md rules, knowledge layout, architecture boundaries, lint and type gates, API and generated-client contracts, test scaffolding… | PacificStudio/ | 268 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 52 | Medium-native long-form research, author-assistance, writing, editing, review, topic discovery, publication matching, and packaging workflow. | flaqai/ | 752 | — | ~5.6k | Automated safety check: Pass | MIT | 14 days ago |
| 53 | Screen LLM input and output with TypeSafe Jev Noul hazard batteries plus a harm Score; policy in code returns pass, review, or block. | cobusgreyling/ | 134 | — | ~544 | Automated safety check: Warn | MIT | 18 days ago |
| 54 | WandB-specific PerforatedAI integration guardrail skill. An agent skill from PerforatedAI/PerforatedAI. | PerforatedAI/ | 237 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 55 | 55.AI Engineer A skill your agent uses when building production LLM applications — designing RAG pipelines, choosing vector databases, implementing agent orchestration, optimizing cost, or adding AI safety… | kid-sid/ | 189 | — | ~3.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 56 | Choose bounded implementation scope and existing owning modules before Guardrails runtime, configuration, client, resource, streaming, Agents, or SDK compatibility changes and feedback fixes. | openai/ | 105 | — | ~951 | Automated safety check: Pass | MIT | today |
| 57 | Build content moderation applications with Azure AI Content Safety SDK for Java. | microsoft/ | 3.1k | 6 repos | ~2.1k | Automated safety check: Pass | MIT | 2 days ago |
| 58 | Azure AI Content Safety SDK for Python. An agent skill from microsoft/skills. | microsoft/ | 3.1k | 6 repos | ~2.2k | Automated safety check: Pass | MIT | 2 days ago |
| 59 | Implement and review trust-platform HMI schema/value/write contracts with safety guardrails. | johannesPettersson80/ | 221 | — | ~611 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 60 | A skill your agent uses when reviewing, auditing, scoring, or improving a real or synthetic MATLAB AI tutor transcript, tutoring prompt, generated lesson, exercise, feedback sequence, or skill… | matlab/ | 181 | — | ~1.4k | Automated safety check: Pass | Unknown | 27 days ago |
| 61 | Draft a Guardrails PR from its complete final diff and carry out authorized CI and review follow-up, including stacked PRs and takeovers. | openai/ | 105 | — | ~1k | Automated safety check: Pass | MIT | today |
| 62 | NVIDIA's runtime safety framework for LLM applications. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 2 repos | ~1.9k | Automated safety check: Warn | MIT | 3 mo ago |
| 63 | Applies the reasoning style of Geoffrey Hinton, deep learning pioneer and 2018 Turing Award winner. | K-Dense-AI/ | 282 | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 64 | Add strict Python project tooling for a specx service. An agent skill from maksimzayats/specx. | maksimzayats/ | 202 | — | ~965 | Automated safety check: Pass | MIT | 2 mo ago |
| 65 | 65.Specialize A skill your agent uses when the user wants to tailor a workflow for a specific industry, domain, or vertical with specialized expertise, terminology, and guardrails. | sharpdeveye/ | 592 | — | ~614 | Automated safety check: Pass | MIT | 5 mo ago |
| 66 | A skill your agent uses for all frontend and web UI tasks by default to apply a business-agnostic desktop web visual style system with medium-strength guardrails; if an existing design system is… | yusifeng/ | 195 | — | ~902 | Automated safety check: Pass | MIT | 2 mo ago |
| 67 | Configure the shared guard against catastrophic shell commands in local AI agents. | davidondrej/ | 4.1k | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 68 | Secure AI coding agents (Claude Code, Cursor, Codex, Copilot) with permission boundaries, secret protection, code review gates, and safe sandbox configurations for team environments. | sickn33/ | 47k | 2 repos | ~3.3k | Automated safety check: Notes | MIT | yesterday |
| 69 | A skill your agent uses when reasoning about generative AI, adversarial machine learning, neural network security, algorithmic fairness, or deep learning fundamentals. | K-Dense-AI/ | 282 | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 70 | 70.LLM Gate LLM-powered quality verification using prompt hooks. An agent skill from rohitg00/pro-workflow. | rohitg00/ | 2.9k | — | ~757 | Automated safety check: Pass | No licence | 9 days ago |
| 71 | 71.Prompt Guard Meta's 86M prompt injection and jailbreak detector. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 1 repo | ~2.4k | Automated safety check: Warn | MIT | 3 mo ago |
| 72 | Creates actionable alignment frameworks that give teams a shared North Star (direction), values (guardrails), and decision tenets (behavioral standards). | lyndonkl/ | 164 | — | ~1.6k | Automated safety check: Pass | No licence | 1 mo ago |
| 73 | Generate documents with writefile when retrieval tools fail, with explicit guardrails against task context drift | HKUDS/ | 7.7k | — | ~2.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 74 | Apex code quality guardrails for Salesforce development. An agent skill from github/awesome-copilot. | github/ | 40k | 1 repo | ~1.8k | Automated safety check: Pass | MIT | today |
| 75 | AI agent and LLM system engineering reference covering single-agent dev (ReAct, tool calling, plan-execute), multi-agent coordination (swarm, role decomposition, file locking), LLM security (prompt… | telagod/ | 243 | — | ~691 | Automated safety check: Pass | MIT | 2 mo ago |
| 76 | Design spec with 98 rules for building CLI tools that AI agents can safely use. | sickn33/ | 47k | 2 repos | ~3.3k | Automated safety check: Pass | MIT | yesterday |
| 77 | 77.Metrics Builds your north star metric, a metric tree with owned input metrics and guardrails, and a review cadence, as a one-page metrics spec you can paste into a doc. | menkesu/ | 429 | — | ~5k | Automated safety check: Pass | Unknown | 2 days ago |
| 78 | 78.Loop Library Find, compare, adapt, and design bounded AI-agent feedback loops with explicit checks, stop rules, guardrails, and handoffs. | sickn33/ | 47k | 1 repo | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 79 | Authorized Android/iOS application reverse engineering and security testing: APK/IPA analysis, runtime instrumentation (Frida/Objection), SSL-pinning and jailbreak/root-detection bypass, per OWASP… | sickn33/ | 47k | 1 repo | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 80 | Wires Promptfoo and DeepTeam into CI/CD for automated, repeatable red-teaming of LLM apps against OWASP LLM Top 10, OWASP Agentic, and MITRE ATLAS presets, failing the build when jailbreak or… | mukul975/ | 34k | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 81 | Implements input/output validation guardrails for LLM applications using NVIDIA NeMo Guardrails (Colang), custom Python validators for PII detection, and the Guardrails AI framework, intercepting… | mukul975/ | 34k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 82 | Senior engineer CLI expertise for AI agents — workflows, safety guardrails, gotchas, and anti-patterns across cloud, IaC, containers, databases, dev tools, and platforms | SylphAI-Inc/ | 111 | — | ~1.7k | Automated safety check: Pass | MIT | 15 days ago |
| 83 | 83.Specx Tests Add or refine tests for specx Python services. An agent skill from maksimzayats/specx. | maksimzayats/ | 202 | — | ~1.9k | Automated safety check: Pass | MIT | 2 mo ago |
| 84 | Autonomous AI code generation safety guardrail register: static AST analysis, forbidden import filters, and zero-day vulnerability checks. | sickn33/ | 47k | 1 repo | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 85 | Deploy frontend and full-stack apps on Vercel with previews, edge functions, environment promotion, and production guardrails. | sickn33/ | 47k | 1 repo | ~1.9k | Automated safety check: Notes | MIT | yesterday |
| 86 | Deploys a baseline landing zone foundation for a Google Cloud Organization, establishing security guardrails using Organization Policies, resource hierarchy folders and projects, billing… | google/ | 21k | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | today |
| 87 | A skill your agent uses when managing Alibaba Cloud Content Moderation (Green) via OpenAPI/SDK, including the user needs content moderation resource and policy operations, including… | cinience/ | 397 | — | ~728 | Automated safety check: Pass | MIT | 1 mo ago |
| 88 | A skill your agent uses when writing, reviewing, or refactoring Manor code to avoid overcomplication, make surgical changes, surface assumptions, and define verifiable success criteria. | manor-os/ | 162 | — | ~816 | Automated safety check: Pass | MIT | 1 mo ago |
| 89 | 89.AI Security A skill your agent uses when assessing AI/ML systems for prompt injection, jailbreak vulnerabilities, model inversion risk, data poisoning exposure, or agent tool abuse. | alirezarezvani/ | 28k | 1 repo | ~4.5k | Automated safety check: Warn | MIT | 1 mo ago |
| 90 | Use before creating a worktree, script, or ad hoc orchestration, handling generated Bazel artifacts, changing native add-ons, or adding serviceradarcore tests. | carverauto/ | 921 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | today |
| 91 | Application security defense knowledge for builders. An agent skill from telagod/code-abyss. | telagod/ | 243 | — | ~777 | Automated safety check: Pass | MIT | 2 mo ago |
| 92 | 92.Control Control bkit automation level (L0-L4), view trust score, and manage guardrails. | ww-w-ai/ | 601 | — | ~1.6k | Automated safety check: Notes | Apache-2.0 | 11 days ago |
| 93 | Generates BYO custom safety policies for NVIDIA Nemotron content-safety guardrails — Nemotron-Content-Safety-Reasoning-4B (text) and multimodal Nemotron-3-Content-Safety. | NVIDIA/ | 3.5k | 1 repo | ~4.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 94 | Run an improve-my-MCP campaign: an autoresearch-style loop that measures the MCP agent experience with the eval harness, picks the highest-impact tool problem from production data, makes one bounded… | PostHog/ | 40k | — | ~1.5k | Automated safety check: Pass | Unknown | today |
| 95 | AI text humanization and 윤문 (post-editing) specialist that detects and removes AI tells while preserving meaning, facts, and figures. | modu-ai/ | 1.2k | — | ~4.7k | Automated safety check: Pass | Apache-2.0 | today |
| 96 | Creates a reusable use case specification file that defines the business problem, stakeholders, and measurable success criteria for model customization, as recommended by the AWS Responsible AI Lens. | awslabs/ | 915 | — | ~1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents563
- Deep learning415
- Embeddings386
- LLM inference and serving372
- Prompt engineering360
- Retrieval-augmented generation358
- Fine-tuning309
- LLM evaluation308
- Speech recognition and synthesis308
- Structured output and tool calling276
- LLM cost and token optimization259
- LLM API integration255
- Model routing and gateways255
- LLM observability240
- Computer vision203
- Model hubs and datasets180
- GPU and accelerator computing176
- Diffusion and image models166
- Natural language processing131
- Reinforcement learning66
- AI interpretability23