Topic · AI & LLM Engineering
Best LLM guardrails skills, page 3
LLM guardrails skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | A skill your agent uses when a Scenario generation is blocked, refused, or returns a moderation or sensitive-content error, when a video is rejected after rendering or its audio track is flagged… | scenario-labs/ | 913 | — | ~3.6k | Automated safety check: Pass | MIT | 2 days ago |
| 98 | 98.Guard Safety guardrails — blocks destructive commands (rm -rf, DROP TABLE, force-push, git reset --hard) and optionally restricts file edits to a specific directory. | Houseofmvps/ | 123 | — | ~790 | Automated safety check: Notes | MIT | 3 mo ago |
| 99 | 99.Save Tokens Token-saving session automation — statusline, prompt guard, precompact handoffs, session rotation, and handoff commands for Claude Code | aspenkit/ | 102 | — | ~1.6k | Automated safety check: Pass | MIT | 26 days ago |
| 100 | Help users accelerate their product development cycles by implementing high-intensity rituals, ruthless scoping, and technical guardrails that allow teams to ship faster without sacrificing quality. | RefoundAI/ | 1.4k | — | ~1.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 101 | 101.Agent Workflow A skill your agent uses when any Maestro command is invoked — provides foundational workflow design principles across prompt engineering, context management, tool orchestration, agent architecture… | sharpdeveye/ | 592 | — | ~2.1k | Automated safety check: Pass | MIT | 5 mo ago |
| 102 | 102.Wallets How to create, manage, and use Ethereum wallets. An agent skill from austintgriffith/ethskills. | austintgriffith/ | 294 | — | ~1.9k | Automated safety check: Notes | No licence | 1 mo ago |
| 103 | Use this skill before running, or recommending, any Docker command that deletes, wipes, resets, or otherwise irreversibly changes state — even if the user just says to "clean up", "clear the cache"… | docker/ | 539 | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 104 | Defend AI systems against prompt injection and indirect prompt attacks using input controls, tool permissions, output validation, and isolation boundaries. | sickn33/ | 47k | 2 repos | ~4.2k | Automated safety check: Warn | MIT | yesterday |
| 105 | 105.Stuart Russell Applies the reasoning of Stuart Russell, AI safety expert, UC Berkeley professor, and co-author of 'Artificial Intelligence: A Modern Approach'. | K-Dense-AI/ | 282 | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 106 | 106.Yann Lecun This skill channels the reasoning of Yann LeCun, Chief AI Scientist at Meta and Turing Award winner. | K-Dense-AI/ | 282 | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 107 | 107.Yoshua Bengio Applies the reasoning, AI safety frameworks, and deep learning principles of Yoshua Bengio (Turing Award winner, Mila). | K-Dense-AI/ | 282 | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 108 | 108.Social Card Gen Generate platform-specific social post variants (Twitter/X, LinkedIn, Reddit) from one source input. | BrianRWagner/ | 440 | — | ~1.9k | Automated safety check: Pass | No licence | 6 mo ago |
| 109 | Turn a retirement portfolio into sustainable lifetime income: sequence-of-returns risk, the 4% rule and its assumptions, Guyton-Klinger-style guardrails, RMD calculation from the Uniform Lifetime… | JoelLewis/ | 205 | — | ~3.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 110 | 110.Careful Safety guardrails for destructive commands. An agent skill from mr-daedalium/ostack-saas. | mr-daedalium/ | 114 | 1 repo | ~545 | Automated safety check: Notes | MIT | 6 mo ago |
| 111 | AI/LLM defensive security reference: prompt-injection defense, OWASP LLM Top 10 defensive mapping, MCP and agentic tool-call hardening, training-data poisoning detection, model-output validation and… | modu-ai/ | 1.2k | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | today |
| 112 | 112.Existing Repo Analyze existing repositories, maintain structure, setup guardrails and best practices | alinaqi/ | 707 | — | ~4k | Automated safety check: Notes | MIT | 14 days ago |
| 113 | Amazon Aurora MySQL — creates, modifies, and advises on Aurora MySQL clusters specifically (MySQL-compatible engine, Aurora serverless, parallel query). | aws/ | 2.8k | — | ~4.3k | Automated safety check: Pass | Apache-2.0 | today |
| 114 | Amazon Aurora PostgreSQL — creates, modifies, and advises on Aurora PostgreSQL clusters specifically (PostgreSQL-compatible engine, Aurora serverless, express configuration, pgvector, Babelfish). | aws/ | 2.8k | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | today |
| 115 | 115.Content Review A skill your agent uses when needing to understand content moderation policies, avoid content removal, or successfully navigate Xiaohongshu's content review process | vivy-yi/ | 474 | — | ~1.2k | Automated safety check: Pass | No licence | 8 mo ago |
| 116 | A skill your agent uses when reviewing, designing, or modifying Java enterprise systems that may support intermediary services, hosting services, online platforms, marketplaces, content moderation… | jabrena/ | 446 | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 117 | Full red flag table with all guardrail patterns. An agent skill from grandamenium/cortextos. | grandamenium/ | 101 | — | ~978 | Automated safety check: Pass | MIT | 15 days ago |
| 118 | 118.Cloud Security Cloud posture security across AWS, Azure, and GCP — IAM least privilege, public exposure, encryption, logging coverage, landing-zone guardrails. | borghei/ | 881 | — | ~3.5k | Automated safety check: Pass | MIT | yesterday |
| 119 | Design a product metrics dashboard — North Star, input metrics, and guardrails — that a team actually uses to make decisions. | borghei/ | 881 | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 120 | Apply ML/AI project delivery guidance for data exploration, feasibility, experimentation, testing, responsible AI, and operating ML systems. | managedcode/ | 138 | — | ~1k | Automated safety check: Pass | MIT | yesterday |
| 121 | Detect and defend against indirect prompt injection hidden in web pages, documents, and images consumed by an agent, via content extraction (HTML/PDF/OCR), normalization, and scanning with LLM… | mukul975/ | 34k | — | ~2.8k | Automated safety check: Warn | Apache-2.0 | 1 mo ago |
| 122 | Runs NVIDIA garak probe suites (jailbreak, prompt injection, data leakage, toxicity, and more) against an LLM endpoint - Hugging Face models, OpenAI-compatible APIs, or Bedrock - then interprets the… | mukul975/ | 34k | — | ~2.9k | Automated safety check: Warn | Apache-2.0 | 1 mo ago |
| 123 | 123.Azure Policy Guidance for Azure Policy — enforcing and auditing governance and security guardrails at scale across Azure with definitions, initiatives, assignments, and remediation tasks. | vinayaklatthe/ | 175 | — | ~1.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 124 | Builds generative AI applications on Amazon Bedrock. An agent skill from aws/agent-toolkit-for-aws. | aws/ | 2.8k | — | ~8.6k | Automated safety check: Pass | Apache-2.0 | today |
| 125 | Ship every change the way the user requires, with the guardrails. | MengTo/ | 6.7k | — | ~2.1k | Automated safety check: Notes | MIT | 2 days ago |
| 126 | A skill your agent uses when a learner asks for help with MATLAB homework, labs, projects, graded assignments, take-home exams, quizzes, or any programming task where academic integrity, course… | matlab/ | 181 | — | ~1.4k | Automated safety check: Pass | Unknown | 27 days ago |
| 127 | 127.Framing Attacks Catalogue of prompt framings that determine whether an agent refuses or performs specification search, and the harness for probing them. | brycewang-stanford/ | 4.5k | — | ~1.4k | Automated safety check: Pass | Unknown | 3 days ago |
| 128 | 128.Polygraph Behavioral trust grades (A–F) for MCP servers. An agent skill from BankrBot/skills. | BankrBot/ | 1.2k | — | ~3.4k | Automated safety check: Pass | No licence | 3 days ago |
| 129 | Operate DingTalk messaging APIs through UXC with a curated OpenAPI schema, app-token bearer auth, and robot/service-group guardrails. | holon-run/ | 116 | — | ~1.6k | Automated safety check: Pass | MIT | 23 days ago |
| 130 | Run browser automation through @playwright/mcp over UXC stdio MCP, with daemon-friendly session reuse and safe action guardrails. | holon-run/ | 116 | — | ~1.1k | Automated safety check: Pass | MIT | 23 days ago |
| 131 | Operate Slack Web API through UXC with a curated OpenAPI schema, bearer-token auth, and messaging-core guardrails. | holon-run/ | 116 | — | ~1.9k | Automated safety check: Pass | MIT | 23 days ago |
| 132 | 132.Vibeguard Lightweight anti-hallucination workflow for task kickoff, review prioritization, and regression retrospectives. | majiayu000/ | 286 | — | ~723 | Automated safety check: Pass | MIT | today |
| 133 | Guidelines and non-negotiable engineering invariants for modifying opencode-swarm. | ZaxbyHub/ | 490 | — | ~4.7k | Automated safety check: Pass | MIT | today |
| 134 | Guardrail patterns for opencode-swarm — pattern structure, bypass surfaces, regex anti-patterns, and test conventions for checkDestructiveCommand() | ZaxbyHub/ | 490 | — | ~3.5k | Automated safety check: Pass | MIT | today |
| 135 | 135.Dt Obs Genai Analyze & debug GenAI/LLM apps: token cost & caching by prompt, model & provider; latency/errors; agent & tool loops/failures; conversations; guardrails; evaluations; OpenTelemetry/dt-evals setup. | Dynatrace/ | 161 | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | 7 days ago |
| 136 | Deploy frontend and full-stack apps on Vercel with previews, edge functions, environment promotion, and production guardrails. | BagelHole/ | 1.1k | — | ~1.8k | Automated safety check: Notes | MIT | 4 mo ago |
| 137 | Output templates and scoring rubrics for multi-agent team workflows. | slgoodrich/ | 138 | — | ~2.2k | Automated safety check: Pass | Unknown | 6 mo ago |
| 138 | Monthly, file work only, no browser at all. An agent skill from markfulton/ai-employees. | markfulton/ | 495 | — | ~13k | Automated safety check: Pass | MIT | yesterday |
| 139 | ISO 42001 AI Management System (AIMS) compliance. An agent skill from borghei/Claude-Skills. | borghei/ | 881 | — | ~7.4k | Automated safety check: Pass | MIT | yesterday |
| 140 | Run a worker-backed code review using the workflow prompt pack's strongest review lenses: approach correctness, architecture ownership, correctness, guardrails, validation, compatibility, and… | closedloop-ai/ | 122 | — | ~16k | Automated safety check: Pass | Apache-2.0 | today |
| 141 | A skill your agent uses when an application owns tool or LLM/provider call sites and needs to wrap them with NeMo Relay scopes and managed execution APIs for lifecycle events, middleware, or… | NVIDIA/ | 192 | — | ~1k | Automated safety check: Pass | Apache-2.0 | today |
| 142 | NVIDIA RAG Blueprint — deploy, configure, troubleshoot, and manage. | NVIDIA/ | 3.5k | — | ~2.8k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 143 | A Head of AI Ethics interviewer that simulates an interview focused on responsible AI, AI safety, and trust & safety practices. | PrepLabsAI/ | 112 | — | ~5.4k | Automated safety check: Pass | MIT | yesterday |
| 144 | Content moderation with Claude: pre-filter vs LLM-classify, categories, thresholds, HITL. | softspark/ | 179 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents563
- Deep learning415
- Embeddings386
- LLM inference and serving372
- Prompt engineering360
- Retrieval-augmented generation358
- Fine-tuning309
- LLM evaluation308
- Speech recognition and synthesis308
- Structured output and tool calling276
- LLM cost and token optimization259
- LLM API integration255
- Model routing and gateways255
- LLM observability240
- Computer vision203
- Model hubs and datasets180
- GPU and accelerator computing176
- Diffusion and image models166
- Natural language processing131
- Reinforcement learning66
- AI interpretability23