Agent skill

Expert Ops

by ReJeCtAll in ReJeCtAll/ExpertTeam-Codex

基础设施运维专家入口。用于 Codex CLI 的 $expert-ops 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex.

MITAuto-check passedDevOps & Cloud

Install Expert Ops

skills CLI
$ npx skills add ReJeCtAll/ExpertTeam-Codex --skill expert-ops -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ReJeCtAll/ExpertTeam-Codex expert-ops --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ReJeCtAll/ExpertTeam-Codex.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.codex/skills/expert-ops .claude/skills/expert-ops && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
expert-ops
GitHub stars
113
Token cost
~625 tokens
SKILL.md length
93 words
Files
2
Skills in repo
12
Repo updated
First seen
Licence
MIT

At a glance

基础设施运维专家入口。用于 Codex CLI 的 $expert-ops 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex.

  • Works in 6 steps: 建立事实基线:读取现有架构、云账户范围、环境、SLO、流量、资源利用率、账单、备份… → 识别风险与优先级:按影响、概率、可检测性和恢复难度标记 P0/P1/P2。 → 设计目标方案:给出架构、配置、实施顺序、依赖、成本和取舍。 → …
  • Tasks that involve Infrastructure as code
  • SKILL.md covers 定位, Codex CLI 调用方式, Agent and 参数化路由, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Expert Ops is an agent skill from ReJeCtAll/ExpertTeam-Codex. 基础设施运维专家入口。用于 Codex CLI 的 $expert-ops 调用。 适用于监控告警、可观测性、云基础设施与 IaC、安全加固、成本优化、备份恢复、容量规划和完整基础设施健康评估。 触发词:运维、SRE、Prometheus、Grafana、Terraform、Ansible、云架构、安全加固、成本优化、备份恢复、容量规划

Its SKILL.md is about 630 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

It sits in DevOps & Cloud, covering Infrastructure as code, Monitoring and alerting and Site reliability engineering. It works with Terraform, Ansible, Prometheus and Grafana. The repository describes itself as: 一套可直接安装到 ~/.codex 的专家配置,面向 Codex CLI / Codex App 桌面版的 Skill 调用方式 提供 3 个专家团、3 个单专家与 1 个总路由入口。 The licence is MIT.

When your agent uses it

  • Tasks that involve Infrastructure as code
  • Tasks that involve Monitoring and alerting
  • Tasks that involve Site reliability engineering

Example prompts

  • “/expert-ops”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. 建立事实基线:读取现有架构、云账户范围、环境、SLO、流量、资源利用率、账单、备份和事故记录。缺少数据时明确假设,不伪造指标。
  2. 识别风险与优先级:按影响、概率、可检测性和恢复难度标记 P0/P1/P2。
  3. 设计目标方案:给出架构、配置、实施顺序、依赖、成本和取舍。
  4. 制定变更计划:包含预检查、分批发布、观察窗口、回滚条件和责任边界。
  5. 验证结果:使用配置校验、计划预览、健康检查、故障演练、恢复演练和指标对比验证。
  6. 形成报告:输出结论、证据、行动项、负责人建议和 7/30/90 天路线图。

What it can do on your machine

Read from SKILL.md and the folder at commit 59c573b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Expert Ops loads about 625 tokens when it runs. Until then it costs about 45 tokens; SKILL.md has 93 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~45
When it runs · the whole SKILL.md, loaded when a task matches
~625

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ReJeCtAll/ExpertTeam-Codex at commit 59c573b, republished under its MIT licence (© ReJeCtAll). 93 words, ~625 tokens.

Download SKILL.mdSave it as .claude/skills/expert-ops/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
expert-ops
description
基础设施运维专家入口。用于 Codex CLI 的 $expert-ops 调用。 适用于监控告警、可观测性、云基础设施与 IaC、安全加固、成本优化、备份恢复、容量规划和完整基础设施健康评估。 触发词:运维、SRE、Prometheus、Grafana、Terraform、Ansible、云架构、安全加固、成本优化、备份恢复、容量规划

Expert Ops - 基础设施运维专家

你现在启动单专家模式的 基础设施运维专家,Agent ID 为 infrastructure-operations-expert。

定位

面向部署上线后的可靠性与运营问题,输出可验证、可回滚、可持续维护的基础设施方案。默认覆盖可观测性、安全、成本和灾难恢复,不把生产变更当作普通代码修改。

Codex CLI 调用方式

text
$expert-ops <infrastructure operations request>
$expert-ops --monitor <monitoring and alerting request>
$expert-ops --infra <infrastructure and IaC request>
$expert-ops --security <security audit and hardening request>
$expert-ops --cost <cost analysis and optimization request>
$expert-ops --backup <backup and disaster recovery request>
$expert-ops --capacity <capacity planning request>
$expert-ops --full <complete infrastructure health assessment>

注意:Codex CLI 当前使用 $skill-name 调用 Skill,不一定识别 /expert-ops Slash Command。

Agent

如环境支持 Agents,优先读取并采用:

  • ~/.codex/agents/infrastructure-operations-expert.md

如环境不支持独立 Agent 调度,则由当前会话按本 Skill 的规则执行。

参数化路由

  • --monitor:Prometheus/Grafana、日志、追踪、SLO、告警分级、通知与降噪。
  • --infra:Terraform/CloudFormation/Ansible、网络、计算、存储、数据库、容器和自动扩缩容。
  • --security:漏洞管理、补丁、最小权限、密钥管理、审计日志、事件响应和合规差距分析。
  • --cost:资源利用率、规格调整、预留容量、生命周期策略、FinOps 指标和 ROI。
  • --backup:RPO/RTO、备份策略、加密、异地副本、恢复演练和完整性验证。
  • --capacity:增长模型、峰值假设、资源水位、扩容触发器、技术路线图和投资需求。
  • --full:覆盖上述全部维度,输出完整健康评估。

未提供参数时,根据用户意图选择单一维度;同时涉及三个及以上维度时使用完整评估。

执行流程

  1. 建立事实基线:读取现有架构、云账户范围、环境、SLO、流量、资源利用率、账单、备份和事故记录。缺少数据时明确假设,不伪造指标。
  2. 识别风险与优先级:按影响、概率、可检测性和恢复难度标记 P0/P1/P2。
  3. 设计目标方案:给出架构、配置、实施顺序、依赖、成本和取舍。
  4. 制定变更计划:包含预检查、分批发布、观察窗口、回滚条件和责任边界。
  5. 验证结果:使用配置校验、计划预览、健康检查、故障演练、恢复演练和指标对比验证。
  6. 形成报告:输出结论、证据、行动项、负责人建议和 7/30/90 天路线图。

安全与生产边界

  • 默认先做只读发现;没有用户明确授权时,不直接修改生产环境。
  • 任何可能中断服务、删除资源、改变网络/权限、覆盖数据或增加长期成本的操作,都必须先说明风险、影响范围和回滚方案。
  • 凭证、密码、Token、Webhook 和私钥必须使用环境变量或密钥管理服务,禁止硬编码。
  • 输出云服务、Terraform provider、Kubernetes 或安全基线配置前,先核对目标版本和官方文档;无法核实时标注版本假设。
  • 合规输出只能描述控制项覆盖和证据缺口,不能仅凭清单宣称通过 SOC 2、ISO 27001 等认证。
  • 成本节省和 ROI 必须列出计算口径;缺少账单或利用率数据时使用区间估算并标注假设。

交付要求

最终输出应包含:

  • TL;DR:当前健康状态和最高优先级风险。
  • 事实与假设:数据来源、时间范围、缺失信息。
  • 方案:架构或配置,以及关键取舍。
  • 行动项:P0/P1/P2、预计收益、工作量和依赖。
  • 变更安全:预检查、回滚、验证和观察指标。
  • 路线图:适用时给出 7/30/90 天计划。

配置和脚本应说明适用环境、依赖和待替换变量;不要把示例描述为未经适配即可直接用于生产。

协作边界

  • 需要开发运维平台、Operator、内部工具或业务代码时,转交 $expert-software。
  • 需要产品层面的容量投资、预算优先级或商业目标决策时,转交 $expert-product。
  • 需要运维控制台或监控大屏的交互与视觉设计时,转交 $expert-design。

请使用与用户原始需求一致的语言输出。

© ReJeCtAll, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in .codex/skills/expert-ops of ReJeCtAll/ExpertTeam-Codex.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit 59c573b

Compare with similar skills

Expert Ops next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Expert Ops compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Expert Ops this skillReJeCtAll/ExpertTeam-Codex113—~625Automated safety check: PassMIT
Gke Alert Configurationgoogle/skills21k—~5.3kAutomated safety check: PassApache-2.0
Alerting Irmgrafana/skills2791 repos~1.9kAutomated safety check: PassApache-2.0
Promqlgrafana/skills2791 repos~1.1kAutomated safety check: PassApache-2.0
Monitoring Observabilityahmedasmar/devops-claude-skills203—~3.9kAutomated safety check: PassNone
Admingrafana/skills279—~1.5kAutomated safety check: PassApache-2.0

Similar skills

  • Official

    Configures alerting policies in Terraform for Google Kubernetes Engine (GKE) clusters, workloads, and services using PromQL and Google Cloud Managed Service for Prometheus.

    21k GitHub stars~5.3k tokensUpdated today
    DevOps & CloudAuto-check passed
  • Alerting Irm

    grafana/skills

    Official

    Configure Grafana Alerting, Incident Response Management (IRM), and SLOs end-to-end — provisions Grafana-managed and data-source-managed alert rules, contact points (Slack/PagerDuty/email/webhook)…

    279 GitHub starsUsed in 1 repo~1.9k tokens
    DevOps & CloudAuto-check passed
  • Promql

    grafana/skills

    Official

    Write, validate, and optimize PromQL for Prometheus / Grafana Mimir / Grafana Cloud Metrics.

    279 GitHub starsUsed in 1 repo~1.1k tokens
    DevOps & CloudAuto-check passed
  • Monitoring Observability

    ahmedasmar/devops-claude-skills

    Monitoring and observability strategy, implementation, and troubleshooting.

    203 GitHub stars~3.9k tokensUpdated 5 mo ago
    DevOps & CloudAuto-check passed
  • Admin

    grafana/skills

    Official

    Manage Grafana Cloud accounts — organizations, stacks, RBAC roles and assignments, SSO/SAML/OAuth/GitHub auth, service accounts for CI/CD, user invites, team membership, and API-driven provisioning.

    279 GitHub stars~1.5k tokensUpdated yesterday
    DevOps & CloudAuto-check passed
  • Official

    Configures best-practice, high-signal alerting policies for Cloud Run resources on Google Cloud (services, jobs, and worker pools) based on seasoned SRE practices.

    21k GitHub stars~1.7k tokensUpdated today
    DevOps & CloudAuto-check passed

More from ReJeCtAll/ExpertTeam-Codex

All 12 skills in this repo
  • Product Playbook

    ReJeCtAll/ExpertTeam-Codex

    产品战略团队完整手册。覆盖需求文档撰写、路线图管理、竞品分析、用户研究综合、指标评审、Sprint规划、干系人沟通和产品头脑风暴全流程。当涉及任何产品管理相关的请求时自动触发。

    113 GitHub starsUsed in 1 repo~998 tokens
    Auto-check passed
  • Expert Database

    ReJeCtAll/ExpertTeam-Codex

    数据库优化专家入口。用于 Codex CLI 的 $expert-database 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex.

    113 GitHub stars~692 tokensUpdated 3 mo ago
    Auto-check passed
  • Expert Design

    ReJeCtAll/ExpertTeam-Codex

    设计原型专家团入口。用于 Codex CLI 的 $expert-design 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex.

    113 GitHub stars~441 tokensUpdated 3 mo ago
    Auto-check passed
  • Expert Product

    ReJeCtAll/ExpertTeam-Codex

    产品战略团队专家团入口。用于 Codex CLI 的 $expert-product 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex.

    113 GitHub stars~430 tokensUpdated 3 mo ago
    Auto-check passed
  • Expert Security

    ReJeCtAll/ExpertTeam-Codex

    安全专家入口。用于 Codex CLI 的 $expert-security 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex.

    113 GitHub stars~780 tokensUpdated 3 mo ago
    Auto-check passed
  • Expert Software

    ReJeCtAll/ExpertTeam-Codex

    软件开发团队专家团入口。用于 Codex CLI 的 $expert-software 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex.

    113 GitHub stars~467 tokensUpdated 3 mo ago
    Auto-check passed

Categories

Questions about Expert Ops

What does Expert Ops do?

基础设施运维专家入口。用于 Codex CLI 的 $expert-ops 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex. Expert Ops is an agent skill from ReJeCtAll/ExpertTeam-Codex.

When should I use Expert Ops?

Expert Ops fits situations like: tasks that involve Infrastructure as code; tasks that involve Monitoring and alerting; tasks that involve Site reliability engineering.

How do I install Expert Ops in Claude Code?

Run `npx skills add ReJeCtAll/ExpertTeam-Codex --skill expert-ops -a claude-code`. Or copy the skill folder (.codex/skills/expert-ops in ReJeCtAll/ExpertTeam-Codex) into .claude/skills/expert-ops in your project. Claude Code loads it when a task matches its description.

How do I install Expert Ops in Codex?

Run `npx skills add ReJeCtAll/ExpertTeam-Codex --skill expert-ops -a codex`. Or copy the skill folder (.codex/skills/expert-ops in ReJeCtAll/ExpertTeam-Codex) into .agents/skills/expert-ops in your project. Codex loads it when a task matches its description.

Can I use Expert Ops in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ReJeCtAll/ExpertTeam-Codex --skill expert-ops -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/expert-ops, .gemini/skills/expert-ops, .github/skills/expert-ops and .opencode/skills/expert-ops in your project.

What does Expert Ops need to run?

SKILL.md names no scripts, command-line tools or credentials: Expert Ops is instructions for the agent only.

Does Expert Ops access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Expert Ops safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Expert Ops use?

Expert Ops is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Expert Ops use?

About 625 tokens (SKILL.md is roughly 2.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Expert Ops?

Skills that share tags, products or a category with Expert Ops: Gke Alert Configuration (google/skills, 21k stars), Alerting Irm (grafana/skills, 279 stars), Promql (grafana/skills, 279 stars) and Monitoring Observability (ahmedasmar/devops-claude-skills, 203 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Expert Ops?

ReJeCtAll (a GitHub user) maintains it in ReJeCtAll/ExpertTeam-Codex, which has 113 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on July 8, 2026.

Source: ReJeCtAll/ExpertTeam-Codex on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.