Analyze Trajectory
yologdev/yoyo-evolve
Diagnoses a recurring failure such as a stuck task, repeated CI error or frequent reverts by sending sub-agents through the logs and returning one root-cause diagnosis.
Diagnose and plan fixes for errors/bugs with Codex-first multi-agent collaboration (Codex + Opus 4.6 + Agent Teams).
$ npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install DeL-TaiseiOzaki/claude-code-orchestra troubleshoot --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/DeL-TaiseiOzaki/claude-code-orchestra.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/troubleshoot .claude/skills/troubleshoot && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "troubleshoot" agent skill from https://github.com/DeL-TaiseiOzaki/claude-code-orchestra/tree/main/.claude/skills/troubleshoot into .claude/skills/troubleshoot/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "troubleshoot", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/DeL-TaiseiOzaki/claude-code-orchestra/tree/main/.claude/skills/troubleshootType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install DeL-TaiseiOzaki/claude-code-orchestra troubleshoot --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/DeL-TaiseiOzaki/claude-code-orchestra.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/troubleshoot .agents/skills/troubleshoot && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "troubleshoot" agent skill from https://github.com/DeL-TaiseiOzaki/claude-code-orchestra/tree/main/.claude/skills/troubleshoot into .agents/skills/troubleshoot/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "troubleshoot", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install DeL-TaiseiOzaki/claude-code-orchestra troubleshoot --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/DeL-TaiseiOzaki/claude-code-orchestra.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/troubleshoot .cursor/skills/troubleshoot && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "troubleshoot" agent skill from https://github.com/DeL-TaiseiOzaki/claude-code-orchestra/tree/main/.claude/skills/troubleshoot into .cursor/skills/troubleshoot/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "troubleshoot", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/DeL-TaiseiOzaki/claude-code-orchestra.git --path .claude/skills/troubleshoot--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install DeL-TaiseiOzaki/claude-code-orchestra troubleshoot --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/DeL-TaiseiOzaki/claude-code-orchestra.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/troubleshoot .gemini/skills/troubleshoot && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "troubleshoot" agent skill from https://github.com/DeL-TaiseiOzaki/claude-code-orchestra/tree/main/.claude/skills/troubleshoot into .gemini/skills/troubleshoot/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "troubleshoot", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install DeL-TaiseiOzaki/claude-code-orchestra troubleshootInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/DeL-TaiseiOzaki/claude-code-orchestra.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/troubleshoot .github/skills/troubleshoot && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "troubleshoot" agent skill from https://github.com/DeL-TaiseiOzaki/claude-code-orchestra/tree/main/.claude/skills/troubleshoot into .github/skills/troubleshoot/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "troubleshoot", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install DeL-TaiseiOzaki/claude-code-orchestra troubleshoot --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/DeL-TaiseiOzaki/claude-code-orchestra.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/troubleshoot .opencode/skills/troubleshoot && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "troubleshoot" agent skill from https://github.com/DeL-TaiseiOzaki/claude-code-orchestra/tree/main/.claude/skills/troubleshoot into .opencode/skills/troubleshoot/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "troubleshoot", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
troubleshootDiagnose and plan fixes for errors/bugs with Codex-first multi-agent collaboration (Codex + Opus 4.6 + Agent Teams).
Troubleshoot is an agent skill from DeL-TaiseiOzaki/claude-code-orchestra. Diagnose and plan fixes for errors/bugs with Codex-first multi-agent collaboration (Codex + Opus 4.6 + Agent Teams). Codex CLI is consulted in EVERY phase for deep code reasoning, hypothesis evaluation, and fix validation. Phase 1: Error reproduction & context gathering (Opus subagent 1M context + Codex initial analysis + Claude user interaction). Phase 2: Parallel diagnosis (Agent Teams: Root Cause Analyst [Codex-driven] + Impact Investigator [Opus + Codex risk analysis]). Phase 3: Fix plan synthesis, Codex…
Its SKILL.md is about 7.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including reference files (for example `references/bug-report-template.md`, `references/debug-patterns.md` and `references/diagnosis-template.md`).
It sits in Agent Workflows, covering Subagents and Root cause analysis. The licence is MIT.
3 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit ef0d8f8. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships script files (Python), which the agent can run.
Shell commands in SKILL.md call:
python3claudecodexbashgitFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Troubleshoot loads about 7.6k tokens when it runs, and up to ~9.5k if it reads all its reference files. Until then it costs about 153 tokens; SKILL.md has 1,839 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from DeL-TaiseiOzaki/claude-code-orchestra at commit ef0d8f8, republished under its MIT licence (© DeL-TaiseiOzaki). 1,839 words, ~7,587 tokens.
.claude/skills/troubleshoot/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.Codex-first error/bug diagnosis skill leveraging Codex deep reasoning, Opus 1M context, and Agent Teams.
Preflight: ensure codex CLI is current (see codex-system skill).
This skill handles the diagnosis phases (Phase 1-3) with a Codex-first approach: Codex CLI is consulted proactively in every phase for pattern recognition, hypothesis evaluation, root cause reasoning, and fix validation. Fix implementation and review are done via /team-execute.
/troubleshoot <error description> <- This skill (diagnosis & fix planning)
| After approval
/team-execute <- Parallel fix implementation (Phase 1)
| After completion
Phase 2 REVIEW <- Parallel review (regression check)Phase 1: REPRODUCE & UNDERSTAND (Opus 1M context + Codex Initial Analysis + Claude Lead)
Opus subagent analyzes the error context, Codex generates initial hypotheses,
Claude gathers details from the user
|
Phase 2: DIAGNOSE (Agent Teams -- Parallel, Codex-driven)
Root Cause Analyst (Codex mandatory) <-> Impact Investigator (Opus + Codex) communicate bidirectionally
Both teammates consult Codex for deep reasoning throughout analysis
|
Phase 3: FIX PLAN & APPROVE (Codex Validation + Claude Lead + User)
Integrate diagnosis results, validate fix plan with Codex, get user approvalReproduce the error and gather full context with Opus subagent's 1M context, then consult Codex for initial hypothesis generation, while Claude interacts with the user.
Main orchestrator context is precious. Large-scale error context analysis is delegated to Opus subagent (1M context). Codex is consulted early for pattern recognition and hypothesis generation.
Resolve this bug's deterministic workspace once. The title becomes file and directory names, so give it a short English descriptor of the bug -- not the user's raw wording, which the Language Protocol keeps out of paths:
python3 .claude/skills/_shared/workspace.py --skill troubleshoot --title "{short English title}" --createThis prints one JSON object: slug, team_name, and paths (bug_report, context, root_cause, impact, diagnosis, state_input, team_dir). Exit 0 resolved/created; 1 bad args; 2 applies only to --verify (used later in Phase 3); 3 the workspace directories could not be created. Use {slug}, {team_name}, and every paths.* value from this JSON verbatim for the rest of this skill -- do not re-derive them by hand in a later phase.
Ask the user to provide:
First run the bundled script for the mechanical capture — it runs the failing
command under a deadline, records stdout/stderr/exit code + extracted traceback
to a log file keyed by --label, and gathers recent git history (plus optional
last-commit context for a stack-trace file):
python3 .claude/skills/troubleshoot/repro.py "<repro-command>" \
--label {slug}-initial [--file <path-from-stack-trace>] [--timeout 120]Always pass --label {slug}-initial here: the log path is
.claude/logs/troubleshoot-repro-{label}.log and an unlabelled run reuses one
shared file, so the Phase 3 fix-verification run (Step 2 task 3) would otherwise
overwrite the original failure evidence this whole diagnosis rests on.
Exit codes: 0 capture completed; 1 bad arguments (including an unusable
--label or --bisect-good ref, checked before the command runs); 2 the
observed exit code differs from --expect-exit (not used in Phase 1); 3 the
repro command timed out or the log could not be written. A failing repro command
is the expected case and is still exit 0 — its result is the JSON exit_code.
Read the JSON fields: exit_code, timed_out, stdout_tail, stderr_tail,
traceback, traceback_format, git_available, git_error, recent_commits,
blame, blame_error, bisect, log_file, artifacts. Two fields exist to
stop a null being over-read: traceback is only extracted for CPython
tracebacks (traceback_format: "python"), so null there means "no Python
traceback" — a Node/Go/pytest-assertion stack is in stderr_tail. And
git_available: false with a git_error means history could not be read at
all; that is not the same as "no relevant recent history".
On timed_out: true (exit 3) the command has no usable result: raise the
--timeout, narrow the repro command, or treat the hang itself as the bug —
do not proceed as if the capture succeeded.
To scope a regression, add --bisect-good <last-known-good-ref>. It reports the
bisect object (candidate_commits, candidate_count, path_filter,
bisect_command) — the commits an actual git bisect would search, plus the
command to start it. The script never checks out a commit itself, so driving the
bisect stays the Impact Investigator's call in Phase 2.
Then hand that captured context to general-purpose-opus for the
judgment part — do NOT re-run the command or re-fetch git history:
Task tool:
subagent_type: "general-purpose-opus"
prompt: |
Analyze this reproduced error (already captured by repro.py):
Error: {error message / stack trace}
repro.py JSON: {exit_code, traceback, recent_commits, log_file}
Tasks:
1. Read all files mentioned in the traceback; trace the execution flow
leading to the error and identify the immediate cause (what line fails).
2. Look for related tests and whether they pass/fail.
3. Check if similar patterns exist elsewhere in the codebase.
Use Glob, Grep, and Read tools to investigate thoroughly.
Save analysis to `{paths.context}` (from Phase 1 Step 0).
Return concise summary (5-7 key findings).Consult Codex for initial hypothesis generation before creating the Bug Report. Write the prompt to a file, then invoke the wrapper:
Objective: Analyze this error and generate initial hypotheses for root cause.
Context:
- Error: {error message / stack trace}
- Failing location: {file:line from Opus subagent analysis}
- Execution flow: {call chain from Opus subagent analysis}
Constraints:
- Focus on root cause categories (state mutation, boundary, concurrency, dependency, type/contract)
- Rank hypotheses by likelihood
- Suggest specific code areas to investigate for each hypothesis
Output format:
## Error Pattern Recognition
## Hypotheses (ranked by likelihood)
## Investigation Plan (per hypothesis)
## Known Similar Patternspython3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-initial.md --label troubleshoot-initial.claude/skills/_shared/codex_consult.py exits 0 when Codex answered normally, 2 if the Codex CLI is not installed, 3 if Codex failed or timed out -- check the JSON ok field and read response_file for the answer (error/stderr_file explain a failure). Every later Codex consultation in this skill follows this same write-prompt-then-invoke pattern without repeating these exit codes.
Use Codex's analysis to strengthen the Initial Hypotheses section of the Bug Report.
Combine error details + codebase analysis + Codex initial hypotheses into a Bug Report following the template contract in references/bug-report-template.md. Save it to {paths.bug_report} (from Step 0), then validate it:
python3 .claude/skills/_shared/validate_doc.py --contract bug-report --file {paths.bug_report}references/bug-report-template.md is the single source of truth for the
required sections; the bug-report contract is pinned to that template by
tests/test_validate_doc.py. Do not work from a section list retyped here --
that drift is exactly what this fix removed. Exit 0 means every required section
is present; exit 2 means one is missing, and the JSON sections_missing names
it. Fill the gap before proceeding; exit 1 means the file does not exist.
Both Phase 2 teammates read this file, and Phase 3's --verify gate requires
it, so it must exist on disk -- not only in this conversation.
Launch Root Cause Analyst and Impact Investigator in parallel via Agent Teams with bidirectional communication. Both teammates MUST consult Codex for deep reasoning tasks.
Key difference from subagents: Teammates can communicate with each other. Root Cause Analyst's findings change Impact Investigator's scope, and Impact Investigator's context informs root cause analysis.
Create an agent team named `{team_name}` for troubleshooting: {slug}
Spawn two teammates:
1. **Root Cause Analyst** — Uses Codex CLI as PRIMARY analysis engine for deep code reasoning
Prompt: "You are the Root Cause Analyst for bug: {slug}.
Your job: Identify the definitive root cause of this error through deep code analysis.
Codex CLI is your PRIMARY tool for reasoning about code behavior.
Bug Report: read `{paths.bug_report}` (written and validated in Phase 1 Step 3).
Tasks:
1. Trace the execution flow step by step from entry point to error
2. Evaluate each hypothesis from the Bug Report:
- Gather evidence FOR and AGAINST each hypothesis
- Eliminate hypotheses that contradict the evidence
3. Identify the root cause (not just the symptom):
- What is the underlying defect?
- Why does it manifest as this specific error?
- Under what conditions does it trigger?
4. Propose fix approaches (at least 2 alternatives):
- Approach A: {description, pros, cons}
- Approach B: {description, pros, cons}
- Recommended approach with rationale
## Codex Analysis Protocol (MANDATORY)
You MUST consult Codex for EACH of the following analysis tasks.
Do NOT skip Codex consultation — it is the primary reasoning engine for this role.
Each consultation below follows the same shape: write the prompt to a file,
then run `python3 .claude/skills/_shared/codex_consult.py --prompt-file <path> --label <label>`
and read the JSON `response_file`.
### 1. Execution Flow Tracing
For complex control flow, write the prompt below to a file, then consult Codex:
Objective: Trace the execution flow from {entry point} to {error location}.
Context:
- Entry point: {file:function}
- Error location: {file:line}
- Key intermediate functions: {list}
Constraints:
- Track state transformations at each step
- Identify where assumptions are violated
Output format:
## Execution Flow (step by step)
## State Transformations
## Assumption Violations
## Critical Decision Points
python3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-flow.md --label troubleshoot-flow
### 2. Hypothesis Evaluation
For each hypothesis, write the prompt below to a file, then consult Codex to evaluate evidence:
Objective: Evaluate hypothesis "{hypothesis}" against collected evidence.
Context:
- Hypothesis: {description}
- Evidence FOR: {list}
- Evidence AGAINST: {list}
- Code context: {relevant code snippets}
Constraints:
- Apply logical reasoning, not pattern matching
- Consider alternative explanations for the evidence
Output format:
## Verdict (CONFIRMED / ELIMINATED / INCONCLUSIVE)
## Reasoning
## Remaining Unknowns
python3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-hypothesis.md --label troubleshoot-hypothesis
### 3. Fix Approach Design
Write the prompt below to a file, then consult Codex for trade-off analysis of fix alternatives:
Objective: Design and compare fix approaches for root cause: {root cause description}.
Context:
- Root cause: {description}
- Affected code: {file:line}
- Current behavior: {description}
- Desired behavior: {description}
Constraints:
- Propose at least 2 approaches
- Evaluate: correctness, minimal invasiveness, maintainability, performance
- Consider backward compatibility
Output format:
## Approach A: {name}
## Approach B: {name}
## Comparison Matrix
## Recommendation with Rationale
python3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-fix-design.md --label troubleshoot-fix-design
### 4. Fix Correctness Verification
Before finalizing, write the prompt below to a file, then consult Codex to verify the proposed fix:
Objective: Verify that the proposed fix correctly resolves the root cause.
Context:
- Root cause: {description}
- Proposed fix: {description}
- Edge cases identified: {list}
Constraints:
- Check that the fix addresses the root cause, not just symptoms
- Verify behavior under all identified trigger conditions
- Check for new failure modes introduced by the fix
Output format:
## Correctness Assessment (CORRECT / INCOMPLETE / INCORRECT)
## Edge Case Coverage
## New Failure Modes (if any)
## Confidence Level
python3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-fix-verify.md --label troubleshoot-fix-verify
Save analysis to `{paths.root_cause}` (from Phase 1 Step 0).
Communicate with Impact Investigator teammate:
- Share root cause findings that expand the affected scope
- Request context about specific code paths or history
- Confirm or refute hypotheses based on shared evidence
IMPORTANT — Work Log:
When ALL your tasks are complete, write your work log to
{paths.team_dir}root-cause-analyst.md per the shared
format: .claude/skills/_shared/work-log-format.md
Keep all five core sections, `## Tasks Completed` included -- the Lead
validates this log with `validate_doc.py --contract work-log`, which
rejects a log that drops it.
Role-specific sections (between Tasks Completed and Communication with
Teammates) for this role:
## Hypotheses Evaluated
- [confirmed/eliminated] {hypothesis}: {evidence}
## Root Cause
- Defect: {description}
- Location: {file:line}
- Trigger condition: {when it occurs}
## Proposed Fixes
- Approach A: {description} — {pros/cons}
- Approach B: {description} — {pros/cons}
- Recommended: {which and why}
## Codex Consultations
- {question asked to Codex}: {key insight from response}
"
2. **Impact Investigator** — Uses Opus with Git history, codebase search, WebSearch, and Codex for risk analysis
Prompt: "You are the Impact Investigator for bug: {slug}.
Your job: Determine the full scope and impact of this bug, and gather context for the fix.
Consult Codex for regression risk reasoning and fix safety analysis.
Bug Report: read `{paths.bug_report}` (written and validated in Phase 1 Step 3).
Tasks:
1. Trace the bug's origin in git history:
- git log / git bisect to find the introducing commit
- What change caused this? Was it intentional?
2. Assess blast radius:
- What other code paths call the affected function?
- What features/users are impacted?
- Are there related bugs or similar patterns elsewhere?
3. Research external context:
- Is this a known issue in a dependency? (WebSearch)
- Are there upstream fixes or workarounds?
- Check issue trackers, changelogs, migration guides
4. Evaluate regression risk:
- What tests cover the affected area?
- What could break if we change this code?
- Are there downstream consumers to consider?
How to research:
- Use Git commands (git log, git blame, git bisect) for history
- Use Grep/Glob for codebase impact analysis
- Use WebSearch for external known issues:
WebSearch: '{library} {error message} issue fix'
## Codex Risk Analysis Protocol (MANDATORY)
You MUST consult Codex for regression risk reasoning and fix safety analysis.
Each consultation below follows the same shape: write the prompt to a file,
then run `python3 .claude/skills/_shared/codex_consult.py --prompt-file <path> --label <label>`
and read the JSON `response_file`.
### Regression Risk Reasoning
Write the prompt below to a file, then consult Codex to evaluate what could break if the proposed change is applied:
Objective: Evaluate regression risk if {proposed change} is applied to {file:line}.
Context:
- Current behavior: {description}
- Proposed change: {description}
- Callers of affected function: {list}
- Existing test coverage: {description}
Constraints:
- Consider all callers and downstream consumers
- Identify implicit contracts that may be violated
- Assess backward compatibility impact
Output format:
## Risk Assessment (HIGH / MEDIUM / LOW)
## Affected Code Paths
## Implicit Contracts at Risk
## Recommended Safeguards
python3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-regression.md --label troubleshoot-regression
### Fix Safety Analysis
Write the prompt below to a file, then consult Codex to verify the proposed fix does not introduce new issues:
Objective: Analyze whether the proposed fix introduces new issues or side effects.
Context:
- Root cause: {from Root Cause Analyst}
- Proposed fix: {description}
- Blast radius: {affected code paths}
- Dependencies: {upstream/downstream}
Constraints:
- Check for new edge cases created by the fix
- Verify thread safety if applicable
- Check for performance implications
Output format:
## Safety Assessment (SAFE / CAUTION / UNSAFE)
## New Issues Identified
## Side Effects
## Mitigation Recommendations
python3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-fix-safety.md --label troubleshoot-fix-safety
Save findings to `{paths.impact}` (from Phase 1 Step 0).
Communicate with Root Cause Analyst teammate:
- Share git history context that informs root cause
- Share external findings (known issues, upstream fixes)
- Request clarification on which code paths to investigate
IMPORTANT — Work Log:
When ALL your tasks are complete, write your work log to
{paths.team_dir}impact-investigator.md per the shared
format: .claude/skills/_shared/work-log-format.md
Keep all five core sections, `## Tasks Completed` included -- the Lead
validates this log with `validate_doc.py --contract work-log`, which
rejects a log that drops it.
Role-specific sections (between Tasks Completed and Communication with
Teammates) for this role:
## Git History
- Introducing commit: {hash} — {description}
- Related commits: {list}
## Blast Radius
- Affected code paths: {list}
- Affected features/users: {list}
## External Research
- {source}: {finding and relevance}
## Regression Risk
- Existing test coverage: {description}
- Risk areas: {what could break}
## Codex Risk Analysis
- Regression risk assessment: {Codex's verdict and reasoning}
- Fix safety assessment: {Codex's verdict and reasoning}
"
Wait for both teammates to complete their tasks.Example interaction flow:
Root Cause Analyst: "The error occurs because parse_config() returns None when key is missing"
-> Impact Investigator: "Checking git blame -- this was changed in commit abc123"
-> Impact Investigator: "Found 5 other callers of parse_config() that don't handle None"
-> Root Cause Analyst: "Expanding fix scope -- need to either fix callers or fix parse_config()"
-> Root Cause Analyst: "Codex recommends: fix parse_config() to raise KeyError instead of returning None"
-> Impact Investigator: "Codex risk analysis confirms: all 5 callers already have try/except for KeyError"
-> Root Cause Analyst: "Root cause confirmed. Codex verified fix correctness. Fix approach: restore KeyError in parse_config()"Without Agent Teams, this discovery loop would require multiple sequential subagent rounds.
Integrate Agent Teams diagnosis results, validate the fix plan with Codex, and request user approval.
Gate Phase 3 on the Phase 1/2 artifacts before reading anything, so a teammate that stopped early cannot be mistaken for one that finished:
python3 .claude/skills/_shared/workspace.py --skill troubleshoot --slug {slug} --verify
python3 .claude/skills/_shared/validate_doc.py --contract work-log \
--dir {paths.team_dir} --expect-files 2The first call exits 0 when bug_report, context, root_cause, and impact are all present and non-trivial; exit 2 means one is missing or empty (read verify.missing / verify.empty). The second exits 0 only when both teammate logs exist and satisfy the work-log contract; exit 2 means a log is missing (error: "expected 2 files, found N") or malformed (files_failed > 0, with sections_missing per file). "Wait for both teammates to complete" is a self-report; these two commands are the check. Resolve every gap before continuing.
Read outputs from Phase 2:
{paths.root_cause} -- Root cause analysis{paths.impact} -- Impact assessmentBefore presenting to the user, validate the fix plan with Codex. Write the prompt to a file, then invoke the wrapper:
Objective: Validate this fix plan for completeness and correctness.
Context:
- Root cause: {from Root Cause Analyst}
- Proposed fix: {recommended approach}
- Blast radius: {from Impact Investigator}
- Fix tasks: {task list}
Constraints:
- Check for missing edge cases
- Verify the fix addresses the root cause (not just symptoms)
- Identify potential new issues the fix could introduce
- Suggest additional test cases if needed
Output format:
## Validation Result (PASS / NEEDS_REVISION)
## Missing Coverage
## Potential New Issues
## Additional Test Cases Recommended
## Revised Task List (if needed)python3 .claude/skills/_shared/codex_consult.py --prompt-file .claude/logs/codex/prompt-troubleshoot-plan-validation.md --label troubleshoot-plan-validationIf Codex returns NEEDS_REVISION, update the fix plan before presenting to user.
Create task list using TodoWrite:
{
"content": "Fix {specific task}",
"activeForm": "Fixing {specific task}",
"status": "pending"
}Task breakdown should follow references/debug-patterns.md.
Typical fix task structure:
Write failing test -- Reproduce the bug as a test case
Apply fix -- Implement the root cause fix
Verify fix -- Re-run the original repro command with the expectation asserted by the script rather than read by eye, and with its own label so the Phase 1 failure log survives:
python3 .claude/skills/troubleshoot/repro.py "<repro-command>" \
--label {slug}-fix-verify --expect-exit 0Exit 0 is the verification. Exit 2 means the fix is not verified
(error: "expected exit 0, got N"); exit 3 means it timed out and nothing
was verified at all. Do not report a verified fix on any exit code but 0.
Check regressions -- Run the quality gates:
bash .claude/skills/_shared/verify.shRead the JSON: overall is pass / fail / no_gates. Exit 0 is a pass; exit 2 is a gate failure or no_gates -- inspect log_file and the per-tool tools object. no_gates means zero gates actually ran, which is a contract violation, not a pass: fall back to the project's own verification commands and confirm manually, and pass --allow-no-gates only when you have done so deliberately. Quote the tools object rather than re-typing each status, so a skipped gate is never reported as a pass.
Fix collateral damage -- Address blast radius items (if any)
Add bug context to .claude/STATE.md for cross-session persistence using the
shared writer script and .claude/rules/agent-state.md.
Gather these fields from the diagnosis:
Write the input JSON to {paths.state_input} (from Phase 1 Step 0):
{
"title": "{slug}",
"sections": [
{"heading": "Context", "content": "- Error: ...\n- Root cause: ...\n- Affected files: ..."},
{"heading": "Fix Approach", "content": "- {approach}"},
{"heading": "Regression Risks", "content": "- {risks}"},
{"heading": "Decisions", "content": "- {Decision 1}: {rationale}"}
]
}Run dry-run, review the preview, then apply:
python3 .claude/skills/_shared/append_state_block.py \
--type bug-fix --input {paths.state_input}
# Review the preview file path in the JSON output, then:
python3 .claude/skills/_shared/append_state_block.py \
--type bug-fix --input {paths.state_input} --applyVerify "ok": true and "progress_tracker_preserved": true in the output.
Exit code 2 means the state structure is invalid; stop before writing.
Compose the diagnosis and fix plan following the template contract in
references/diagnosis-template.md. Write the composed presentation to
{paths.diagnosis} (resolved in Phase 1 Step 0, never hand-built) and validate
its structure before presenting it:
python3 .claude/skills/_shared/validate_doc.py --contract diagnosis \
--file {paths.diagnosis}
python3 .claude/skills/_shared/workspace.py --skill troubleshoot \
--slug {slug} --verify --require diagnosisreferences/diagnosis-template.md is the section source of truth (the
diagnosis contract is pinned to it by tests/test_validate_doc.py). Exit 0
means every required section is present; exit 2 means one is missing and
sections_missing names it -- most often Alternative Approaches Considered,
the section a rushed presentation drops. Then present the validated document to
the user. Structure is all this gate checks: whether the diagnosis is correct
remains the Codex validation in Step 1.5 plus the user's approval.
Paths resolved once in Phase 1 Step 0 (.claude/skills/_shared/workspace.py --skill troubleshoot):
| File | Author | Purpose |
|---|---|---|
{paths.bug_report} | Lead | Bug Report (Phase 1 synthesis) |
{paths.context} | Opus Subagent | Initial error context analysis |
{paths.root_cause} | Root Cause Analyst | Root cause analysis (Codex-driven) |
{paths.impact} | Impact Investigator | Impact assessment (with Codex risk analysis) |
.claude/STATE.md (updated) | Lead | Cross-session bug fix context |
| Task list (internal) | Lead | Fix implementation tracking |
Run artifacts under .claude/logs/, keyed by {slug} so successive runs do not
overwrite each other. The repro logs are not part of --verify; the diagnosis is
checkable on demand with --require diagnosis (it is not a default required key,
because in Phases 1-2 it does not exist yet):
| File | Author | Purpose |
|---|---|---|
.claude/logs/troubleshoot-repro-{slug}-initial.log | repro.py | Phase 1 failure capture |
.claude/logs/troubleshoot-repro-{slug}-fix-verify.log | repro.py | Phase 3 fix-verification capture |
{paths.diagnosis} | Lead | Phase 3 presentation, validated with --contract diagnosis and --require diagnosis |
/team-execute/team-execute Phase 2 competing hypotheses pattern)© DeL-TaiseiOzaki, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 4 other files (references) in .claude/skills/troubleshoot of DeL-TaiseiOzaki/claude-code-orchestra.
Open the folder on GitHubat commit ef0d8f8
Troubleshoot next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Troubleshoot this skillDeL-TaiseiOzaki/claude-code-orchestra | 199 | — | ~7.6k | Automated safety check: Pass | MIT | |
| Analyze Trajectoryyologdev/yoyo-evolve | 1.9k | — | ~3.6k | Automated safety check: Pass | MIT | |
| Bug Hunt SwarmDimillian/Skills | 4k | — | ~1.6k | Automated safety check: Pass | MIT | |
| Diagnosing Superpowers SessionsjnMetaCode/superpowers-zh | 8.3k | — | ~858 | Automated safety check: Pass | MIT | |
| Researchwarpdotdev/common-skills | 610 | 1 repos | ~1.3k | Automated safety check: Pass | MIT | |
| Bug Hunt Swarmsickn33/agentic-awesome-skills | 47k | 1 repos | ~1.9k | Automated safety check: Pass | MIT |
yologdev/yoyo-evolve
Diagnoses a recurring failure such as a stuck task, repeated CI error or frequent reverts by sending sub-agents through the logs and returning one root-cause diagnosis.
Dimillian/Skills
Parallel read-only multi-agent root-cause investigation for bugs, regressions, crashes, flaky behavior, or unexplained failures.
jnMetaCode/superpowers-zh
Investigates what went wrong in a superpowers session by reading its transcript, reports findings with path and line citations, and can draft a GitHub issue or redacted bundle.
warpdotdev/common-skills
Delegate noisy investigation to one or more subagents so the orchestrator's context stays clean, then work from the distilled answer.
sickn33/agentic-awesome-skills
Parallel read-only multi-agent root-cause investigation for bugs, regressions, crashes, flaky behavior, or unexplained failures.
ZaxbyHub/opencode-swarm
Apply when implementing features, fixing bugs, debugging errors, investigating failures, tracing root causes, reviewing tech debt, tracing issues, planning fixes, or completing any task.
DeL-TaiseiOzaki/claude-code-orchestra
Codex CLI handles planning, design, and complex code implementation.
DeL-TaiseiOzaki/claude-code-orchestra
Record a project design decision into .claude/docs/DESIGN.md through the shared typed writer.
DeL-TaiseiOzaki/claude-code-orchestra
Unified feature planning & implementation skill — replaces the old /add-feature and /start-feature skills (both trigger phrases still apply here).
DeL-TaiseiOzaki/claude-code-orchestra
Create a detailed implementation plan for a feature or task.
DeL-TaiseiOzaki/claude-code-orchestra
Comprehensive onboarding for new or returning contributors. An agent skill from DeL-TaiseiOzaki/claude-code-orchestra.
DeL-TaiseiOzaki/claude-code-orchestra
Save session activity, rebuild rolling PROGRESS.md, and compact stale working blocks in .claude/STATE.md.
Categories
Diagnose and plan fixes for errors/bugs with Codex-first multi-agent collaboration (Codex + Opus 4.6 + Agent Teams). Troubleshoot is an agent skill from DeL-TaiseiOzaki/claude-code-orchestra.6 + Agent Teams).
Troubleshoot fits situations like: tasks that involve Subagents; tasks that involve Root cause analysis.
Run `npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a claude-code`. Or copy the skill folder (.claude/skills/troubleshoot in DeL-TaiseiOzaki/claude-code-orchestra) into .claude/skills/troubleshoot in your project. Claude Code loads it when a task matches its description.
Run `npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a codex`. Or copy the skill folder (.claude/skills/troubleshoot in DeL-TaiseiOzaki/claude-code-orchestra) into .agents/skills/troubleshoot in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add DeL-TaiseiOzaki/claude-code-orchestra --skill troubleshoot -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/troubleshoot, .gemini/skills/troubleshoot, .github/skills/troubleshoot and .opencode/skills/troubleshoot in your project.
Going by SKILL.md and its folder, Troubleshoot needs Python for the scripts in its folder and the command-line tools its instructions call (python3, claude, codex, bash and git). Our summary lists: Python 3.
SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Troubleshoot is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 7.6k tokens (SKILL.md is roughly 30k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.9k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Troubleshoot: Analyze Trajectory (yologdev/yoyo-evolve, 1.9k stars), Bug Hunt Swarm (Dimillian/Skills, 4k stars), Diagnosing Superpowers Sessions (jnMetaCode/superpowers-zh, 8.3k stars) and Research (warpdotdev/common-skills, 610 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
DeL-TaiseiOzaki (a GitHub user) maintains it in DeL-TaiseiOzaki/claude-code-orchestra, which has 199 GitHub stars. The repository holds 15 skills in this directory. The repository was last updated on September 20, 2026.
Source: DeL-TaiseiOzaki/claude-code-orchestra on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.