Agent skill

Octocode Code Research

by bgauryy in bgauryy/octocode

Researches code with evidence: traces callers, imports and cross-repo links, diagnoses failures and reports findings with exact file and line references and a confidence label.

MITAuto-check passedDevelopment

Install Octocode Code Research

skills CLI
$ npx skills add bgauryy/octocode --skill octocode-research -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install bgauryy/octocode octocode-research --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/bgauryy/octocode.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/octocode-research .claude/skills/octocode-research && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
octocode-research
GitHub stars
949
Token cost
~1.5k tokens
SKILL.md length
603 words
Files
25 (incl. scripts, references)
Skills in repo
12
Repo updated
First seen
Licence
MIT

At a glance

Researches code with evidence: traces callers, imports and cross-repo links, diagnoses failures and reports findings with exact file and line references and a confidence label.

  • Works in 7 steps: Open with one line: corpus, actual vs… → Call it a bug only when evidence shows a… → Root cause needs mechanism, trigger,… → …
  • Tracing what calls a function and what breaks if you change it
  • SKILL.md covers The rules, Workflows, Tooling and Output, plus 1 more section
  • Calls npx and node

What it does

This skill makes the agent prove claims about code before it answers or patches. It follows a ladder of frame, classify, model, search, read exact, prove, decide or patch, and verify, scaled to the claim: a small lookup gets a cheap read and an honest confidence label, while a delete, a merge verdict or a root cause gets the whole ladder. Skipped stages are named.

The rules include opening with one line on corpus, task class and surfaces used, calling something a bug only when a supported contract was violated, and treating an empty search result as a blind lane, not proof of absence. Findings are tracked as claim, evidence, confidence and next check, and the agent asks before broad contracts, deletes or renames.

Reference files give routes for the local repo, external repositories and npm packages, cross-repo connections, debugging, change planning and pull request review analysis. The work stops at a budget of 3-5 decisive iterations or about 15 minutes. Writing the docs themselves goes to octocode-documentation, and building skill folders to octocode-skills.

When your agent uses it

  • Tracing what calls a function and what breaks if you change it
  • Finding the root cause of a failure with evidence from the code
  • Researching an external repository or npm package before depending on it
  • Planning a change before writing code and validating it afterwards

Example prompts

  • “Research this: where is the retry logic for failed webhooks and who calls it?”
  • “Use octocode to find out what breaks if I rename the parseConfig function.”
  • “Find the root cause of the intermittent timeout in the upload service, with file and line evidence.”
  • “Look into how the upstream library handles token refresh and compare it to our wrapper.”

Requirements

  • The Octocode toolset, used through MCP or its CLI

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. Open with one line: corpus, actual vs desired, task class, mode, surfaces used and skipped.
  2. Call it a bug only when evidence shows a supported contract was violated.
  3. Root cause needs mechanism, trigger, violated contract, divergence boundary, and a killed alternate.
  4. Use the strongest handle you already hold. For nontrivial claims check two of: structure, stream, connections.
  5. A snippet is a lead. Empty means this lane can't see it, not it isn't there. Say which you mean.
  6. Track claim → evidence → confidence → next check. Cite exact anchors, and only checks that actually ran.
  7. Ask before broad contracts, deletes/renames, thin evidence, or a third unrelated search space. Patch after proof.

What it can do on your machine

Read from SKILL.md and the folder at commit c265e3f. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/, which the agent can run.

    Shell commands in SKILL.md call:

    • npx
    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Octocode Code Research loads about 1.5k tokens when it runs, and up to ~15k if it reads all its reference files. Until then it costs about 167 tokens; SKILL.md has 603 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~167
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~15k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from bgauryy/octocode at commit c265e3f, republished under its MIT licence (© bgauryy). 603 words, ~1,473 tokens.

Download SKILL.mdSave it as .claude/skills/octocode-research/SKILL.md (or your agent's skills folder). This skill also uses 24 other files; get the full folder from GitHub.
name
octocode-research
description
Use when code must be checked, not assumed: trace connections (callers, imports, cross-repo wiring, what breaks if I change it), locate behavior, map a system, diagnose a failure, RCA. Also for external repositories, npm packages, upstream/prior art, and general research. Plan a coding flow before writing, validate it after. Triggers on 'research this' or 'use octocode'. Gives exact file:line/PR/commit evidence with confidence. Skip trivial edits whose blast radius is already known. Not for authoring or copyediting the docs themselves, or for building skill folders: docs deliverable → octocode-documentation; SKILL.md folders → octocode-skills.

Octocode Research

Evidence before assertion. Find the anchor, read the exact bytes, prove the claim, then answer or patch.

text
FRAME → CLASSIFY → MODEL → SEARCH → READ EXACT → PROVE → DECIDE/PATCH → VERIFY

That's the full shape, not a checklist to march through. Scale it to the claim: a small lookup gets a cheap read and an honest confidence label; a delete, a merge verdict, or a root cause earns the whole ladder. Skip stages you can already answer — but say which ones you skipped.

The rules

  1. Open with one line: corpus, actual vs desired, task class, mode, surfaces used and skipped.
  2. Call it a bug only when evidence shows a supported contract was violated.
  3. Root cause needs mechanism, trigger, violated contract, divergence boundary, and a killed alternate.
  4. Use the strongest handle you already hold. For nontrivial claims check two of: structure, stream, connections.
  5. A snippet is a lead. Empty means this lane can't see it, not it isn't there. Say which you mean.
  6. Track claim → evidence → confidence → next check. Cite exact anchors, and only checks that actually ran.
  7. Ask before broad contracts, deletes/renames, thin evidence, or a third unrelated search space. Patch after proof.

Stop when: grounded evidence answers the framed question and the alternate is dead; no cheap next step can change the conclusion; the budget is hit (default 3-5 decisive iterations or ~15 minutes); the last iterations changed no state; retries stay thin, or a license/product/architecture call belongs to the user; a gate blocks (broad contract, delete/rename, clone or run untrusted code, unapproved artifact write); or a skill edit measured flat/worse — revert through references/improve-loop.md. Report the remaining gaps instead of padding certainty.

Workflows

Start with references/algorithm.md (routing, evidence grades) and references/problem-framing.md (is this a bug, feature, enhancement, or still unknown?). Then pick one route by what you're looking at — load references/workflows.md when you need the per-route detail, the load budget per task size, or the handoff receipt between routes:

SituationRoute
This repo, checkout, installed dependencyreferences/workflow-local.md
External repository, npm package, upstream projectreferences/workflow-external.md
Cross-repo connections, local clue → upstream, or remote code needing AST/LSP proofreferences/workflow-combination.md
Rank an ecosystem — several candidate repos/packagesreferences/github-landscape.md
Something fails and you need the causereferences/workflow-debug.md
Plan a coding flow, implement, migrate, patch behaviorreferences/workflow-change.md
Reshape structure or names, keep behaviorreferences/workflow-refactor.md
Validate after a change; review a PR or local diffreferences/workflow-pr-review.md
Trace connections — callers, imports, references, reachabilityreferences/code-research.md
Show full SKILL.md (212 more words)Show less

Proof depth for any of them: references/code-research.md. General research — Map / Validate / Investigate / Plan across code, packages, docs, and history: references/research-flow.md.

Review runs in three parts: use references/workflow-pr-review.md for target, guidelines, and risk sizing; then references/workflow-pr-review-analysis.md for sizing depth, flow proof, and finding shape; then references/workflow-pr-review-report.md for the verification-gated recommendation and the optional written document.

Reach for these only when they earn it: references/loop-mode.md (evidence keeps flipping), references/long-research.md (durable decision brief), references/researcher-mindset.md (budgets, fan-out, campaign planning).

Load a reference when the current step needs it. Loading all of them is a failure mode.

Tooling

Prefer Octocode MCP tools when exposed. Otherwise npx octocode tools <name> — same 15 tools, same schemas, no loss.

bash
npx octocode context --minimal                         # what's available
npx octocode tools <name> --scheme --json --compact    # read fields — never guess
npx octocode tools <name> --queries '<json>' --compact # run it

Batch up to five queries per call. Orient cheap (tree, discovery) before exact reads. Follow returned next.* and cursors instead of re-deriving them.

Read references/octocode.md when transport, tool choice, auth, gates (ENABLE_LOCAL, ENABLE_CLONE, ENABLE_RELEASES), materialization, diagnostics, or exit codes are unclear.

Output

Finding · Evidence · Confidence · Next. Decisions add verdict, risks, exact anchors, verification, and the smallest safe fix. Report gaps instead of padding certainty.

octocode-brainstorming (worth building?) · octocode-rfc-generator (design contract) · octocode-graph-eval (goal→KPI) · octocode-documentation (docs deliverable) · octocode-skills (skill folders) · octocode-subagent (fan-out) · octocode-roast (critique tone).

When changing this skill, run node scripts/check-description.mjs (description contract; --help for flags) and gate accept/revert with references/improve-loop.md.

© bgauryy, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 24 other files (scripts, references) in skills/octocode-research of bgauryy/octocode.

  • SKILL.md
  • README.md
  • agents/openai.yaml
  • references/algorithm.md
  • references/code-research.md
  • references/github-landscape.md
  • references/improve-loop.md
  • references/long-research.md
  • references/loop-mode.md
  • references/octocode.md
  • references/problem-framing.md
  • references/research-flow.md
  • references/researcher-mindset.md
  • references/workflow-change.md
  • references/workflow-combination.md
  • references/workflow-debug.md
  • references/workflow-external.md
  • references/workflow-local.md
  • references/workflow-pr-review-analysis.md
  • … and 6 more

Open the folder on GitHubat commit c265e3f

Compare with similar skills

Octocode Code Research next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Octocode Code Research compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Octocode Code Research this skillbgauryy/octocode949—~1.5kAutomated safety check: PassMIT
Orca Run Replayiflytek/skillhub5.2k4 repos~3kAutomated safety check: PassApache-2.0
Fix Issueyonatangross/orchestkit292—~6.2kAutomated safety check: NotesMIT
Cursor Agents Workflowlangfuse/langfuse36k—~1.7kAutomated safety check: PassCustom licence
Project Pull Requestswimmwatch/cloakbrowser-mcp164—~1kAutomated safety check: PassMIT
Triagearcee-ai/nac281—~2kAutomated safety check: PassApache-2.0

Similar skills

  • Orca Run Replay

    iflytek/skillhub

    Answers questions about a past agent run from its recording, using causal graphs and replay, instead of reconstructing events from memory.

    5.2k GitHub starsUsed in 4 repos~3k tokens
    Agent WorkflowsAuto-check passed
  • Fix Issue

    yonatangross/orchestkit

    Fixes GitHub issues using parallel analysis agents for root cause investigation, code exploration, and regression detection.

    292 GitHub stars~6.2k tokensUpdated yesterday
    DevelopmentAuto-check: notes
  • Cursor Agents Workflow

    langfuse/langfuse

    Human handoff, Linear branch names, reviewable (non-draft) PRs, the cursor GitHub label, Claude, Greptile, and Codex review comments, preview test steps, proof of work posted on the GitHub PR, and…

    36k GitHub stars~1.7k tokensUpdated today
    DevelopmentAuto-check passed
  • Project Pull Request

    swimmwatch/cloakbrowser-mcp

    Create, update, prepare, or review a cloakbrowser-mcp GitHub Pull Request only when the user explicitly requests PR work.

    164 GitHub stars~1k tokensUpdated 2 days ago
    DevelopmentAuto-check passed
  • Triage

    arcee-ai/nac

    Triage a GitHub repository's open issues by finding exact duplicates, rejecting evidenceably off-base requests, requesting concrete clarification, applying only existing labels, and opening a linked…

    281 GitHub stars~2k tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Flowstudio Power Automate Debug

    github/awesome-copilot

    Official

    Debug failing Power Automate cloud flows using the FlowStudio MCP server.

    40k GitHub starsUsed in 2 repos~5k tokens
    DevelopmentAuto-check passed

More from bgauryy/octocode

All 12 skills in this repo
  • Runs blind pairwise comparisons of Octocode against a gh-based baseline over markdown research questions, scored by total characters through the model rather than self-report.

    949 GitHub stars~2.1k tokensUpdated 2 days ago
    Auto-check passed
  • Writes, repairs and copyedits project docs against the Google developer documentation style guide, verifying claims in the repository before stating them.

    949 GitHub stars~2k tokensUpdated 2 days ago
    Auto-check passed
  • Octocode Mannequin

    bgauryy/octocode

    Poses and animates a 22-bone anatomical humanoid rig with joint range-of-motion limits, using a Node CLI, a Three.js viewer and WebMCP tools an agent can drive live.

    949 GitHub stars~1.2k tokensUpdated 2 days ago
    Auto-check passed
  • Octocode Skills Manager

    bgauryy/octocode

    Finds, rates, reviews, creates, improves, installs and syncs Agent Skill folders from local workspaces, registries or remote sources, with a user gate before any write.

    949 GitHub stars~1.2k tokensUpdated 2 days ago
    Auto-check passed
  • Octocode Brainstorming

    bgauryy/octocode

    Walks an idea through framing, diverging into options, researching evidence and stress-testing before converging on a build, prototype, narrow or park decision.

    949 GitHub stars~1.3k tokensUpdated 2 days ago
    Auto-check passed
  • Octocode Chrome Devtools

    bgauryy/octocode

    A skill your agent uses when a live page needs Chrome DevTools/CDP evidence: network failures, console errors, performance, DOM/CSS actionability, screenshots/PDF, cookies/storage…

    949 GitHub stars~1.6k tokensUpdated 2 days ago
    Auto-check passed

Questions about Octocode Code Research

What does Octocode Code Research do?

Researches code with evidence: traces callers, imports and cross-repo links, diagnoses failures and reports findings with exact file and line references and a confidence label. This skill makes the agent prove claims about code before it answers or patches. It follows a ladder of frame, classify, model, search, read exact, prove, decide or patch, and verify, scaled to the claim: a small lookup gets a cheap read and an honest confidence label, while a delete, a merge verdict or a root cause gets the whole ladder.

When should I use Octocode Code Research?

Octocode Code Research fits situations like: tracing what calls a function and what breaks if you change it; finding the root cause of a failure with evidence from the code; researching an external repository or npm package before depending on it; planning a change before writing code and validating it afterwards.

How do I install Octocode Code Research in Claude Code?

Run `npx skills add bgauryy/octocode --skill octocode-research -a claude-code`. Or copy the skill folder (skills/octocode-research in bgauryy/octocode) into .claude/skills/octocode-research in your project. Claude Code loads it when a task matches its description.

How do I install Octocode Code Research in Codex?

Run `npx skills add bgauryy/octocode --skill octocode-research -a codex`. Or copy the skill folder (skills/octocode-research in bgauryy/octocode) into .agents/skills/octocode-research in your project. Codex loads it when a task matches its description.

Can I use Octocode Code Research in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add bgauryy/octocode --skill octocode-research -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/octocode-research, .gemini/skills/octocode-research, .github/skills/octocode-research and .opencode/skills/octocode-research in your project.

What does Octocode Code Research need to run?

Going by SKILL.md and its folder, Octocode Code Research needs the command-line tools its instructions call (npx and node). Our summary lists: The Octocode toolset, used through MCP or its CLI.

Does Octocode Code Research access the network?

SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Octocode Code Research safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Octocode Code Research use?

Octocode Code Research is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Octocode Code Research use?

About 1.5k tokens (SKILL.md is roughly 5.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 14k tokens, read only when the agent opens those files.

What are the alternatives to Octocode Code Research?

Skills that share tags, products or a category with Octocode Code Research: Orca Run Replay (iflytek/skillhub, 5.2k stars), Fix Issue (yonatangross/orchestkit, 292 stars), Cursor Agents Workflow (langfuse/langfuse, 36k stars) and Project Pull Request (swimmwatch/cloakbrowser-mcp, 164 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Octocode Code Research?

bgauryy (a GitHub user) maintains it in bgauryy/octocode, which has 949 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on October 9, 2026.

Source: bgauryy/octocode on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.