Agent skill

Autoresearch

by yangyuan-zhen in yangyuan-zhen/PolyWeather

[OMX] Stateful validator-gated research loop with native-hook persistence

AGPL-3.0Auto-check passedAgent Workflows

Install Autoresearch

skills CLI
$ npx skills add yangyuan-zhen/PolyWeather --skill autoresearch -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install yangyuan-zhen/PolyWeather autoresearch --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/yangyuan-zhen/PolyWeather.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.codex/skills/autoresearch .claude/skills/autoresearch && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
autoresearch
GitHub stars
316
Token cost
~786 tokens
SKILL.md length
335 words
Files
1
Skills in repo
26
Repo updated
First seen
Licence
AGPL-3.0

At a glance

[OMX] Stateful validator-gated research loop with native-hook persistence

  • Works in 4 steps: Init chooses validation mode. Pick… → Persist mode state in… → Completion is artifact-gated. The loop… → …
  • Tasks that involve Autonomous loops
  • SKILL.md covers Boundary with planning research, Use when, Do not use when and Core contract, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Autoresearch is an agent skill from yangyuan-zhen/PolyWeather. [OMX] Stateful validator-gated research loop with native-hook persistence

Its SKILL.md is about 790 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Agent Workflows, covering Autonomous loops. The repository describes itself as: polymarket Intelligent Weather Quant Analysis Bot. The licence is AGPL-3.0.

When your agent uses it

  • Tasks that involve Autonomous loops

Example prompts

  • “/autoresearch”

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Init chooses validation mode. Pick exactly one
  2. Persist mode state in .omx/state/.../autoresearch-state.json including
  3. Completion is artifact-gated. The loop does not stop because the model says “done”, because a stop hook fired once, or because several…
  4. Direct CLI launch is gone. Use $deep-interview --autoresearch for intake and $autoresearch for execution.

What it can do on your machine

Read from SKILL.md and the folder at commit 43e658b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are json).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Autoresearch loads about 786 tokens when it runs. Until then it costs about 22 tokens; SKILL.md has 335 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~22
When it runs · the whole SKILL.md, loaded when a task matches
~786

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from yangyuan-zhen/PolyWeather at commit 43e658b, republished under its AGPL-3.0 licence (© yangyuan-zhen). 335 words, ~786 tokens.

Download SKILL.mdSave it as .claude/skills/autoresearch/SKILL.md (or your agent's skills folder).
name
autoresearch
description
[OMX] Stateful validator-gated research loop with native-hook persistence

Autoresearch

Autoresearch is the skill-first replacement for the deprecated omx autoresearch command. It keeps the useful measured-research loop, but it now runs as a native-hook stateful workflow instead of a direct CLI or tmux launch surface.

Boundary with planning research

Use $autoresearch when the research output itself is a bounded deliverable that must pass an explicit validator. Do not recommend it for ordinary pre-planning docs lookup or general best-practice checks; use $best-practice-research for that. If $autoresearch is intentionally run before architecture planning, its approved artifact should feed evidence into $ralplan; it should not become a final architecture/component unless the user explicitly asks for ongoing research automation.

Use when

  • You want a Ralph-ish persistent research loop
  • The task should keep nudging until explicit validation evidence exists
  • You want init-time choice between script validation and prompt+architect validation

Do not use when

  • You want the old omx autoresearch command surface (hard-deprecated)
  • You want detached tmux or split-pane launch parity
  • You have not decided the validation regime yet

Core contract

  1. Init chooses validation mode. Pick exactly one:
    • mission-validator-script
    • prompt-architect-artifact
  2. Persist mode state in .omx/state/.../autoresearch-state.json including:
    • validation_mode
    • completion_artifact_path
    • mission_validator_command or validator_prompt
    • optional output_artifact_path
  3. Completion is artifact-gated. The loop does not stop because the model says “done”, because a stop hook fired once, or because several turns were no-ops.
  4. Direct CLI launch is gone. Use $deep-interview --autoresearch for intake and $autoresearch for execution.

Completion artifact contract

mission-validator-script

The completion artifact must exist and record a passing validator result, for example:

json
{
  "status": "passed",
  "passed": true,
  "summary": "metric improved beyond baseline"
}
prompt-architect-artifact

The completion artifact must include both an architect approval verdict and an output artifact path, for example:

json
{
  "validator_prompt": "Review the research output against the mission.",
  "architect_review": { "verdict": "approved" },
  "output_artifact_path": ".omx/specs/autoresearch-demo/report.md"
}
  1. Run $deep-interview --autoresearch to clarify mission + evaluator.
  2. Materialize .omx/specs/autoresearch-{slug}/mission.md, sandbox.md, and result.json.
  3. Start $autoresearch with the chosen validation mode stored in mode state.
  4. Let stop-hook / auto-nudge continue until the completion artifact satisfies the chosen validation mode.
  5. Finish only after the validator artifact is complete.

Migration note

  • omx autoresearch is hard-deprecated.
  • No direct CLI launch.
  • No tmux split-pane launch.
  • No noop-count completion gate.

© yangyuan-zhen, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .codex/skills/autoresearch of yangyuan-zhen/PolyWeather.

Open the folder on GitHubat commit 43e658b

Compare with similar skills

Autoresearch next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Autoresearch compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Autoresearch this skillyangyuan-zhen/PolyWeather316—~786Automated safety check: PassAGPL-3.0
Show Me Your Work Decision Logcursor/plugins11k8 repos~1.6kAutomated safety check: PassNone
Autoresearch Iteration Loopuditgoenka/autoresearch6.5k1 repos~2kAutomated safety check: PassMIT
Install Loop Engineeringcobusgreyling/loop-engineering11k1 repos~648Automated safety check: PassMIT
LoopyForward-Future/loopy3.2k—~3.9kAutomated safety check: PassMIT
AI Performance Improvement Plantanweai/pua20k2 repos~6.9kAutomated safety check: PassMIT

Similar skills

  • Official

    Keeps a TSV decision log for long or unattended agent runs, one row per decision with what, why, evidence and result, so a reviewer can check the work later.

    11k GitHub starsUsed in 8 repos~1.6k tokens
    Agent WorkflowsAuto-check passed
  • Autoresearch Iteration Loop

    uditgoenka/autoresearch

    Runs an autonomous modify, verify, keep-or-discard loop against any metric, with subcommands for planning, debugging, fixing, security audits, shipping and more.

    6.5k GitHub starsUsed in 1 repo~2k tokens
    Agent WorkflowsAuto-check passed
  • Install Loop Engineering

    cobusgreyling/loop-engineering

    Installs Loop Engineering into a project through the single @cobusgreyling/loop CLI, scaffolding a report-only loop and a readiness score.

    11k GitHub starsUsed in 1 repo~648 tokens
    Agent WorkflowsAuto-check passed
  • Loopy

    Forward-Future/loopy

    Discover, find, compare, audit, repair, adapt, craft, run, debrief, save, and prepare repeatable AI-agent loops for publication.

    3.2k GitHub stars~3.9k tokensUpdated 1 mo ago
    Agent WorkflowsAuto-check passed
  • Pushes an agent to exhaust every option, investigate before asking and take initiative beyond the literal request, instead of giving up or waiting passively.

    20k GitHub starsUsed in 2 repos~6.9k tokens
    Agent WorkflowsAuto-check passed
  • LoopX Self Repair

    loopx-project/loopx

    Diagnoses surprising LoopX behavior, such as stale recommendations or tiny progress, assigns it to the responsible layer and repairs it at the lowest durable level.

    6.2k GitHub stars~2.2k tokensUpdated today
    Agent WorkflowsAuto-check passed

More from yangyuan-zhen/PolyWeather

All 26 skills in this repo
  • AI Slop Cleaner

    yangyuan-zhen/PolyWeather

    [OMX] Run an anti-slop cleanup/refactor/deslop workflow. An agent skill from yangyuan-zhen/PolyWeather.

    316 GitHub stars~2.2k tokensUpdated 20 days ago
    Auto-check passed
  • Analyze

    yangyuan-zhen/PolyWeather

    [OMX] Run read-only deep repository analysis and return a ranked synthesis with explicit confidence, concrete file references, and clear evidence-vs-inference boundaries.

    316 GitHub stars~1.6k tokensUpdated 20 days ago
    Auto-check passed
  • Best Practice Research

    yangyuan-zhen/PolyWeather

    [OMX] Bounded best-practice research wrapper using official/upstream evidence first

    316 GitHub stars~1.4k tokensUpdated 20 days ago
    Auto-check passed
  • Cancel

    yangyuan-zhen/PolyWeather

    [OMX] Cancel any active OMX mode (autopilot, ralph, ultrawork, ecomode, ultraqa, swarm, ultrapilot, pipeline, team)

    316 GitHub stars~3.7k tokensUpdated 20 days ago
    Auto-check passed
  • Configure Notifications

    yangyuan-zhen/PolyWeather

    [OMX] Configure OMX notifications - unified entry point for all platforms

    316 GitHub stars~2.7k tokensUpdated 20 days ago
    Auto-check passed
  • Design

    yangyuan-zhen/PolyWeather

    [OMX] Canonical repo-local DESIGN.md workflow for product, UI/UX, and frontend decision source of truth

    316 GitHub stars~1.8k tokensUpdated 20 days ago
    Auto-check passed

Categories

Questions about Autoresearch

What does Autoresearch do?

[OMX] Stateful validator-gated research loop with native-hook persistence. Autoresearch is an agent skill from yangyuan-zhen/PolyWeather.

When should I use Autoresearch?

Autoresearch fits situations like: tasks that involve Autonomous loops.

How do I install Autoresearch in Claude Code?

Run `npx skills add yangyuan-zhen/PolyWeather --skill autoresearch -a claude-code`. Or copy the skill folder (.codex/skills/autoresearch in yangyuan-zhen/PolyWeather) into .claude/skills/autoresearch in your project. Claude Code loads it when a task matches its description.

How do I install Autoresearch in Codex?

Run `npx skills add yangyuan-zhen/PolyWeather --skill autoresearch -a codex`. Or copy the skill folder (.codex/skills/autoresearch in yangyuan-zhen/PolyWeather) into .agents/skills/autoresearch in your project. Codex loads it when a task matches its description.

Can I use Autoresearch in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add yangyuan-zhen/PolyWeather --skill autoresearch -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/autoresearch, .gemini/skills/autoresearch, .github/skills/autoresearch and .opencode/skills/autoresearch in your project.

What does Autoresearch need to run?

SKILL.md names no scripts, command-line tools or credentials: Autoresearch is instructions for the agent only.

Does Autoresearch access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Autoresearch safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Autoresearch use?

Autoresearch is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Autoresearch use?

About 786 tokens (SKILL.md is roughly 3.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Autoresearch?

Skills that share tags, products or a category with Autoresearch: Show Me Your Work Decision Log (cursor/plugins, 11k stars), Autoresearch Iteration Loop (uditgoenka/autoresearch, 6.5k stars), Install Loop Engineering (cobusgreyling/loop-engineering, 11k stars) and Loopy (Forward-Future/loopy, 3.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Autoresearch?

yangyuan-zhen (a GitHub user) maintains it in yangyuan-zhen/PolyWeather, which has 316 GitHub stars. The repository holds 26 skills in this directory. The repository was last updated on September 20, 2026.

Source: yangyuan-zhen/PolyWeather on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.