Agent skill

Android Benchmark Comparison

by chrisbanes in chrisbanes/skills

A skill your agent uses when comparing physical Android benchmark configurations, investigating inconsistent rankings, or selecting an Android default from measured results.

Apache-2.0Auto-check passedMobile

Install Android Benchmark Comparison

skills CLI
$ npx skills add chrisbanes/skills --skill android-benchmark-comparison -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install chrisbanes/skills android-benchmark-comparison --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/chrisbanes/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/android-benchmark-comparison .claude/skills/android-benchmark-comparison && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
android-benchmark-comparison
GitHub stars
1.1k
Token cost
~971 tokens
SKILL.md length
486 words
Files
1
Skills in repo
19
Repo updated
First seen
Licence
Apache-2.0

At a glance

A skill your agent uses when comparing physical Android benchmark configurations, investigating inconsistent rankings, or selecting an Android default from measured results.

  • Works in 7 steps: State the decision, configurations,… → Control and record relevant device… → Balance or reverse run order and repeat… → …
  • Comparing physical Android benchmark configurations
  • SKILL.md covers Core principle, Procedure and Boundaries
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Android Benchmark Comparison is an agent skill from chrisbanes/skills. Use when comparing physical Android benchmark configurations, investigating inconsistent rankings, or selecting an Android default from measured results. Do not use for code-level Compose performance diagnosis without a configuration comparison.

Its SKILL.md is about 970 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Mobile. It works with Android. The repository describes itself as: Skills for Kotlin, Jetpack Compose, and Android development. The licence is Apache-2.0.

When your agent uses it

  • Comparing physical Android benchmark configurations
  • Investigating inconsistent rankings
  • Selecting an Android default from measured results
  • Code-level Compose performance diagnosis without a configuration comparison

Example prompts

  • “/android-benchmark-comparison”

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. State the decision, configurations, workloads, metric definitions, and
  2. Control and record relevant device conditions, including device model and
  3. Balance or reverse run order and repeat the comparison. Report the spread
  4. When rankings reverse or variability is material, defer a firm default
  5. Calculate summaries from unrounded observations, then round only for
  6. Separate controlled-experiment evidence from normal user performance. If
  7. Finish with the raw-evidence location, completed-case counts, variability,

What it can do on your machine

Read from SKILL.md and the folder at commit f872f97. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Android Benchmark Comparison loads about 971 tokens when it runs. Until then it costs about 69 tokens; SKILL.md has 486 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~69
When it runs · the whole SKILL.md, loaded when a task matches
~971

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from chrisbanes/skills at commit f872f97, republished under its Apache-2.0 licence (© chrisbanes). 486 words, ~971 tokens.

Download SKILL.mdSave it as .claude/skills/android-benchmark-comparison/SKILL.md (or your agent's skills folder).
name
android-benchmark-comparison
description
Use when comparing physical Android benchmark configurations, investigating inconsistent rankings, or selecting an Android default from measured results. Do not use for code-level Compose performance diagnosis without a configuration comparison.

Android benchmark comparison

Core principle

Treat a physical Android configuration comparison as a reproducible experiment: verify comparable workloads and device conditions before interpreting a ranking or choosing a default.

Procedure

  1. State the decision, configurations, workloads, metric definitions, and repetitions. Preserve exact build identity, configuration, raw results, and traces; then verify every intended case and iteration ran. Distinguish missing, failed, and excluded runs; do not compare only the fastest survivors.
  2. Control and record relevant device conditions, including device model and state, thermal and power mode, display brightness, background load, and network or input conditions. Keep device-specific commands and CPU masks in the project's runbook.
  3. Balance or reverse run order and repeat the comparison. Report the spread and whether the ordering holds; do not discard slow iterations after seeing the result.
  4. When rankings reverse or variability is material, defer a firm default decision until the reversal is resolved. Give the complete next comparison, not just its first blocker: confirm the same named cases and iterations, repeat with balanced or reversed run order, and inspect trace data whose timestamps overlap each measured interval for placement, contention, or thermal changes. Fixed-performance mode does not prove CPU placement. If an affinity experiment was attempted, discover the device topology and verify placement during the measured interval. Restore the recorded original affinity settings after the experiment and verify that restoration before another run; do not merely note that restoration needs checking. Label verified affinity runs as controlled comparisons. If the original settings or any other check are unavailable, state the gap and keep any default choice explicitly provisional.
  5. Calculate summaries from unrounded observations, then round only for presentation. Name the aggregation explicitly: the mean of per-run percentiles is not a percentile of pooled observations. Choose an aggregation that answers the stated decision; do not prescribe one statistic universally.
  6. Separate controlled-experiment evidence from normal user performance. If several conditions changed together, report the comparison as more controlled but do not attribute its whole difference to one control. Use CPU frame-duration evidence to inform a visual quality/performance decision, without claiming it measures GPU shader time.
  7. Finish with the raw-evidence location, completed-case counts, variability, trace findings, controls and restoration status, plus the bounded decision or remaining uncertainty. When a reversal is unresolved, state the full sequence still needed: matching coverage, balanced or reversed order, measured-interval trace inspection, and restoration of any changed affinity settings. Name run order explicitly in the recommendation: a "balanced comparison" alone does not tell the team to balance or reverse run order. Label any earlier default choice provisional.
Show full SKILL.md (57 more words)Show less

Boundaries

  • A single stable benchmark run can support a narrow observation, but not a robust configuration ranking.
  • Do not turn a device-specific CPU mask, brightness value, iteration count, or summary statistic into a permanent default.
  • When traces or repeat coverage cannot resolve a reversal, keep the default unchanged or make a provisional decision with that limitation explicit.

© chrisbanes, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/android-benchmark-comparison of chrisbanes/skills.

Open the folder on GitHubat commit f872f97

Compare with similar skills

Android Benchmark Comparison next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Android Benchmark Comparison compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Android Benchmark Comparison this skillchrisbanes/skills1.1k—~971Automated safety check: PassApache-2.0
Compose Multiplatform Patternsmonta-app/ocpp-emulator1805 repos~2kAutomated safety check: PassApache-2.0
Phone HarnessShawnPana/phone-harness3.2k—~4.3kAutomated safety check: PassMIT
Stylesarindamxd/camerax-android1324 repos~2.3kAutomated safety check: PassApache-2.0
Verified Emailarindamxd/camerax-android1324 repos~4.7kAutomated safety check: PassApache-2.0
Argent Metro Debuggerbbplayer-app/BBPlayer1.1k—~3.4kAutomated safety check: PassMIT

Similar skills

  • Compose Multiplatform Patterns

    monta-app/ocpp-emulator

    Compose Multiplatform and Jetpack Compose patterns for KMP projects — state management, navigation, theming, performance, and platform-specific UI.

    180 GitHub starsUsed in 5 repos~2k tokens
    MobileAuto-check passed
  • Phone Harness

    ShawnPana/phone-harness

    Control the user's phone — an iPhone through the Mac's iPhone Mirroring window, an Android over adb, or a rented cloud Android: open apps, tap, type, swipe, read the screen.

    3.2k GitHub stars~4.3k tokensUpdated 10 days ago
    MobileAuto-check passed
  • Styles

    arindamxd/camerax-android

    A skill your agent uses to integrate the Jetpack Compose Styles API into an Android project.

    132 GitHub starsUsed in 4 repos~2.3k tokens
    MobileAuto-check passed
  • Verified Email

    arindamxd/camerax-android

    Provides a complete workflow for implementing verified email retrieval on Android Credential Manager API.

    132 GitHub starsUsed in 4 repos~4.7k tokens
    MobileAuto-check passed
  • Argent Metro Debugger

    bbplayer-app/BBPlayer

    Debug a JS runtime via CDP using argent debugger tools. An agent skill from bbplayer-app/BBPlayer.

    1.1k GitHub stars~3.4k tokensUpdated today
    MobileAuto-check passed
  • 统一 Android XML 资源命名:layout、drawable、mipmap、color、values、id 的前缀与 snakecase 规则;把颜色/圆角/描边/状态编码进文件名(bg、border、textcolor…selector)。

    1.6k GitHub stars~1.3k tokensUpdated 1 mo ago
    MobileAuto-check passed

More from chrisbanes/skills

All 19 skills in this repo
  • Gradle Run

    chrisbanes/skills

    A skill your agent uses when planning to execute Gradle through gradle, ./gradlew, or a custom gradlew wrapper script, or diagnosing a Gradle build, compact workflow ledger, repeated failure…

    1.1k GitHub stars~1.2k tokensUpdated 4 days ago
    Auto-check passed
  • Compose Animations

    chrisbanes/skills

    A skill your agent uses when writing or reviewing Jetpack Compose motion: visibility enter/exit, animating one property toward a target, color or size transitions, multiple properties from one…

    1.1k GitHub stars~1.1k tokensUpdated 4 days ago
    Auto-check passed
  • A skill your agent uses when writing or reviewing Jetpack Compose UI tests, screenshot tests or baseline-recording evidence, previews, semantics assertions, fake image loading, keyboard input, focus…

    1.1k GitHub stars~1.8k tokensUpdated 4 days ago
    Auto-check passed
  • Release Kotlin Library

    chrisbanes/skills

    A skill your agent uses when preparing, publishing, or checking readiness for a Kotlin library release, including verifying its gradle-maven-publish-plugin prerequisite, reconciling changelogs and…

    1.1k GitHub stars~1.9k tokensUpdated 4 days ago
    Auto-check: notes
  • Run GitHub Project

    chrisbanes/skills

    A skill your agent uses when asked to set up, review, or operate a repository's GitHub Project workflow, including ready claims, role-labelled human work, unknown remote mutation outcomes, Todo…

    1.1k GitHub stars~1.5k tokensUpdated 4 days ago
    Auto-check passed
  • Compose Component Design

    chrisbanes/skills

    A skill your agent uses when designing or reviewing reusable Jetpack Compose component APIs with modifier parameters, root layout placement, caller-provided variable content, primitive content…

    1.1k GitHub starsUsed in 1 repo~718 tokens
    Auto-check passed

Works with

Categories

Questions about Android Benchmark Comparison

What does Android Benchmark Comparison do?

A skill your agent uses when comparing physical Android benchmark configurations, investigating inconsistent rankings, or selecting an Android default from measured results. Android Benchmark Comparison is an agent skill from chrisbanes/skills. Use when comparing physical Android benchmark configurations, investigating inconsistent rankings, or selecting an Android default from measured results.

When should I use Android Benchmark Comparison?

Android Benchmark Comparison fits situations like: comparing physical Android benchmark configurations; investigating inconsistent rankings; selecting an Android default from measured results; code-level Compose performance diagnosis without a configuration comparison.

How do I install Android Benchmark Comparison in Claude Code?

Run `npx skills add chrisbanes/skills --skill android-benchmark-comparison -a claude-code`. Or copy the skill folder (skills/android-benchmark-comparison in chrisbanes/skills) into .claude/skills/android-benchmark-comparison in your project. Claude Code loads it when a task matches its description.

How do I install Android Benchmark Comparison in Codex?

Run `npx skills add chrisbanes/skills --skill android-benchmark-comparison -a codex`. Or copy the skill folder (skills/android-benchmark-comparison in chrisbanes/skills) into .agents/skills/android-benchmark-comparison in your project. Codex loads it when a task matches its description.

Can I use Android Benchmark Comparison in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add chrisbanes/skills --skill android-benchmark-comparison -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/android-benchmark-comparison, .gemini/skills/android-benchmark-comparison, .github/skills/android-benchmark-comparison and .opencode/skills/android-benchmark-comparison in your project.

What does Android Benchmark Comparison need to run?

SKILL.md names no scripts, command-line tools or credentials: Android Benchmark Comparison is instructions for the agent only.

Does Android Benchmark Comparison access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Android Benchmark Comparison safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Android Benchmark Comparison use?

Android Benchmark Comparison is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Android Benchmark Comparison use?

About 971 tokens (SKILL.md is roughly 3.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Android Benchmark Comparison?

Skills that share tags, products or a category with Android Benchmark Comparison: Compose Multiplatform Patterns (monta-app/ocpp-emulator, 180 stars), Phone Harness (ShawnPana/phone-harness, 3.2k stars), Styles (arindamxd/camerax-android, 132 stars) and Verified Email (arindamxd/camerax-android, 132 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Android Benchmark Comparison?

chrisbanes (a GitHub user) maintains it in chrisbanes/skills, which has 1,092 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on October 4, 2026.

Source: chrisbanes/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.