BrowserStack Live Testing
handsontable/handsontable
Opens a live BrowserStack session on a real Android, iOS or desktop browser for a local or public URL, tunneling localhost through Cloudflare when needed.
Drive an already-reserved Kobiton device from a natural-language intent.
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install jeremylongshore/tons-of-skills-marketplace drive-automation-session --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/testing/kobiton-automate/skills/drive-automation-session .claude/skills/drive-automation-session && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "drive-automation-session" agent skill from https://github.com/jeremylongshore/tons-of-skills-marketplace/tree/main/plugins/testing/kobiton-automate/skills/drive-automation-session into .claude/skills/drive-automation-session/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "drive-automation-session", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/jeremylongshore/tons-of-skills-marketplace/tree/main/plugins/testing/kobiton-automate/skills/drive-automation-sessionType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install jeremylongshore/tons-of-skills-marketplace drive-automation-session --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .agents/skills && cp -r skills-src/plugins/testing/kobiton-automate/skills/drive-automation-session .agents/skills/drive-automation-session && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "drive-automation-session" agent skill from https://github.com/jeremylongshore/tons-of-skills-marketplace/tree/main/plugins/testing/kobiton-automate/skills/drive-automation-session into .agents/skills/drive-automation-session/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "drive-automation-session", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install jeremylongshore/tons-of-skills-marketplace drive-automation-session --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/plugins/testing/kobiton-automate/skills/drive-automation-session .cursor/skills/drive-automation-session && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "drive-automation-session" agent skill from https://github.com/jeremylongshore/tons-of-skills-marketplace/tree/main/plugins/testing/kobiton-automate/skills/drive-automation-session into .cursor/skills/drive-automation-session/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "drive-automation-session", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/jeremylongshore/tons-of-skills-marketplace.git --path plugins/testing/kobiton-automate/skills/drive-automation-session--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install jeremylongshore/tons-of-skills-marketplace drive-automation-session --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/plugins/testing/kobiton-automate/skills/drive-automation-session .gemini/skills/drive-automation-session && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "drive-automation-session" agent skill from https://github.com/jeremylongshore/tons-of-skills-marketplace/tree/main/plugins/testing/kobiton-automate/skills/drive-automation-session into .gemini/skills/drive-automation-session/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "drive-automation-session", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install jeremylongshore/tons-of-skills-marketplace drive-automation-sessionInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .github/skills && cp -r skills-src/plugins/testing/kobiton-automate/skills/drive-automation-session .github/skills/drive-automation-session && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "drive-automation-session" agent skill from https://github.com/jeremylongshore/tons-of-skills-marketplace/tree/main/plugins/testing/kobiton-automate/skills/drive-automation-session into .github/skills/drive-automation-session/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "drive-automation-session", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install jeremylongshore/tons-of-skills-marketplace drive-automation-session --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/plugins/testing/kobiton-automate/skills/drive-automation-session .opencode/skills/drive-automation-session && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "drive-automation-session" agent skill from https://github.com/jeremylongshore/tons-of-skills-marketplace/tree/main/plugins/testing/kobiton-automate/skills/drive-automation-session into .opencode/skills/drive-automation-session/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "drive-automation-session", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
drive-automation-sessionDrive an already-reserved Kobiton device from a natural-language intent.
Drive Automation Session is an agent skill from jeremylongshore/tons-of-skills-marketplace. Drive an already-reserved Kobiton device from a natural-language intent. Opens an automation Appium session directly against the Kobiton WebDriver hub, runs an observe-decide-act loop with one action per iteration, pauses to ask the user when stuck (same-action repetition, screen unchanged, or model self-declared blocker), and returns the session id. Use when the user says "drive the device to X", describes a flow they want exercised on a reserved device, or asks to "automate this intent on Kobiton". Complements…
Its SKILL.md is about 5.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 11 other files, including scripts and reference files (for example `references/capabilities.md`, `references/endpoint-reference.md` and `references/loop-discipline.md`). Compatibility notes: Cross-platform — talks WebDriver HTTP directly via Node's built-in node:https; no platform-specific binary dependency. Requires Node.js = 18, a persistent…
It sits in Mobile, covering Mobile testing and debugging. It works with Model Context Protocol. The repository describes itself as: Model-agnostic agent-skills platform with a harness-free canonical layer, verified adapters, and the ccpi package manager. Explore at tonsofskills.com. The licence is MIT.
Read from SKILL.md and the folder at commit cfae287. It shows what the files ask for, not the result of running them.
Pre-approves these tools, so the agent can use them without asking each time:
ReadEditBash(node:*)Bash(bash:*)Bash(pwsh:*)Bash(mkdir:*)Bash(mv:*)Bash(date:*)Bash(echo:*)Bash(printf:*)…and 3 more on the same allowed-tools line.
From allowed-tools in the SKILL.md frontmatter.
Ships 5 files in scripts/ (JavaScript), which the agent can run.
Shell commands in SKILL.md call:
nodebashjqclaudepwshFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
portal.kobiton.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Cross-platform — talks WebDriver HTTP directly via Node's built-in node:https; no platform-specific binary dependency. Requires Node.js >= 18, a persistent local filesystem (the observe-decide-act loop passes state through local turn files), and ~/.kobiton/.credentials written by /automate:setup — scripts/appium.js reads that file directly and never calls the MCP getCredential tool, so it is required, not a fallback. An authenticated Kobiton MCP connection is additionally useful for device selection.
From compatibility in the SKILL.md frontmatter.
Drive Automation Session loads about 5.9k tokens when it runs, and up to ~17k if it reads all its reference files. Until then it costs about 184 tokens; SKILL.md has 2,344 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from jeremylongshore/tons-of-skills-marketplace at commit cfae287, republished under its MIT licence (© jeremylongshore). 2,344 words, ~5,942 tokens.
.claude/skills/drive-automation-session/SKILL.md (or your agent's skills folder). This skill also uses 8 other files; get the full folder from GitHub.Drive a Kobiton device from a natural-language intent. The caller provides:
kobiton-store:vXXXXX) or a browser name for web sessions.The skill opens an automation-type Appium session (with appium:newCommandTimeout: 1800 so the session survives pauses while the model asks for guidance), runs an observe-decide-act loop, and returns the session id. The returned session is consumable by the Kobiton MCP tools getSession, getSessionArtifacts, and saveTestCase exactly like any other automation session.
Tool naming. This doc refers to Kobiton MCP tools by their bare names (
getSession,reserveDevice,terminateSession, …). The exact registered name depends on how the host loaded the MCP server (e.g. Claude Code as a plugin exposesmcp__plugin_automate_kobiton__getSession; a repo-local.mcp.jsonorclaude mcp addexposesmcp__kobiton__getSession; other hosts differ). Models fuzzy-match on the bare name, so use it and let the host resolve the prefix.
This skill complements — and does NOT replace — run-interactive-session. They serve different session types:
| Skill | Session type | Best for |
|---|---|---|
run-interactive-session | CLI session (~/.kobiton/bin/kobiton) | Human-driven exploration, ad-hoc inspection |
drive-automation-session | Automation session (direct Appium HTTP) | AI-driven flows, saveable as a test case |
Runs on any OS, but needs two things from the host: a persistent local filesystem — the observe-decide-act loop passes state between turns through iter-N.*.json files, so that file handoff is the loop — and ~/.kobiton/.credentials, which scripts/appium.js reads directly (see below). A host with a filesystem but no way to run /automate:setup still can't finish; route those users to create-test-run. Node.js 18+; no native binary (scripts/appium.js uses node:https). See the Skill compatibility matrix in CLAUDE.md.
scripts/appium.js reads ~/.kobiton/.credentials (written by /automate:setup) directly on every invocation and stops with no-credentials if it's missing — this file is required, not a fallback, and the script never calls the MCP getCredential tool. Reading it directly is deliberate: credentials never pass through argv, env, or the host transcript.kobiton-store:vXXXXX, an .apk / .ipa to upload, or a browser name (for web testing). The skill helps locate or upload it in Step 0.Before reserving anything, the skill clarifies what the user wants. Do NOT auto-pick a device unless the user's intent unambiguously implies a platform AND the user did not constrain the choice.
If the user didn't specify a device in their intent:
listDevices with the relevant platform filter and present the top 3-5 options with name + OS version + availability.getDeviceStatus before reserving.reserveDevice. Capture the UDID, deviceName, platformVersion, automationName for Step 2's caps render.If the user's intent strongly implies a platform (e.g., "test the iOS App Store flow", "open Chrome on Android"), it's OK to auto-pick — but state which device you're picking and why in one line before reserving, so the user can redirect.
For an app testing session:
@build.apk, "the latest staging build", kobiton-store:vXXXXX), use it. Verify the store reference exists via listApps if uncertain..apk/.ipa, an existing kobiton-store:vXXXXX reference, or none if the intent is browser-only)."uploadAppToStore → confirmAppUpload → poll getAppParsingStatus until the state is terminal (OK = ready; stop on a FAILURE_* value). confirmAppUpload only kicks off async parsing on the backend — it does NOT guarantee the app is ready, so the session can fail to install if you skip the poll.For a web testing session, skip the app step — --testingType web --browserName chrome|safari is enough.
If the user didn't say how they want to observe the session, ask before creating it:
"Open the device live view to watch the agent drive (foreground), or run the session in the background?"
Two options:
<portal>/devices/launch?id=<deviceId>&view=device-only (same shape as run-automation-suite Step 5; the <deviceId> is the one captured during Step 0 reservation) and launch it via the chromeless launcher chain (see Step 3b for the chromeless-vs-default-browser branching). The URL is also surfaced as text so the user can click manually if the launcher can't open Chrome and no default-browser fallback applies.Remember the choice; act on it in Step 3 right after the session is created. Do NOT auto-open without asking — some users prefer silent runs, especially on shared screens.
| Input | Required | Notes |
|---|---|---|
intent | yes | One to a few sentences describing what the user wants done on the device |
udid | yes | The device the caller has reserved |
platformName | yes | Android or iOS (look up via getDeviceStatus if unknown) |
platformVersion | yes | from getDeviceStatus |
deviceName | yes | from getDeviceStatus |
automationName | recommended | UiAutomator2 (Android) or XCUITest (iOS) |
app OR browserName | yes | kobiton-store:vXXXXX for app testing, browser name for web testing |
testingType | no | app (default) or web |
MAX_ITERS | no | env var override; default 100. Hard ceiling on iteration count — pure safety net against runaway cycles, not a stuck-detection mechanism. |
appium.js reads ~/.kobiton/.credentials directly on every invocation. If the file is missing, tell the user to run /automate:setup and stop:
if [ ! -s ~/.kobiton/.credentials ]; then
echo "Credentials not found. Run /automate:setup first." >&2
exit 1
fiReuse run-automation-suite's renderer. The --newCommandTimeout 1800 flag is the key piece for this skill.
SKILL_DIR=$(dirname "$0") # this skill's directory
RENDER=$(realpath "$SKILL_DIR/../run-automation-suite/scripts/render-capabilities.js")
# Unique temp file — two concurrent sessions (two terminals / conversations)
# must not clobber each other's rendered caps before Step 3 reads them.
CAPS_TMP=$(mktemp "${TMPDIR:-/tmp}/drive-automation-session-caps.XXXXXX.json")
node "$RENDER" \
--platformName "$PLATFORM_NAME" \
--udid "$UDID" \
--deviceName "$DEVICE_NAME" \
--platformVersion "$PLATFORM_VERSION" \
--automationName "$AUTOMATION_NAME" \
--app "$APP" \
--testingType "${TESTING_TYPE:-app}" \
--newCommandTimeout 1800 \
--scriptlessCapture \
> "$CAPS_TMP"appium.js reads ~/.kobiton/.credentials on each invocation. No flags, no env vars.
SESSION_ID=$(node "$SKILL_DIR/scripts/appium.js" \
--method POST --url /session \
--req-body "@$CAPS_TMP" \
| jq -r '.value.sessionId // .sessionId')
SESSION_DIR=".kobiton/sessions/$SESSION_ID"
mkdir -p "$SESSION_DIR"
mv "$CAPS_TMP" "$SESSION_DIR/caps.json"
printf '%s session=%s started intent=%q\n' "$(date -u +%FT%TZ)" "$SESSION_ID" "$INTENT" > "$SESSION_DIR/session.log"
trap 'cleanup' EXIT INT TERM
cleanup() {
[ -z "$SESSION_ID" ] && return 0
# DELETE /wd/hub/session/<id> ends the WebDriver session cleanly; Kobiton
# records state=COMPLETE. appium.js already treats 404 as idempotent success
# (the session was already gone). appium.js always exits 0 by policy, so we
# check stderr instead of $? to detect a real failure.
delete_err=$(node "$SKILL_DIR/scripts/appium.js" \
--method DELETE --url "/session/$SESSION_ID" 2>&1 >/dev/null)
if [ -z "$delete_err" ]; then
printf '%s session=%s end-via-trap status=COMPLETE\n' "$(date -u +%FT%TZ)" "$SESSION_ID" >> "$SESSION_DIR/session.log"
else
printf '%s session=%s end-via-trap delete-failed (session will time out per appium:newCommandTimeout) err=%s\n' "$(date -u +%FT%TZ)" "$SESSION_ID" "$delete_err" >> "$SESSION_DIR/session.log"
fi
}Build the device-only URL from the device id captured during Step 0 reservation and the portal base URL (e.g., https://portal.kobiton.com) — the AI host has the portal value from earlier MCP responses in the conversation context. Same URL shape as run-automation-suite Step 5.
LIVE_VIEW_URL="<portal>/devices/launch?id=<deviceId>&view=device-only"
printf 'Live view: %s\n' "$LIVE_VIEW_URL"(The post-run replay URL is <portal>/sessions/<kobiton-session-id>/explorer — a different page, NOT the live view. Surface that one only after the session completes if the user wants to review the recording.)
If the user chose background in Step 0, stop here — the URL is logged and the loop proceeds. They can paste it into a browser later if they get curious.
If the user chose foreground, reuse run-automation-suite's launcher chain instead of re-implementing it. The pattern (chromeless Chrome for users who haven't expressed a preference or prefer Chrome; default-browser fallback otherwise) is identical to run-automation-suite Step 5.
Launch via the chromeless helper (default — gated on saved browser preference). Check auto memory for a saved browser preference:
Pick width and height from the reserved device's form factor (the device name from Step 0):
| Device class | Detect by name (case-insensitive) | Portrait W × H | Landscape W × H |
|---|---|---|---|
| Tablet | iPad, Galaxy Tab, Pixel Tablet, Surface, MatePad, or any model with Tab / Pad | 780 × 920 | 920 × 780 |
| Fold (unfolded) | Fold, Z Fold, Pixel Fold | 880 × 920 | 920 × 880 |
| Phone (default) | none of the above | 540 × 920 | 920 × 540 |
Swap width and height if getSession / rendered caps report orientation=LANDSCAPE — default is portrait.
Invoke per host OS. Run in the background (Claude Code: Bash tool with run_in_background: true; other hosts: append & and disown) so the resize-polling loop doesn't block iteration 1. The launcher's stdout/stderr will surface later when it completes; the URL printed above is the user's fallback if the launcher silently fails.
| OS | Command |
|---|---|
| macOS | bash $SKILL_DIR/../run-automation-suite/scripts/chromeless-launcher.sh --url "$LIVE_VIEW_URL" --width <W> --height <H> |
| Windows | pwsh $SKILL_DIR/../run-automation-suite/scripts/chromeless-launcher-windows.ps1 -Url "$LIVE_VIEW_URL" -Width <W> -Height <H> |
| Linux | bash $SKILL_DIR/../run-automation-suite/scripts/chromeless-launcher.sh --url "$LIVE_VIEW_URL" --width <W> --height <H> (launch-only on Linux — no auto-resize) |
Launcher exit codes (surface in the background-task completion event):
0 — Chrome was launched (resize may or may not have succeeded; the warning is informational).2 — Chrome / Chromium is not installed. The user already has $LIVE_VIEW_URL from the printf above; they can paste it into their default browser, or you can offer the default-browser fallback table below as a follow-up.64 — usage error in the launcher invocation. Surface to the user.Default-browser fallback (when the saved preference is Safari/Firefox/Default, or chromeless exited 2). Ask the user once which browser to use (save to auto memory):
| Choice | Command |
|---|---|
| Google Chrome | open -na "Google Chrome" --args --new-window "$LIVE_VIEW_URL" |
| Safari | open -a "Safari" "$LIVE_VIEW_URL" |
| Firefox | open -a "Firefox" "$LIVE_VIEW_URL" |
| Default browser | open "$LIVE_VIEW_URL" |
On Linux, fall back to xdg-open "$LIVE_VIEW_URL" (browser selection isn't supported).
If neither path is available, the URL is already on stdout from the printf above — the user can copy-paste it.
There is no Bash while. The skill is turn-based: each turn, you (the AI host) increment ITER and pick exactly one of three branches:
| Branch | Command | When to pick |
|---|---|---|
| screen | node appium.js screen --session-id $SID --session-dir $DIR | The screen state has likely changed (after a successful act, or at session start, or to verify mid-flow). |
| act | node appium.js <argv> --session-dir $DIR | You know what to do based on the most recent iter-K.xml you observed. |
| control | node appium.js control --done|--blocked --reason "..." --session-dir $DIR | Goal reached, or you're genuinely stuck. Ends the cycle. |
Each branch consumes one ITER. The script always exits 0; the host detects failures by checking whether iter-N.error.json was written.
export ITER=$((ITER + 1))
# Pick ONE of the three. Credentials inherit from env (Step 1).
# Branch A: observe — captures BOTH iter-N.xml and iter-N.png by default.
# Native overlays / OS dialogs only show in the PNG (the webview source
# XML doesn't see them). --xml-only skips the screenshot to save tokens
# when you trust the XML is complete; --png-only skips the source when
# you're verifying layout / image rendering only.
node "$SKILL_DIR/scripts/appium.js" screen \
--session-id "$SESSION_ID" --session-dir "$SESSION_DIR"
# Stdout: {"hash":"<sha256>","xmlBytes":N,"pngBytes":M}
# Track the hash across turns in your conversation context — repetition is a
# signal, never a forced stop. See references/loop-discipline.md "Stuck patterns".
# Branch B: act — execute an Appium call. Writes iter-N.request.json + either
# iter-N.response.json (success) or iter-N.error.json (any failure). The
# error file contains the raw Appium body or {error, message} for usage
# errors; the host reads it to decide what to do next turn.
node "$SKILL_DIR/scripts/appium.js" <YOUR-ARGV> --session-dir "$SESSION_DIR"
# Examples of <YOUR-ARGV>:
# --method POST --url /session/$SESSION_ID/element --req-body '{"using":"xpath","value":"//Button"}'
# actions --session-id $SESSION_ID --type swipe --from-x 540 --from-y 1800 --to-x 540 --to-y 600
# touch-perform --session-id $SESSION_ID --steps @/tmp/steps.json
# --method POST --url /session/$SESSION_ID/element/$EL_ID/clear --req-body '{}'
# Branch C: end the cycle.
node "$SKILL_DIR/scripts/appium.js" control \
--done --reason "Settings reached and Bluetooth toggled" \
--session-dir "$SESSION_DIR"
# OR --blocked --reason "Two modals stacked; cannot dismiss"
# ---- Per-turn bookkeeping (runs after the chosen branch) ----
# Log failures so session.log has a timeline. appium.js writes iter-N.error.json
# on any failure — Appium HTTP error, network blip, usage error like missing
# flag or malformed JSON. The host reads the file next turn and re-plans.
ITER_PAD=$(printf '%03d' "$ITER")
if [ -f "$SESSION_DIR/iter-$ITER_PAD.error.json" ]; then
printf 'iter=%d error\n' "$ITER" >> "$SESSION_DIR/session.log"
fi
# Iteration ceiling — safety net only. Not a stuck-detection mechanism.
if [ "$ITER" -ge "${MAX_ITERS:-100}" ]; then
printf 'iter=%d max-iters-reached limit=%d\n' "$ITER" "${MAX_ITERS:-100}" >> "$SESSION_DIR/session.log"
exit 0
fiWhat to pick next turn depends on what just happened:
screen (the screen probably changed).screen just ran → next turn act (you have a fresh observation).act returned no such element / invalid selector / bad-input → next turn act again with a corrected call. The screen didn't change (the action didn't fire), so the most recent iter-K.xml is still valid; re-read it from disk if you need to, but don't burn an ITER on a fresh screen.act returned stale element reference → next turn screen (the element id is from a prior state; you need a fresh observation to get current element ids).act returned HTTP 5xx / network timeout → next turn act again with the same call (retry). If it fails twice, control --blocked.control --done.loop-discipline.md "Stuck patterns") → control --blocked.The host is responsible for tracking which iter-K.xml represents the current screen state. After a successful act, the previous iter-K.xml is stale until the next screen. After a failed act, the previous iter-K.xml is still current.
When the chosen branch is act, the argv body is constructed from the most recent iter-K.xml — selectors come from attributes you can see in the XML, not from guesses. See references/endpoint-reference.md "Building Appium calls from the observed XML" for the find-element → element-id workflow, selector-strategy preference order (accessibility id > id > relative xpath > css selector > class name), and the coordinates-from-bounds fallback for gestures with no element target.
The full loop discipline (artifact paths, error feedback, termination conditions) lives in references/loop-discipline.md. The endpoint catalog (allowlisted vs not, helpers vs generic) lives in references/endpoint-reference.md.
The Bash trap (set up in Step 3) issues DELETE /wd/hub/session/$SESSION_ID on exit (loop end, error, user stop). That's the clean way to end the session — Kobiton records the session state as COMPLETE, which is what saveTestCase and analytics expect.
Do NOT call the terminateSession MCP tool here. It ends the session via Kobiton's platform-side kill path, which marks the session state as TERMINATED — distinct from COMPLETE and treated as an abnormal exit by the recording pipeline. Reserve terminateSession for the genuinely-stuck case where the WebDriver DELETE is unreachable (network down, Appium hub unresponsive) AND the user explicitly asks to force-kill the session.
If the DELETE somehow fails silently, the session will time out on the Kobiton side per appium:newCommandTimeout (30 min, set in Step 2). Better to wait for that timeout than to mark the session TERMINATED.
Emit the session id back to the caller. The session is now consumable by:
getSession({sessionId}) — session metadata, device info, results.getSessionArtifacts({sessionId}) — video, logs, screenshots, reports.saveTestCase({sessionId}) — persist as a re-runnable test case (sibling feature, out of scope here).See references/endpoint-reference.md for the full Appium endpoint catalog — allowlisted (capturable by saveTestCase) vs not, helpers (actions, touch-perform) vs generic mode. The AI host picks endpoints from that table when emitting iter-N.call.json.
The script does not enforce blocker thresholds. The AI host tracks the hash from screen across turns in its conversation context and decides when to emit control --blocked. See references/loop-discipline.md "Stuck patterns" for concrete examples — same-call repetition, screen oscillation (A→B→A→B), credentials prompt, CAPTCHA, lazy load (use a no-op observe to wait), network spinner.
The only hard programmatic stop is MAX_ITERS=100 (override per session), which is a safety net against runaway cycles — not a stuck-detection mechanism.
| Symptom | Likely cause | Fix |
|---|---|---|
No credentials available... | No ~/.kobiton/.credentials (an authenticated MCP connection does not substitute — appium.js only reads the file) | Run /automate:setup, which uses the MCP connection to fetch and write the file |
iter-N.error.json has status: 401 on session create | Credentials stale | Re-run /automate:setup to refresh ~/.kobiton/.credentials; check the portal URL |
iter-N.error.json has a platform-cap message on session create | newCommandTimeout: 1800 rejected by Kobiton | Lower the timeout in references/capabilities.md |
iter-N.error.json body has value.error: "no such element" | Selector matched nothing | Re-plan next turn with a different strategy / selector |
iter-N.error.json body has value.error: "invalid session id" | Kobiton platform-side session ended | Emit control --blocked (or just let MAX_ITERS catch it); trap cleans up |
kobiton:aiToolName capability is auto-detected by render-capabilities.js (Claude / Codex / Copilot / Gemini). No flag needed..kobiton/sessions/<session-id>/ — workspace-relative, not /tmp (consistent with run-interactive-session).tools/*.yaml and does not add any MCP tool. It only consumes existing MCP tools (reserveDevice, getSession, terminateSession, etc.).© jeremylongshore, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 8 other files (scripts, references) in plugins/testing/kobiton-automate/skills/drive-automation-session of jeremylongshore/tons-of-skills-marketplace.
Open the folder on GitHubat commit cfae287
Drive Automation Session next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Drive Automation Session this skilljeremylongshore/tons-of-skills-marketplace | 2.8k | — | ~5.9k | Automated safety check: Pass | MIT | |
| BrowserStack Live Testinghandsontable/handsontable | 22k | — | ~950 | Automated safety check: Pass | Custom licence | |
| Magicnet Device DebuggingLIghtJUNction/MagicNet | 186 | — | ~2.9k | Automated safety check: Notes | MIT | |
| Scrcpy GotchasJuanCF/scrcpy-mcp | 116 | — | ~5.8k | Automated safety check: Pass | MIT | |
| iOS Accessibility Testingconorluddy/xclaude-plugin | 183 | — | ~5k | Automated safety check: Pass | MIT | |
| Generate Appclaw Flowappclawhq/AppClaw | 117 | — | ~3.4k | Automated safety check: Notes | Apache-2.0 |
handsontable/handsontable
Opens a live BrowserStack session on a real Android, iOS or desktop browser for a local or public URL, tunneling localhost through Cloudflare when needed.
LIghtJUNction/MagicNet
Debug and verify MagicNet on a real Android root device. An agent skill from LIghtJUNction/MagicNet.
JuanCF/scrcpy-mcp
A skill your agent uses when interacting with Android devices via scrcpy-mcp tools (tap, swipe, inputtext, keyevent, uidump, uifindelement, appstart, etc.).
conorluddy/xclaude-plugin
Guides WCAG 2.1 and VoiceOver accessibility testing for iOS apps, working from the accessibility tree and not from screenshots.
appclawhq/AppClaw
Generate YAML flow files for AppClaw mobile automation. An agent skill from appclawhq/AppClaw.
WahdanZ/SpockAdb
Debug Android apps on a connected device or emulator through the Spock ADB MCP server (tools named android).
jeremylongshore/tons-of-skills-marketplace
Execute this skill enables AI assistant to conduct a security-focused code review using the security-agent plugin.
jeremylongshore/tons-of-skills-marketplace
Build this skill automates the adaptation of pre-trained machine learning models using transfer learning techniques.
jeremylongshore/tons-of-skills-marketplace
Execute proactive auto-loading: automatically detects and loads agents.md files.
jeremylongshore/tons-of-skills-marketplace
Aggregate and centralize performance metrics from applications, systems, databases, caches, and services.
jeremylongshore/tons-of-skills-marketplace
Execute this skill enables AI assistant to analyze capacity requirements and plan for future growth.
jeremylongshore/tons-of-skills-marketplace
Process use when you need to work with database indexing. An agent skill from jeremylongshore/tons-of-skills-marketplace.
Works with
Categories
Drive an already-reserved Kobiton device from a natural-language intent. Drive Automation Session is an agent skill from jeremylongshore/tons-of-skills-marketplace. Drive an already-reserved Kobiton device from a natural-language intent.
Drive Automation Session fits situations like: the user says drive the device to X; describes a flow they want exercised on a reserved device; asks to automate this intent on Kobiton.
Run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a claude-code`. Or copy the skill folder (plugins/testing/kobiton-automate/skills/drive-automation-session in jeremylongshore/tons-of-skills-marketplace) into .claude/skills/drive-automation-session in your project. Claude Code loads it when a task matches its description.
Run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a codex`. Or copy the skill folder (plugins/testing/kobiton-automate/skills/drive-automation-session in jeremylongshore/tons-of-skills-marketplace) into .agents/skills/drive-automation-session in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill drive-automation-session -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/drive-automation-session, .gemini/skills/drive-automation-session, .github/skills/drive-automation-session and .opencode/skills/drive-automation-session in your project.
Going by SKILL.md and its folder, Drive Automation Session needs JavaScript for the scripts in its folder and the command-line tools its instructions call (node, bash, jq, claude and pwsh). Our summary lists: Node.js. Its frontmatter pre-approves these tools: Read, Edit, Bash(node:*), Bash(bash:*), Bash(pwsh:*), Bash(mkdir:*), Bash(mv:*), Bash(date:*), Bash(echo:*), Bash(printf:*), Bash(jq:*), Bash(open:*), Bash(xdg-open:*). Compatibility (from SKILL.md): Cross-platform — talks WebDriver HTTP directly via Node's built-in node:https; no platform-specific binary dependency. Requires Node.js >= 18, a persistent local filesystem (the observe-decide-act loop passes state through local turn files), and ~/.kobiton/.credentials written by /automate:setup — scripts/appium.js reads that file directly and never calls the MCP getCredential tool, so it is required, not a fallback. An authenticated Kobiton MCP connection is additionally useful for device selection..
SKILL.md names 1 domain. In commands or code: portal.kobiton.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Drive Automation Session is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 5.9k tokens (SKILL.md is roughly 24k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 11k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Drive Automation Session: BrowserStack Live Testing (handsontable/handsontable, 22k stars), Magicnet Device Debugging (LIghtJUNction/MagicNet, 186 stars), Scrcpy Gotchas (JuanCF/scrcpy-mcp, 116 stars) and iOS Accessibility Testing (conorluddy/xclaude-plugin, 183 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
jeremylongshore (a GitHub user) maintains it in jeremylongshore/tons-of-skills-marketplace, which has 2,827 GitHub stars. The repository holds 3,342 skills in this directory. The repository was last updated on October 10, 2026.
Source: jeremylongshore/tons-of-skills-marketplace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.