Agent skill

Browser Use

by Prismer-AI in Prismer-AI/PrismerCloud

LLM-driven browser automation via the pre-installed browser-use library (chromium already in the image).

MITAuto-check passedProductivity & Automation

Install Browser Use

skills CLI
$ npx skills add Prismer-AI/PrismerCloud --skill browser-use -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Prismer-AI/PrismerCloud browser-use --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Prismer-AI/PrismerCloud.git skills-src && mkdir -p .claude/skills && cp -r skills-src/sdk/cloud/catalog/skills/browser-use .claude/skills/browser-use && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
browser-use
GitHub stars
1.6k
Token cost
~1.1k tokens
SKILL.md length
298 words
Files
1
Skills in repo
88
Repo updated
First seen
Licence
MIT

At a glance

LLM-driven browser automation via the pre-installed browser-use library (chromium already in the image).

  • The task needs to interact with web pages beyond a single fetch — multi-step navigation
  • SKILL.md covers When to use, How to run, Task-writing rules (from… and Environment facts, plus 1 more section
  • Calls python; needs PRISMER_API_KEY
  • Structured data extraction

What it does

Browser Use is an agent skill from Prismer-AI/PrismerCloud. LLM-driven browser automation via the pre-installed browser-use library (chromium already in the image). Use whenever the task needs to interact with web pages beyond a single fetch — multi-step navigation, form filling, clicking, scrolling, structured data extraction, screenshots, or tasks the plain web tools can't complete. Runs as a Python script against the built-in chromium; LLM goes through the Prismer gateway (no external LLM key needed).

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Productivity & Automation, covering Browser automation. It works with Python. The licence is MIT.

When your agent uses it

  • The task needs to interact with web pages beyond a single fetch — multi-step navigation
  • Structured data extraction
  • Tasks the plain web tools cant complete

Example prompts

  • “/browser-use”

Requirements

  • Python 3
  • A credential in PRISMER_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit e5d9444. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • PRISMER_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Browser Use loads about 1.1k tokens when it runs. Until then it costs about 115 tokens; SKILL.md has 298 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~115
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Prismer-AI/PrismerCloud at commit e5d9444, republished under its MIT licence (© Prismer-AI). 298 words, ~1,100 tokens.

Download SKILL.mdSave it as .claude/skills/browser-use/SKILL.md (or your agent's skills folder).
name
browser-use
description
LLM-driven browser automation via the pre-installed browser-use library (chromium already in the image). Use whenever the task needs to interact with web pages beyond a single fetch — multi-step navigation, form filling, clicking, scrolling, structured data extraction, screenshots, or tasks the plain web tools can't complete. Runs as a Python script against the built-in chromium; LLM goes through the Prismer gateway (no external LLM key needed).
scope
common

Browser-Use (Web Automation)

The sandbox image ships browser-use (installed in /home/user/.venv) and a full chromium build (playwright). This skill drives them: write a short Python script with Agent(task=..., llm=..., browser=...), run it, report the result.

When to use

  • Multi-step web tasks: search → open → extract → compare → download
  • Form filling / login flows / clicking through pages
  • Structured extraction (tables, listings, prices) that a single fetch can't parse
  • Screenshots / visual confirmation of a page state
  • Anything needing JS-rendered content (SPA pages the fetch tools miss)

For a single plain fetch, prefer the existing web tools (lighter). Reach for browser-use when they come back empty or the task is interactive.

How to run

bash
# chromium executable (version dir may change — resolve dynamically):
CHROME="$(find /home/user/.cache/ms-playwright -name chrome -type f | head -1)"

# LLM = our gateway (OpenAI-compatible). PRISMER_BASE_URL / PRISMER_API_KEY
# are already in the environment; never hardcode them in the script.
cat > /tmp/bu_task.py <<'PY'
import os, asyncio
from browser_use import Agent, ChatOpenAI, Browser

async def main():
    llm = ChatOpenAI(
        base_url=f"{os.environ['PRISMER_BASE_URL']}/api/v1",
        api_key=os.environ['PRISMER_API_KEY'],
        model=os.environ.get('PRISMER_MODEL', 'deepseek-v4-flash'),
        temperature=0.0,
        # REQUIRED for the Prismer gateway: browser-use defaults to forcing
        # JSON-schema structured output (response_format=json_schema) which
        # our upstream models reject with 400 "response_format type
        # unavailable" (verified 2026-08-07). Disable it; the agent still
        # extracts via its normal flow.
        dont_force_structured_output=True,
    )
    browser = Browser(
        headless=True,
        executable_path=os.environ['CHROME'],
    )
    agent = Agent(
        task="<TASK>",  # be specific: steps, URLs, what to extract, output format
        llm=llm,
        browser=browser,
    )
    history = await agent.run(max_steps=30)
    print("RESULT:", history.final_result())
    print("URLS:", history.urls())
    print("ERRORS:", history.errors())
    await browser.close()

asyncio.run(main())
PY
CHROME="$CHROME" /home/user/.venv/bin/python /tmp/bu_task.py
rm -f /tmp/bu_task.py

Task-writing rules (from browser-use docs)

  • Be specific: "Go to https://…, use extract with query 'first 3 quotes and their authors', save to quotes.csv via write_file" — open-ended tasks fail.
  • Name actions: "use search action for …", "use click to open first result in a new tab", "use send_keys with 'Tab Tab Enter'" — the agent maps these to its built-in tools.
  • Error recovery: if navigation is blocked, fall back to a search engine; if a click fails, use keyboard navigation (send_keys).
  • max_steps: default 100; keep 30 for focused tasks, raise for long flows.
  • Anti-bot: if a page blocks automation, try use_cloud=True — NOT available here (no Browser-Use Cloud key). Fall back to search/cache.

Environment facts

ItemValue
Python/home/user/.venv/bin/python (browser-use installed here)
Chromium$(find /home/user/.cache/ms-playwright -name chrome -type f | head -1)
LLMPrismer gateway (PRISMER_BASE_URL + PRISMER_API_KEY, OpenAI-compatible)
ModelPRISMER_MODEL env (default deepseek-v4-flash) — adjust for the task
Telemetrybrowser-use collects anonymous telemetry by default — set ANONYMIZED_TELEMETRY=false

Output contract

Report to the user: history.final_result() (the extracted answer), the URLs visited, and any errors. If history.is_successful() is False, say so plainly with history.errors() — do not invent a completion. Attach screenshots (history.screenshot_paths()) as task assets when they matter.

© Prismer-AI, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in sdk/cloud/catalog/skills/browser-use of Prismer-AI/PrismerCloud.

Open the folder on GitHubat commit e5d9444

Compare with similar skills

Browser Use next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Browser Use compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Browser Use this skillPrismer-AI/PrismerCloud1.6k—~1.1kAutomated safety check: PassMIT
Browser Use Terminalbrowser-use/terminal652—~2.1kAutomated safety check: PassMIT
Evoui BrowserSalmonbird/evoui-browser100—~4.5kAutomated safety check: PassApache-2.0
Browser Automation Edge Casesaden-hive/hive11k—~1.8kAutomated safety check: PassMIT
Skyvern Browser AutomationSkyvern-AI/skyvern23k—~1.9kAutomated safety check: PassAGPL-3.0
AI Search Hubminsight-ai-info/AI-Search-Hub1.3k—~1.3kAutomated safety check: PassNone

Similar skills

  • Browser Use Terminal

    browser-use/terminal

    Direct browser control via the Browser Use Terminal CLI. An agent skill from browser-use/terminal.

    652 GitHub stars~2.1k tokensUpdated 1 mo ago
    Productivity & AutomationAuto-check passed
  • Evoui Browser

    Salmonbird/evoui-browser

    Use Evoui Browser as the default entry point for common, self-terminating web tasks that need a real browser through agent-browser, including navigation, page reading, clicks, forms, login flows…

    100 GitHub stars~4.5k tokensUpdated 1 mo ago
    Productivity & AutomationAuto-check passed
  • Step-by-step procedure for debugging browser automation failures on complex sites such as LinkedIn, Twitter/X, single-page apps and Shadow DOM pages.

    11k GitHub stars~1.8k tokensUpdated today
    Productivity & AutomationAuto-check passed
  • Skyvern Browser Automation

    Skyvern-AI/skyvern

    Automates websites with Skyvern's AI browser agent to fill forms, extract data, download files, log in and run multi-step workflows through SDKs, REST, MCP or a CLI.

    23k GitHub stars~1.9k tokensUpdated today
    Productivity & AutomationAuto-check passed
  • AI Search Hub

    minsight-ai-info/AI-Search-Hub

    Run the AI Search Hub browser automation scripts for Yuanbao, LongCat, Doubao, Qwen, Gemini, Grok, and MiniMax.

    1.3k GitHub stars~1.3k tokensUpdated 5 mo ago
    Productivity & AutomationAuto-check passed
  • Browser Tools

    981377660LMT/algorithm-study

    Interactive browser automation via Chrome DevTools Protocol.

    277 GitHub starsUsed in 1 repo~1.3k tokens
    Productivity & AutomationAuto-check passed

More from Prismer-AI/PrismerCloud

All 88 skills in this repo
  • Prismer Google Workspace

    Prismer-AI/PrismerCloud

    Gives an agent account-scoped access to Gmail, Calendar, Drive, Contacts, Docs and Sheets through the gws CLI or a bundled Python client.

    1.6k GitHub starsUsed in 3 repos~4.2k tokens
    Auto-check passed
  • Prismer Skill Creator

    Prismer-AI/PrismerCloud

    Walks an agent through creating, importing, editing, validating, testing and publishing Prismer Skills with a fixed workflow and bundled scripts.

    1.6k GitHub stars~2.6k tokensUpdated 9 days ago
    Auto-check: notes
  • Himalaya Email CLI

    Prismer-AI/PrismerCloud

    Operates a mailbox from the terminal with the external Himalaya CLI over IMAP, SMTP, Notmuch or Sendmail, separate from any built-in email gateway adapter.

    1.6k GitHub starsUsed in 2 repos~2.3k tokens
    Auto-check passed
  • Prismer Image Generation

    Prismer-AI/PrismerCloud

    Generates one image from a text prompt with a bundled Node.js helper and delivers it once as the attachment to the current Prismer reply.

    1.6k GitHub stars~1.4k tokensUpdated 9 days ago
    Auto-check passed
  • Manim Explainer Videos

    Prismer-AI/PrismerCloud

    Produces 3Blue1Brown-style explainer animations with Manim Community Edition for math, algorithms, equations and architecture diagrams, with planning and rendering references.

    1.6k GitHub starsUsed in 2 repos~3.1k tokens
    Auto-check passed
  • Prismer Role Builder

    Prismer-AI/PrismerCloud

    Creates or updates Prismer role templates from a persona, SOP or job description, and turns a role into a working agent that runs its first task through a bundled script.

    1.6k GitHub stars~2.3k tokensUpdated 9 days ago
    Auto-check: notes

Works with

Questions about Browser Use

What does Browser Use do?

LLM-driven browser automation via the pre-installed browser-use library (chromium already in the image). Browser Use is an agent skill from Prismer-AI/PrismerCloud. LLM-driven browser automation via the pre-installed browser-use library (chromium already in the image).

When should I use Browser Use?

Browser Use fits situations like: the task needs to interact with web pages beyond a single fetch — multi-step navigation; structured data extraction; tasks the plain web tools cant complete.

How do I install Browser Use in Claude Code?

Run `npx skills add Prismer-AI/PrismerCloud --skill browser-use -a claude-code`. Or copy the skill folder (sdk/cloud/catalog/skills/browser-use in Prismer-AI/PrismerCloud) into .claude/skills/browser-use in your project. Claude Code loads it when a task matches its description.

How do I install Browser Use in Codex?

Run `npx skills add Prismer-AI/PrismerCloud --skill browser-use -a codex`. Or copy the skill folder (sdk/cloud/catalog/skills/browser-use in Prismer-AI/PrismerCloud) into .agents/skills/browser-use in your project. Codex loads it when a task matches its description.

Can I use Browser Use in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Prismer-AI/PrismerCloud --skill browser-use -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/browser-use, .gemini/skills/browser-use, .github/skills/browser-use and .opencode/skills/browser-use in your project.

What does Browser Use need to run?

Going by SKILL.md and its folder, Browser Use needs the command-line tools its instructions call (python) and credentials named PRISMER_API_KEY. Our summary lists: Python 3; A credential in PRISMER_API_KEY.

Does Browser Use access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Browser Use safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Browser Use use?

Browser Use is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Browser Use use?

About 1.1k tokens (SKILL.md is roughly 4.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Browser Use?

Skills that share tags, products or a category with Browser Use: Browser Use Terminal (browser-use/terminal, 652 stars), Evoui Browser (Salmonbird/evoui-browser, 100 stars), Browser Automation Edge Cases (aden-hive/hive, 11k stars) and Skyvern Browser Automation (Skyvern-AI/skyvern, 23k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Browser Use?

Prismer-AI (a GitHub organization) maintains it in Prismer-AI/PrismerCloud, which has 1,555 GitHub stars. The repository holds 88 skills in this directory. The repository was last updated on September 30, 2026.

Source: Prismer-AI/PrismerCloud on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.