Agent skill

Brightdata Core Workflow B

by jeremylongshore in jeremylongshore/tons-of-skills-marketplace

Analyze and operate the Bright Data Web Scraper API async snapshot lifecycle with terminal-state and delivery controls.

MITAuto-check passedData & Analytics

Install Brightdata Core Workflow B

skills CLI
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill brightdata-core-workflow-b -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install jeremylongshore/tons-of-skills-marketplace brightdata-core-workflow-b --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/.curated/brightdata-core-workflow-b .claude/skills/brightdata-core-workflow-b && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
brightdata-core-workflow-b
GitHub stars
2.8k
Token cost
~1k tokens
SKILL.md length
353 words
Files
2 (incl. references)
Skills in repo
3,342
Repo updated
First seen
Licence
MIT

At a glance

Analyze and operate the Bright Data Web Scraper API async snapshot lifecycle with terminal-state and delivery controls.

  • Works in 4 steps: Validate the manifest → Trigger once → Poll deliberately → …
  • Collecting an approved batch
  • SKILL.md covers Overview, Prerequisites, Instructions and Tool Discipline, plus 4 more sections
  • Calls curl; reaches api.brightdata.com; needs BRIGHTDATA_API_KEY

What it does

Brightdata Core Workflow B is an agent skill from jeremylongshore/tons-of-skills-marketplace. Analyze and operate the Bright Data Web Scraper API async snapshot lifecycle with terminal-state and delivery controls. Use when collecting an approved batch, polling a snapshot, or downloading large structured results. Trigger with: "run a Bright Data dataset job", "poll a Bright Data snapshot", "download Web Scraper API results".

Its SKILL.md is about 1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/official-docs.md`). Compatibility notes: Requires an approved Bright Data account or offline fixtures, current Bright Data documentation, and an authorized public-data collection purpose

It sits in Data & Analytics, covering Web scraping. It works with Bright Data. The repository describes itself as: Model-agnostic agent-skills platform with a harness-free canonical layer, verified adapters, and the ccpi package manager. Explore at tonsofskills.com. The licence is MIT.

When your agent uses it

  • Collecting an approved batch
  • Polling a snapshot
  • Downloading large structured results
  • With: run a Bright Data dataset job

Example prompts

  • “run a Bright Data dataset job”
  • “poll a Bright Data snapshot”
  • “download Web Scraper API results”
  • “/brightdata-core-workflow-b”

Requirements

  • A credential in BRIGHTDATA_API_KEY
  • Compatibility (from SKILL.md): Requires an approved Bright Data account or offline fixtures, current Bright Data documentation, and an authorized public-data collection purpose
  • Pre-approved tools (allowed-tools): Read, Grep, Write, Edit, Bash(curl:*)

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. Validate the manifest
  2. Trigger once
  3. Poll deliberately
  4. Download and verify

What it can do on your machine

Read from SKILL.md and the folder at commit cfae287. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Grep
    • Write
    • Edit
    • Bash(curl:*)

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.brightdata.com

    Also links to:

    • docs.brightdata.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • BRIGHTDATA_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

  • Compatibility

    Requires an approved Bright Data account or offline fixtures, current Bright Data documentation, and an authorized public-data collection purpose

    From compatibility in the SKILL.md frontmatter.

Context cost

Brightdata Core Workflow B loads about 1k tokens when it runs, and up to ~1.2k if it reads all its reference files. Until then it costs about 90 tokens; SKILL.md has 353 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~90
When it runs · the whole SKILL.md, loaded when a task matches
~1k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from jeremylongshore/tons-of-skills-marketplace at commit cfae287, republished under its MIT licence (© jeremylongshore). 353 words, ~1,007 tokens.

Download SKILL.mdSave it as .claude/skills/brightdata-core-workflow-b/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
brightdata-core-workflow-b
description
Analyze and operate the Bright Data Web Scraper API async snapshot lifecycle with terminal-state and delivery controls. Use when collecting an approved batch, polling a snapshot, or downloading large structured results. Trigger with: "run a Bright Data dataset job", "poll a Bright Data snapshot", "download Web Scraper API results".
allowed-tools
Read, Grep, Write, Edit, Bash(curl:*)
compatibility
Requires an approved Bright Data account or offline fixtures, current Bright Data documentation, and an authorized public-data collection purpose
version
2.0.0
argument-hint
[dataset-run-manifest]
model
inherit
effort
high
license
MIT
author
Jeremy Longshore <jeremy@intentsolutions.io>
tags
saas, web-data, bright-data, core-workflow-b, operations

Bright Data Async Snapshot Pipeline

Overview

Implement the documented trigger, progress, and download lifecycle as a state machine. Keep dataset identifiers and approved inputs separate from the API key; cap polling, validate terminal states, and stream results rather than loading arbitrary payloads into memory.

Prerequisites

  • A reviewed dataset ID and approved public input manifest
  • A named-user Bright Data API key in the runtime secret manager
  • A retention, schema, maximum-record, and maximum-byte policy

Instructions

Step 1: Validate the manifest

Read the requested dataset/inputs and Grep for disallowed target classes or fields. Hash the approved input manifest before submission.

Step 2: Trigger once

Use Bash(curl:*) only against the fixed Bright Data API origin with the API key as a Bearer header.

bash
curl --fail-with-body --request POST \
  'https://api.brightdata.com/datasets/v3/trigger?dataset_id=DATASET_ID' \
  --header "Authorization: Bearer $BRIGHTDATA_API_KEY" \
  --header 'Content-Type: application/json' \
  --data-binary @approved-inputs.json
Step 3: Poll deliberately

Persist the returned snapshot_id; poll GET /datasets/v3/progress/SNAPSHOT_ID with bounded attempts and provider-directed delay. Handle ready, failed, empty, expired, and still-building states explicitly.

Step 4: Download and verify

Stream GET /datasets/v3/snapshot/SNAPSHOT_ID in the approved format. Use parts for large results, hold format/compression parameters constant, enforce byte/record ceilings, and validate the schema before promotion.

Tool Discipline

Use Read and Grep for policy and schema checks. Use Write and Edit only for the manifest, state machine, tests, and redacted receipt. Use Bash(curl:*) for fixed-origin Bright Data API calls after authorization; never print the Bearer value or raw result data.

Show full SKILL.md (134 more words)Show less

Output

  • Input-manifest hash and snapshot identifier
  • Bounded state-transition log without target data or credentials
  • Schema/size validation and a promoted-or-quarantined result

Examples

Submit a small approved URL batch, record the returned snapshot ID, poll until ready, and stream JSON to quarantine. Reject a response whose schema, record count, or byte count exceeds the manifest even if the provider marks it ready.

Error Handling

FailureMeaningResponse
400 validation responseDataset ID or input shape is invalidCorrect the manifest; do not retry unchanged input
429 or too many jobsTenant or dataset concurrency is exhaustedPause new triggers and wait for owned jobs
Snapshot expired or emptyResult cannot be promotedTrigger a newly approved run or investigate inputs

Resources

© jeremylongshore, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in skills/.curated/brightdata-core-workflow-b of jeremylongshore/tons-of-skills-marketplace.

  • SKILL.md
  • references/official-docs.md

Open the folder on GitHubat commit cfae287

Compare with similar skills

Brightdata Core Workflow B next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Brightdata Core Workflow B compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Brightdata Core Workflow B this skilljeremylongshore/tons-of-skills-marketplace2.8k—~1kAutomated safety check: PassMIT
Data Feedsbrightdata/skills264—~2.2kAutomated safety check: PassMIT
Bright Data Best Practicesbrightdata/skills2641 repos~3.6kAutomated safety check: PassMIT
Scrapebrightdata/skills264—~1.2kAutomated safety check: PassMIT
Searchbrightdata/skills264—~1.5kAutomated safety check: PassMIT
Brd Browser Debugbrightdata/skills264—~2kAutomated safety check: PassMIT

Similar skills

  • Data Feeds

    brightdata/skills

    Extract structured data from 40+ supported platforms (Amazon, LinkedIn, Instagram, TikTok, Facebook, YouTube, Reddit, and more) via the Bright Data CLI (bdata pipelines).

    264 GitHub stars~2.2k tokensUpdated 3 days ago
    Data & AnalyticsAuto-check passed
  • Build production-ready Bright Data integrations with best practices baked in.

    264 GitHub starsUsed in 1 repo~3.6k tokens
    Data & AnalyticsAuto-check passed
  • Scrape

    brightdata/skills

    Scrape web content as clean markdown/HTML/JSON via the Bright Data CLI (bdata scrape).

    264 GitHub stars~1.2k tokensUpdated 3 days ago
    Data & AnalyticsAuto-check passed
  • Search

    brightdata/skills

    Search the web via the Bright Data CLI — bdata search for Google/Bing/Yandex SERP, bdata discover for intent-ranked semantic results.

    264 GitHub stars~1.5k tokensUpdated 3 days ago
    Data & AnalyticsAuto-check passed
  • Brd Browser Debug

    brightdata/skills

    Debug Bright Data Scraping Browser sessions using the Browser Sessions API.

    264 GitHub stars~2k tokensUpdated 3 days ago
    Data & AnalyticsAuto-check passed
  • Brightdata SDK

    brightdata/skills

    Web data extraction and discovery using the Bright Data Python SDK.

    264 GitHub stars~5.2k tokensUpdated 3 days ago
    Data & AnalyticsAuto-check passed

More from jeremylongshore/tons-of-skills-marketplace

All 3,342 skills in this repo
  • Performing Security Code Review

    jeremylongshore/tons-of-skills-marketplace

    Execute this skill enables AI assistant to conduct a security-focused code review using the security-agent plugin.

    2.8k GitHub starsUsed in 2 repos~1.3k tokens
    Auto-check: notes
  • Adapting Transfer Learning Models

    jeremylongshore/tons-of-skills-marketplace

    Build this skill automates the adaptation of pre-trained machine learning models using transfer learning techniques.

    2.8k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Agent Context Loader

    jeremylongshore/tons-of-skills-marketplace

    Execute proactive auto-loading: automatically detects and loads agents.md files.

    2.8k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Aggregating Performance Metrics

    jeremylongshore/tons-of-skills-marketplace

    Aggregate and centralize performance metrics from applications, systems, databases, caches, and services.

    2.8k GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • Analyzing Capacity Planning

    jeremylongshore/tons-of-skills-marketplace

    Execute this skill enables AI assistant to analyze capacity requirements and plan for future growth.

    2.8k GitHub stars~947 tokensUpdated today
    Auto-check passed
  • Analyzing Database Indexes

    jeremylongshore/tons-of-skills-marketplace

    Process use when you need to work with database indexing. An agent skill from jeremylongshore/tons-of-skills-marketplace.

    2.8k GitHub stars~2k tokensUpdated today
    Auto-check passed

Works with

Questions about Brightdata Core Workflow B

What does Brightdata Core Workflow B do?

Analyze and operate the Bright Data Web Scraper API async snapshot lifecycle with terminal-state and delivery controls. Brightdata Core Workflow B is an agent skill from jeremylongshore/tons-of-skills-marketplace. Analyze and operate the Bright Data Web Scraper API async snapshot lifecycle with terminal-state and delivery controls.

When should I use Brightdata Core Workflow B?

Brightdata Core Workflow B fits situations like: collecting an approved batch; polling a snapshot; downloading large structured results; with: run a Bright Data dataset job.

How do I install Brightdata Core Workflow B in Claude Code?

Run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill brightdata-core-workflow-b -a claude-code`. Or copy the skill folder (skills/.curated/brightdata-core-workflow-b in jeremylongshore/tons-of-skills-marketplace) into .claude/skills/brightdata-core-workflow-b in your project. Claude Code loads it when a task matches its description.

How do I install Brightdata Core Workflow B in Codex?

Run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill brightdata-core-workflow-b -a codex`. Or copy the skill folder (skills/.curated/brightdata-core-workflow-b in jeremylongshore/tons-of-skills-marketplace) into .agents/skills/brightdata-core-workflow-b in your project. Codex loads it when a task matches its description.

Can I use Brightdata Core Workflow B in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill brightdata-core-workflow-b -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/brightdata-core-workflow-b, .gemini/skills/brightdata-core-workflow-b, .github/skills/brightdata-core-workflow-b and .opencode/skills/brightdata-core-workflow-b in your project.

What does Brightdata Core Workflow B need to run?

Going by SKILL.md and its folder, Brightdata Core Workflow B needs the command-line tools its instructions call (curl) and credentials named BRIGHTDATA_API_KEY. Our summary lists: A credential in BRIGHTDATA_API_KEY. Its frontmatter pre-approves these tools: Read, Grep, Write, Edit, Bash(curl:*). Compatibility (from SKILL.md): Requires an approved Bright Data account or offline fixtures, current Bright Data documentation, and an authorized public-data collection purpose.

Does Brightdata Core Workflow B access the network?

SKILL.md names 2 domains. In commands or code: api.brightdata.com; the agent is likely to contact it when it follows the instructions. As links in the text: docs.brightdata.com. This is read from the text; nothing was executed.

Is Brightdata Core Workflow B safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Brightdata Core Workflow B use?

Brightdata Core Workflow B is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Brightdata Core Workflow B use?

About 1k tokens (SKILL.md is roughly 4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 201 tokens, read only when the agent opens those files.

What are the alternatives to Brightdata Core Workflow B?

Skills that share tags, products or a category with Brightdata Core Workflow B: Data Feeds (brightdata/skills, 264 stars), Bright Data Best Practices (brightdata/skills, 264 stars), Scrape (brightdata/skills, 264 stars) and Search (brightdata/skills, 264 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Brightdata Core Workflow B?

jeremylongshore (a GitHub user) maintains it in jeremylongshore/tons-of-skills-marketplace, which has 2,827 GitHub stars. The repository holds 3,342 skills in this directory. The repository was last updated on October 10, 2026.

Source: jeremylongshore/tons-of-skills-marketplace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.