Agent skill

Wayback API Archive Recovery

by divinevideo in divinevideo/divine-mobile

Recover data from archived REST APIs using Wayback Machine. An agent skill from divinevideo/divine-mobile.

MPL-2.0Auto-check passedBackend & APIs

Install Wayback API Archive Recovery

skills CLI
$ npx skills add divinevideo/divine-mobile --skill wayback-api-archive-recovery -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install divinevideo/divine-mobile wayback-api-archive-recovery --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/divinevideo/divine-mobile.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/wayback-api-archive-recovery .claude/skills/wayback-api-archive-recovery && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
wayback-api-archive-recovery
GitHub stars
266
Token cost
~1.4k tokens
SKILL.md length
318 words
Files
1
Skills in repo
103
Repo updated
First seen
Licence
MPL-2.0

At a glance

Recover data from archived REST APIs using Wayback Machine. An agent skill from divinevideo/divine-mobile.

  • Works in 5 steps: Discover What Was Archived → Understand What Gets Archived → Fetch Archived API Responses → …
  • Recovering data from defunct services like Vine
  • SKILL.md covers Problem, Context / Trigger Conditions, Solution and Verification, plus 3 more sections
  • Calls curl; reaches web.archive.org

What it does

Wayback API Archive Recovery is an agent skill from divinevideo/divine-mobile. Recover data from archived REST APIs using Wayback Machine. Use when: (1) Recovering data from defunct services like Vine, Twitter, or other platforms, (2) Need to find which API endpoints were archived vs which require authentication, (3) Building data recovery pipelines from web.archive.org. Covers CDX API queries to discover archived endpoints, understanding what gets archived (public) vs what doesn't (authenticated), and extracting embedded JSON from archived web pages as a fallback.

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Backend & APIs, covering REST APIs. It works with X (Twitter). The licence is MPL-2.0.

When your agent uses it

  • Recovering data from defunct services like Vine
  • Other platforms
  • Need to find which API endpoints were archived vs which require authentication
  • Building data recovery pipelines from web.archive.org

Example prompts

  • “/wayback-api-archive-recovery”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Discover What Was Archived
  2. Understand What Gets Archived
  3. Fetch Archived API Responses
  4. Extract Embedded Data as Fallback
  5. Infer Missing Data from Interactions

What it can do on your machine

Read from SKILL.md and the folder at commit c3d6f7e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • web.archive.org

    Also links to:

    • github.com
    • archive.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Wayback API Archive Recovery loads about 1.4k tokens when it runs. Until then it costs about 130 tokens; SKILL.md has 318 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~130
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from divinevideo/divine-mobile at commit c3d6f7e, republished under its MPL-2.0 licence (© divinevideo). 318 words, ~1,444 tokens.

Download SKILL.mdSave it as .claude/skills/wayback-api-archive-recovery/SKILL.md (or your agent's skills folder).
name
wayback-api-archive-recovery
description
Recover data from archived REST APIs using Wayback Machine. Use when: (1) Recovering data from defunct services like Vine, Twitter, or other platforms, (2) Need to find which API endpoints were archived vs which require authentication, (3) Building data recovery pipelines from web.archive.org. Covers CDX API queries to discover archived endpoints, understanding what gets archived (public) vs what doesn't (authenticated), and extracting embedded JSON from archived web pages as a fallback.
author
Claude Code
version
1.0.0
date
2025-01-20

Wayback Machine API Archive Recovery

Problem

When recovering data from defunct services, you need to know which API endpoints were archived by the Wayback Machine, and how to access them. Not all endpoints are archived - authenticated endpoints typically aren't captured.

Context / Trigger Conditions

  • Recovering data from a shut-down service (Vine, old Twitter API, defunct platforms)
  • Building an archive crawler for digital preservation
  • Need to reconstruct engagement data (comments, likes, followers) from archives
  • Getting empty results when trying to fetch archived API endpoints

Solution

Step 1: Discover What Was Archived

Use the CDX API to search for archived endpoints:

bash
# Find all archived endpoints under a path prefix
curl "https://web.archive.org/cdx/search/cdx?url=vine.co/api/posts/&matchType=prefix&output=json&fl=original,timestamp&filter=statuscode:200&limit=100"

# Check total pages available
curl "https://web.archive.org/cdx/search/cdx?url=vine.co/api/&matchType=prefix&showNumPages=true"

Key CDX parameters:

  • matchType=prefix - Match URL prefixes with wildcards
  • fl=original,timestamp - Select which fields to return
  • filter=statuscode:200 - Only successful captures
  • collapse=urlkey - Deduplicate by URL
  • limit=N - Cap results
Step 2: Understand What Gets Archived

Typically archived (public endpoints):

  • Comment lists: /api/posts/{id}/comments
  • User profiles: /api/users/profiles/{id}
  • Timelines: /api/timelines/users/{id}
  • Public feeds: /api/timelines/promoted

Typically NOT archived (require authentication):

  • Like lists: /api/posts/{id}/likes
  • Follower/following lists: /api/users/{id}/followers
  • Private data: DMs, settings, notifications
Step 3: Fetch Archived API Responses

Use the id_ modifier to get raw JSON without the Wayback toolbar:

bash
# Without id_ - may return HTML wrapper
curl "https://web.archive.org/web/20161209032217/https://vine.co/api/posts/123/comments"

# With id_ - returns raw JSON
curl "https://web.archive.org/web/20161209032217id_/https://vine.co/api/posts/123/comments"
Step 4: Extract Embedded Data as Fallback

When API endpoints weren't archived, check if data was embedded in archived web pages:

python
# Many SPAs embed initial data in script tags
import re
import json

html = fetch_archived_page(url)

# Look for embedded JSON (common patterns)
patterns = [
    r'window\.POST_DATA\s*=\s*({.*?});',      # Vine
    r'window\.__INITIAL_STATE__\s*=\s*({.*?});',  # Redux apps
    r'<script type="application/json"[^>]*>({.*?})</script>',
]

for pattern in patterns:
    match = re.search(pattern, html, re.DOTALL)
    if match:
        data = json.loads(match.group(1))

Caveat: Embedded data is often partial (e.g., only first 3 comments rendered).

Step 5: Infer Missing Data from Interactions

When follower/following lists aren't available, infer relationships:

python
# If A commented on B's post → A probably follows B
inferred_follows.append({
    'follower': commenter_id,
    'followee': post_creator_id,
    'inference_type': 'commented_on_post',
    'confidence': 0.7
})

# If A mentioned @B → A probably follows B
# If A and B mutually commented → bidirectional follow (collaborators)

Verification

  1. CDX query returns results with status 200
  2. Fetching with id_ modifier returns valid JSON
  3. Parsing embedded data produces expected structure

Example

python
import urllib.request
import json

def find_archived_endpoints(base_url):
    """Find all archived API endpoints for a service."""
    cdx_url = f"https://web.archive.org/cdx/search/cdx?url={base_url}&matchType=prefix&output=json&fl=original,timestamp&filter=statuscode:200"

    with urllib.request.urlopen(cdx_url, timeout=60) as resp:
        data = json.loads(resp.read())

    # Skip header row, extract unique URLs
    endpoints = {}
    for row in data[1:]:
        url, timestamp = row[0], row[1]
        if url not in endpoints:
            endpoints[url] = timestamp

    return endpoints

def fetch_archived_json(url, timestamp):
    """Fetch archived JSON using id_ modifier."""
    archive_url = f"https://web.archive.org/web/{timestamp}id_/{url}"

    with urllib.request.urlopen(archive_url, timeout=30) as resp:
        return json.loads(resp.read())

# Usage
endpoints = find_archived_endpoints("vine.co/api/posts/")
print(f"Found {len(endpoints)} archived endpoints")

# Filter to comment endpoints
comments = {k: v for k, v in endpoints.items() if '/comments' in k}

Notes

  • Wayback Machine rate limits requests - add delays (1+ second) between fetches
  • Archived timestamps vary - same endpoint may have captures from different dates
  • Paginated APIs: check for page, offset, cursor parameters in archived URLs
  • Some archives return 302 redirects - follow them or check for "stale" archives
  • The CDX API itself can be slow for large result sets - use pagination

References

© divinevideo, MPL-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/wayback-api-archive-recovery of divinevideo/divine-mobile.

Open the folder on GitHubat commit c3d6f7e

Compare with similar skills

Wayback API Archive Recovery next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Wayback API Archive Recovery compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Wayback API Archive Recovery this skilldivinevideo/divine-mobile266—~1.4kAutomated safety check: PassMPL-2.0
Twitterapi IokaitoInfra/twitterapi-io451—~3.1kAutomated safety check: PassMIT
Twitterapi Io ReadkaitoInfra/twitterapi-io451—~1.1kAutomated safety check: PassMIT
Bankr SignalsaAAaqwq/AGI-Super-Team1052 repos~3.3kAutomated safety check: PassMIT
X Twitter Scrapersickn33/agentic-awesome-skills47k1 repos~2.1kAutomated safety check: PassMIT
X Twitter Scrapergithub/awesome-copilot40k—~2.2kAutomated safety check: PassMIT

Similar skills

  • Twitterapi Io

    kaitoInfra/twitterapi-io

    Official skill for twitterapi.io — query Twitter/X data (tweets, profiles, followers, advanced search, trends, spaces, communities, lists) and perform authenticated actions (post, reply, like…

    451 GitHub stars~3.1k tokensUpdated today
    Backend & APIsAuto-check passed
  • Twitterapi Io Read

    kaitoInfra/twitterapi-io

    Read-only skill for twitterapi.io — look up public X/Twitter data (tweets, user profiles, followers and followings, advanced search, replies, quotes, trends, lists, communities) through the…

    451 GitHub stars~1.1k tokensUpdated today
    Backend & APIsAuto-check passed
  • Bankr Signals

    aAAaqwq/AGI-Super-Team

    Transaction-verified trading signals on Base blockchain. An agent skill from aAAaqwq/AGI-Super-Team.

    105 GitHub starsUsed in 2 repos~3.3k tokens
    Backend & APIsAuto-check passed
  • X Twitter Scraper

    sickn33/agentic-awesome-skills

    Use Xquik for X data workflows: tweet search, user lookup, follower export, media downloads, monitors, webhooks, REST API, MCP, SDK setup, and approval-gated account actions.

    47k GitHub starsUsed in 1 repo~2.1k tokens
    Backend & APIsAuto-check passed
  • X Twitter Scraper

    github/awesome-copilot

    Official

    Build GitHub Copilot workflows with Xquik X API SDKs, REST endpoints, hosted Apify Actor runs, MCP tools, TweetClaw OpenClaw plugin installs, signed webhooks, tweet search, user lookup, follower…

    40k GitHub stars~2.2k tokensUpdated 2 days ago
    Backend & APIsAuto-check passed
  • X Twitter Scraper

    davila7/claude-code-templates

    X API & Twitter scraper skill for AI coding agents. An agent skill from davila7/claude-code-templates.

    33k GitHub stars~1.8k tokensUpdated yesterday
    Backend & APIsAuto-check passed

More from divinevideo/divine-mobile

All 103 skills in this repo
  • Fix ArgoCD ExternalSecret deployment failing with "namespace X is not permitted in project Y".

    266 GitHub stars~931 tokensUpdated today
    Auto-check passed
  • Art Direct

    divinevideo/divine-mobile

    Art direction for any content — reads text, PDF, Word, HTML, PPT, then proposes 2-3 creative directions with photography style, mood, and visual language.

    266 GitHub stars~4.8k tokensUpdated today
    Auto-check passed
  • Async Await Null Race Condition

    divinevideo/divine-mobile

    Fix "Null check operator used on a null value" errors when an object is set to null during an async await.

    266 GitHub stars~881 tokensUpdated today
    Auto-check passed
  • AWS V4 Signing Custom Headers Gcs

    divinevideo/divine-mobile

    Add custom metadata headers (x-amz-meta-) to AWS v4 signed requests for GCS S3-compatible API.

    266 GitHub stars~1k tokensUpdated today
    Auto-check passed
  • Bash Herestring Newline Secrets

    divinevideo/divine-mobile

    Fix password/secret authentication failures caused by trailing newlines when creating Google Cloud secrets (or similar) with bash here-strings.

    266 GitHub stars~791 tokensUpdated today
    Auto-check passed
  • Fix silent video/media processing failures caused by URL extraction code that filters on file extensions (.mp4, .webm, .webp).

    266 GitHub stars~1.1k tokensUpdated today
    Auto-check passed

Works with

Categories

Questions about Wayback API Archive Recovery

What does Wayback API Archive Recovery do?

Recover data from archived REST APIs using Wayback Machine. An agent skill from divinevideo/divine-mobile. Wayback API Archive Recovery is an agent skill from divinevideo/divine-mobile. Recover data from archived REST APIs using Wayback Machine.

When should I use Wayback API Archive Recovery?

Wayback API Archive Recovery fits situations like: recovering data from defunct services like Vine; other platforms; need to find which API endpoints were archived vs which require authentication; building data recovery pipelines from web.archive.org.

How do I install Wayback API Archive Recovery in Claude Code?

Run `npx skills add divinevideo/divine-mobile --skill wayback-api-archive-recovery -a claude-code`. Or copy the skill folder (.agents/skills/wayback-api-archive-recovery in divinevideo/divine-mobile) into .claude/skills/wayback-api-archive-recovery in your project. Claude Code loads it when a task matches its description.

How do I install Wayback API Archive Recovery in Codex?

Run `npx skills add divinevideo/divine-mobile --skill wayback-api-archive-recovery -a codex`. Or copy the skill folder (.agents/skills/wayback-api-archive-recovery in divinevideo/divine-mobile) into .agents/skills/wayback-api-archive-recovery in your project. Codex loads it when a task matches its description.

Can I use Wayback API Archive Recovery in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add divinevideo/divine-mobile --skill wayback-api-archive-recovery -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/wayback-api-archive-recovery, .gemini/skills/wayback-api-archive-recovery, .github/skills/wayback-api-archive-recovery and .opencode/skills/wayback-api-archive-recovery in your project.

What does Wayback API Archive Recovery need to run?

Going by SKILL.md and its folder, Wayback API Archive Recovery needs the command-line tools its instructions call (curl). Our summary lists: Python 3.

Does Wayback API Archive Recovery access the network?

SKILL.md names 3 domains. In commands or code: web.archive.org; the agent is likely to contact it when it follows the instructions. As links in the text: github.com and archive.org. This is read from the text; nothing was executed.

Is Wayback API Archive Recovery safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Wayback API Archive Recovery use?

Wayback API Archive Recovery is published under the MPL-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Wayback API Archive Recovery use?

About 1.4k tokens (SKILL.md is roughly 5.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Wayback API Archive Recovery?

Skills that share tags, products or a category with Wayback API Archive Recovery: Twitterapi Io (kaitoInfra/twitterapi-io, 451 stars), Twitterapi Io Read (kaitoInfra/twitterapi-io, 451 stars), Bankr Signals (aAAaqwq/AGI-Super-Team, 105 stars) and X Twitter Scraper (sickn33/agentic-awesome-skills, 47k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Wayback API Archive Recovery?

divinevideo (a GitHub organization) maintains it in divinevideo/divine-mobile, which has 266 GitHub stars. The repository holds 103 skills in this directory. The repository was last updated on October 10, 2026.

Source: divinevideo/divine-mobile on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.