MCP Server Builder
anthropics/skills
Guides the design and implementation of Model Context Protocol servers in TypeScript or Python, from tool naming and error messages to evaluation.
Teaches the agent to use the Visual Memory MCP server to cache webpage and application screenshots, matching layout states and avoiding redundant LLM vision calls.
$ npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install putervision/state-memory-mcp vision-memory-mcp --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/putervision/state-memory-mcp.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/vision-memory-mcp .claude/skills/vision-memory-mcp && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "vision-memory-mcp" agent skill from https://github.com/putervision/state-memory-mcp/tree/main/.agents/skills/vision-memory-mcp into .claude/skills/vision-memory-mcp/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-memory-mcp", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/putervision/state-memory-mcp/tree/main/.agents/skills/vision-memory-mcpType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install putervision/state-memory-mcp vision-memory-mcp --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/putervision/state-memory-mcp.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/vision-memory-mcp .agents/skills/vision-memory-mcp && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "vision-memory-mcp" agent skill from https://github.com/putervision/state-memory-mcp/tree/main/.agents/skills/vision-memory-mcp into .agents/skills/vision-memory-mcp/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-memory-mcp", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install putervision/state-memory-mcp vision-memory-mcp --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/putervision/state-memory-mcp.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/vision-memory-mcp .cursor/skills/vision-memory-mcp && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "vision-memory-mcp" agent skill from https://github.com/putervision/state-memory-mcp/tree/main/.agents/skills/vision-memory-mcp into .cursor/skills/vision-memory-mcp/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-memory-mcp", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/putervision/state-memory-mcp.git --path .agents/skills/vision-memory-mcp--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install putervision/state-memory-mcp vision-memory-mcp --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/putervision/state-memory-mcp.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/vision-memory-mcp .gemini/skills/vision-memory-mcp && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "vision-memory-mcp" agent skill from https://github.com/putervision/state-memory-mcp/tree/main/.agents/skills/vision-memory-mcp into .gemini/skills/vision-memory-mcp/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-memory-mcp", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install putervision/state-memory-mcp vision-memory-mcpInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/putervision/state-memory-mcp.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/vision-memory-mcp .github/skills/vision-memory-mcp && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "vision-memory-mcp" agent skill from https://github.com/putervision/state-memory-mcp/tree/main/.agents/skills/vision-memory-mcp into .github/skills/vision-memory-mcp/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-memory-mcp", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install putervision/state-memory-mcp vision-memory-mcp --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/putervision/state-memory-mcp.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/vision-memory-mcp .opencode/skills/vision-memory-mcp && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "vision-memory-mcp" agent skill from https://github.com/putervision/state-memory-mcp/tree/main/.agents/skills/vision-memory-mcp into .opencode/skills/vision-memory-mcp/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-memory-mcp", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
vision-memory-mcpTeaches the agent to use the Visual Memory MCP server to cache webpage and application screenshots, matching layout states and avoiding redundant LLM vision calls.
Vision Memory MCP is an agent skill from putervision/state-memory-mcp. Teaches the agent to use the Visual Memory MCP server to cache webpage and application screenshots, matching layout states and avoiding redundant LLM vision calls.
Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Agent Workflows, covering MCP servers. It works with Model Context Protocol. The repository describes itself as: Persistent, branch-aware workflow state memory MCP server for AI coding assistants. Tracks tasks, accepted decisions, and active blockers to prevent session context bloat and… The licence is MIT.
4 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 0f60ae4. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Vision Memory MCP loads about 1.4k tokens when it runs. Until then it costs about 45 tokens; SKILL.md has 571 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
rkspace .vision-memory-mcp/, .gitignore, .env, and IDE agent rules.Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from putervision/state-memory-mcp at commit 0f60ae4, republished under its MIT licence (© putervision). 571 words, ~1,387 tokens.
.claude/skills/vision-memory-mcp/SKILL.md (or your agent's skills folder).This project utilizes vision-memory-mcp to cache visual states, record layout transitions, and avoid repetitive LLM vision calls.
Whenever you capture a screenshot, examine a webpage, or need to verify a visual state, you MUST run this sequence:
get_session_context to load active transitions and recent visual states.recall_memory to search past visual states by text query or image query.analyze_screenshot with the base64 screenshot before calling any vision LLMs.is_known is true, read the returned description and do NOT call your vision LLM.is_known is false, inspect the image with your vision model, summarize the layout, and register it back by calling analyze_screenshot with both the screenshot and description parameters.record_outcome to build the navigation graph.manage_snapshot (action: "save") when reaching milestones, and manage_snapshot (action: "diff") to check for visual regressions.| Tool Name | Key Inputs | Description |
|---|---|---|
analyze_screenshot | screenshot? (base64), file_path?, description?, items? | Main ingestion (single or batch) and visual state retrieval tool. |
recall_memory | query?, screenshot?, file_path?, strategy?, limit? | Search visual memory by text query or image query (read-only). |
record_outcome | from_state_id, to_state_id?, action, action_type? ('blocker' | 'click' | etc.) | Record UI action transitions or log visual blockers for state-memory. |
get_navigation_paths | from_state_id?, to_state_id?, to_description?, max_hops? | Find historical path or instructions between states. |
predict_next_action | current_state_id, goal_description?, goal_state_id? | Predict best next UI action and grounded element handles (target_selector, target_coords). |
compare_states | state_a_id & state_b_id OR video_a_id & video_b_id | Compare two states visually (has_layout_change) or compare video runs. |
get_session_context | include_recent?, include_frequent? | Get recent/frequent states, transition graphs, disk stats, cache metrics, and version info. |
manage_snapshot | action ('save' | 'diff' | 'export' | 'restore'), name?, archive_json? | Unified snapshot management for visual checkpoints and regression detection. |
manage_visual_spec | action ('set' | 'verify' | 'list'), name?, screenshot?, tolerance? | Register and verify visual design contract baselines (Visual SDD). |
manage_video | action ('ingest' | 'search' | 'timeline'), file_path?, query?, video_id? | Ingest WebM/MP4 recordings, search video keyframes, or retrieve timelines. |
create_evidence_pack | keyframe_state_ids, source_video_id?, linked_state_memory_nodes? | Package immutable evidence packs linking video keyframes to state-memory DAGs. |
export_trajectories | format? ('json' | 'llava' | 'qwen2_vl' | 'joint'), trace_id? | Export multimodal trajectories for model fine-tuning or joint workflow exports. |
undo_visual_mutation | type? ('state' | 'transition' | 'any') | Revert the last visual state ingestion or transition edge addition. |
forget_state | state_id | Purge a specific state and vector embedding for privacy. |
wait_for_visual_state | target_state_id, timeout_ms? | Poll for target visual state until present or timeout occurs. |
To bypass confirmation dialogs when running CLI cache commands or reading/writing brain images, add these allows to your configuration:
~/.gemini/config/config.json): Add these rules to your "globalPermissionGrants" -> "allow" list:"command(vision-memory-mcp)" (Allows running any query/ingest command prefix)"read_file(.*\\.gemini/antigravity/brain/.*)" (Allows reading brain screenshots)"write_file(.*\\.gemini/antigravity/brain/.*)" (Allows saving brain snapshots)Run these commands in the terminal for management and analytics:
vision-memory-mcp init [-y|--yes]: Scaffold workspace .vision-memory-mcp/, .gitignore, .env, and IDE agent rules.vision-memory-mcp init-global: Re-initialize across all projects registered in ~/.vision-memory-mcp/projects.json.vision-memory-mcp doctor: Health check storage writability, sharp bindings, Node runtime, and sub-directory Git repos.vision-memory-mcp audit: Audit sub-directory Git repos, submodules, database locations, and total visual states.vision-memory-mcp inspect: Display stored visual states in an ASCII table.vision-memory-mcp metrics: Calculate cache hit rate, token savings, and ROI.vision-memory-mcp view: Open an interactive force-directed graph visualizer in the browser.vision-memory-mcp export --format [json\|mermaid\|html] --out [file]: Export the memory graph.vision-memory-mcp prune: Purge expired or low-access states.© putervision, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .agents/skills/vision-memory-mcp of putervision/state-memory-mcp.
Open the folder on GitHubat commit 0f60ae4
Vision Memory MCP next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Vision Memory MCP this skillputervision/state-memory-mcp | 111 | — | ~1.4k | Automated safety check: Notes | MIT | |
| MCP Server Builderanthropics/skills | 180k | 64 repos | ~2.3k | Automated safety check: Pass | Apache-2.0 | |
| MCP Server BuildershareAI-lab/learn-claude-code | 78k | 5 repos | ~1.2k | Automated safety check: Pass | MIT | |
| MCP Integration for Pluginsanthropics/claude-plugins-official | 38k | 11 repos | ~3.1k | Automated safety check: Pass | Apache-2.0 | |
| Fastmcp Client CLIPrefectHQ/fastmcp | 28k | 1 repos | ~823 | Automated safety check: Pass | Apache-2.0 | |
| Crush Configurationcharmbracelet/crush | 29k | — | ~3.7k | Automated safety check: Pass | Custom licence |
anthropics/skills
Guides the design and implementation of Model Context Protocol servers in TypeScript or Python, from tool naming and error messages to evaluation.
shareAI-lab/learn-claude-code
Walks through building MCP servers in Python or TypeScript that expose tools, resources and prompts to Claude, with templates, registration and testing.
anthropics/claude-plugins-official
Explains how to bundle Model Context Protocol servers in a Claude Code plugin, covering config files, stdio, SSE, HTTP and WebSocket server types, and authentication.
PrefectHQ/fastmcp
Query and invoke tools on MCP servers using fastmcp list and fastmcp call.
charmbracelet/crush
Explains how to configure the Crush coding agent with crushrc or crush.json, covering providers, models, LSPs, MCP servers, hooks, permissions and config precedence.
mksglu/context-mode
Routes large command, file, API and browser output through context-mode tools so only the needed result enters the agent's context, instead of dumping it via Bash.
putervision/state-memory-mcp
Teaches the agent to use the Behavior MCP server for ~60Hz in-browser behavior trees, triggers, and recordings.
putervision/state-memory-mcp
Teaches the agent to use the state-memory-mcp MCP server to track workflow state, tasks, decisions, blockers, artifacts, plans, milestones, and their semantic relationships in a persistent graph…
putervision/state-memory-mcp
Teaches the agent to process, ingest, analyze, and compare WebM, MP4, and GIF video recordings using vision-memory-mcp and state-memory-mcp.
putervision/state-memory-mcp
Teaches the agent to use the WebCrypt MCP server for AES-256-GCM symmetric encryption, RSA-4096 hybrid encryption, key generation, digital signatures, hashing, and post-quantum cryptography.
putervision/state-memory-mcp
Teaches the agent to use the Strategic Agent Reasoning MCP server for BDI goals, utility scoring, risk evaluation, and replanning.
putervision/state-memory-mcp
Teaches the agent to use the Spatial World Model MCP server to track entities, 3D/2D positions, spatial relationships, object permanence, and movement simulation.
Works with
Categories
Teaches the agent to use the Visual Memory MCP server to cache webpage and application screenshots, matching layout states and avoiding redundant LLM vision calls. Vision Memory MCP is an agent skill from putervision/state-memory-mcp. Teaches the agent to use the Visual Memory MCP server to cache webpage and application screenshots, matching layout states and avoiding redundant LLM vision calls.
Vision Memory MCP fits situations like: tasks that involve MCP servers.
Run `npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a claude-code`. Or copy the skill folder (.agents/skills/vision-memory-mcp in putervision/state-memory-mcp) into .claude/skills/vision-memory-mcp in your project. Claude Code loads it when a task matches its description.
Run `npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a codex`. Or copy the skill folder (.agents/skills/vision-memory-mcp in putervision/state-memory-mcp) into .agents/skills/vision-memory-mcp in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add putervision/state-memory-mcp --skill vision-memory-mcp -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/vision-memory-mcp, .gemini/skills/vision-memory-mcp, .github/skills/vision-memory-mcp and .opencode/skills/vision-memory-mcp in your project.
SKILL.md names no scripts, command-line tools or credentials: Vision Memory MCP is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.
Vision Memory MCP is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.4k tokens (SKILL.md is roughly 5.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Vision Memory MCP: MCP Server Builder (anthropics/skills, 180k stars), MCP Server Builder (shareAI-lab/learn-claude-code, 78k stars), MCP Integration for Plugins (anthropics/claude-plugins-official, 38k stars) and Fastmcp Client CLI (PrefectHQ/fastmcp, 28k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
putervision (a GitHub organization) maintains it in putervision/state-memory-mcp, which has 111 GitHub stars. The repository holds 7 skills in this directory. The repository was last updated on October 3, 2026.
Source: putervision/state-memory-mcp on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.