Read GitHub
AgentTeam-TaichuAI/ScienceClaw
Read and search GitHub repository documentation via gitmcp.io MCP service.
Find official portals, APIs, and download paths for authoritative primary data sources (governments, international organizations, research institutions, etc.).
$ npx skills add MLT-OSS/FirstData --skill firstdata -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install MLT-OSS/FirstData firstdata --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/MLT-OSS/FirstData.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/firstdata .claude/skills/firstdata && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "firstdata" agent skill from https://github.com/MLT-OSS/FirstData/tree/main/skills/firstdata into .claude/skills/firstdata/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "firstdata", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/MLT-OSS/FirstData/tree/main/skills/firstdataType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add MLT-OSS/FirstData --skill firstdata -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install MLT-OSS/FirstData firstdata --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MLT-OSS/FirstData.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/firstdata .agents/skills/firstdata && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "firstdata" agent skill from https://github.com/MLT-OSS/FirstData/tree/main/skills/firstdata into .agents/skills/firstdata/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "firstdata", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add MLT-OSS/FirstData --skill firstdata -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install MLT-OSS/FirstData firstdata --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MLT-OSS/FirstData.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/firstdata .cursor/skills/firstdata && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "firstdata" agent skill from https://github.com/MLT-OSS/FirstData/tree/main/skills/firstdata into .cursor/skills/firstdata/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "firstdata", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/MLT-OSS/FirstData.git --path skills/firstdata--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add MLT-OSS/FirstData --skill firstdata -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install MLT-OSS/FirstData firstdata --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MLT-OSS/FirstData.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/firstdata .gemini/skills/firstdata && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "firstdata" agent skill from https://github.com/MLT-OSS/FirstData/tree/main/skills/firstdata into .gemini/skills/firstdata/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "firstdata", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install MLT-OSS/FirstData firstdataInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add MLT-OSS/FirstData --skill firstdata -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/MLT-OSS/FirstData.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/firstdata .github/skills/firstdata && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "firstdata" agent skill from https://github.com/MLT-OSS/FirstData/tree/main/skills/firstdata into .github/skills/firstdata/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "firstdata", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add MLT-OSS/FirstData --skill firstdata -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install MLT-OSS/FirstData firstdata --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MLT-OSS/FirstData.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/firstdata .opencode/skills/firstdata && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "firstdata" agent skill from https://github.com/MLT-OSS/FirstData/tree/main/skills/firstdata into .opencode/skills/firstdata/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "firstdata", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
firstdataFind official portals, APIs, and download paths for authoritative primary data sources (governments, international organizations, research institutions, etc.).
Firstdata is an agent skill from MLT-OSS/FirstData. Find official portals, APIs, and download paths for authoritative primary data sources (governments, international organizations, research institutions, etc.). Use when users need to know "where to find this data from an official source", "which source is more authoritative", or "how to cite primary data". Covers the live FirstData catalog with authority comparison and site navigation guidance.
Its SKILL.md is about 3.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including reference files (for example `mcp-tool-descriptions-draft.md` and `references/firstdata-register.md`).
It sits in Research & Science, covering MCP servers. It works with Model Context Protocol. The repository describes itself as: The World's Most Comprehensive, Authoritative, and Structured Open Source Data Source Knowledge Base. The licence is MIT.
Read from SKILL.md and the folder at commit e4a6687. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
npxFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
firstdata.deepminer.com.cnAlso links to:
arxiv.orgFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
FIRSTDATA_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Firstdata loads about 3.1k tokens when it runs, and up to ~4.3k if it reads all its reference files. Until then it costs about 102 tokens; SKILL.md has 1,330 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from MLT-OSS/FirstData at commit e4a6687, republished under its MIT licence (© MLT-OSS). 1,330 words, ~3,054 tokens.
.claude/skills/firstdata/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.FirstData is the External Facts Context Layer for AI Agents — a purpose-built, authoritative collection of primary data sources that helps agents locate official origins rather than generating unverified answers.
It does not replace raw data — it acts as an "authoritative data navigator", taking vague user needs as input, recommending the most appropriate primary sources, and providing clear access paths, API information, and download methods so both users and agents can trace back to original evidence. The live catalog size is returned by the MCP get_status tool and should not be hard-coded in agent responses.
Coverage:
When to use: When users need to find official data sources, compare source authority, obtain official URLs/APIs/download paths, or build evidence-chain workflows. FirstData is a source locator, not an answer generator — after receiving results, guide users back to original sources for verification rather than treating them as final answers.
1. Source Locator — Returns the top 3–5 most relevant sources with authority level, matching rationale, access URL, API documentation, and download methods.
2. Site Pathfinder — Provides step-by-step navigation from homepage to target data for complex official websites, including alternative paths and API access methods.
3. Evidence-Ready Workflows — Can be embedded into workflows requiring evidence chains: deep research, policy analysis, investment research, compliance auditing, fact-checking, etc.
Each data source includes structured metadata: authority level (government / international / research / market / commercial / other), access URL, API information, download formats, geographic scope, update frequency, access level, etc.
Typical query scenarios when agents call FirstData via MCP:
| User Need | Query Direction | Expected Output |
|---|---|---|
| "Which official source should I cite for China's 2023 NEV export volume?" | China Customs, National Bureau of Statistics | Official source + authority level + data page URL |
| "Where to download IPO prospectus for a Hong Kong-listed company?" | HKEXnews | Official platform + step-by-step navigation |
| "World Bank vs IMF GDP data — which is better for academic citation?" | World Bank WDI, IMF WEO | Source comparison + authority differences + API docs |
| "Need global climate data with API access" | NASA Earthdata, NOAA CDO | Data source + API docs + access methods |
| "Where is the official data for China's M2 money supply?" | People's Bank of China | Official data portal + update frequency + historical coverage |
Full project background and feature documentation: README
This skill connects to the FirstData MCP server (firstdata.deepminer.com.cn), the project's official hosted API endpoint. An API key (FIRSTDATA_API_KEY) is required for authentication.
If you already have FIRSTDATA_API_KEY set, configure the MCP connection:
npx mcporter config add firstdata https://firstdata.deepminer.com.cn/mcp --header 'Authorization=Bearer ${FIRSTDATA_API_KEY}'Or add manually to your MCP config:
{
"mcpServers": {
"firstdata": {
"type": "streamable-http",
"url": "https://firstdata.deepminer.com.cn/mcp",
"headers": {
"Authorization": "Bearer <FIRSTDATA_API_KEY>"
}
}
}
}If you don't have an API key, see firstdata-register.md for the registration process (two API calls to the FirstData server to obtain a JWT token).
Once connected, call get_status first to verify the MCP connection, inspect the active tool list, and read the current catalog snapshot metadata. Then browse the tool list and select the appropriate tool based on your needs.
The FirstData MCP server provides 6 tools. Below is a reference with usage guidelines, limitations, and examples.
Authorization: Bearer <token> header.POST /api/token/verify) which returns remaining_daily in the response — this is a separate HTTP call, not available through MCP tool invocation.firstdata.deepminer.com.cn). Network latency and server availability affect response times.get_statusPurpose: Check the MCP server version, currently registered tools, and the catalog snapshot available to the server.
Use it first when configuring a new Agent or diagnosing a connection. The response includes catalog source counts, generated timestamp when available, and whether the runtime data directories are present. It does not return credentials.
search_sourcePurpose: Unified data source search tool supporting keyword search, structured filtering, pagination, and multiple output modes.
Limitations:
limit parameter range: 1–200, default: 20).["中国", "GDP"] (173 results) instead of ["中国 GDP"] (0 results). This is by design to preserve multi-word terms like "New Zealand" or "World Bank".domain parameter uses substring matching, not exact enum matching (e.g., "finance" matches "public-finance", "finance", "financial-markets").get_sourcePurpose: Retrieve full details for specific data sources by their IDs.
Limitations:
source_id values do NOT cause an error response (isError: false). Instead, the result array includes {"id": "xxx", "error": "Not found"} for each invalid ID alongside valid results. Callers must check individual items for error fields rather than relying solely on isError.source_ids per request, but performance with large batches (50+) is unverified. As a practical guideline (not a hard limit), consider batching in groups of ~20.fields parameter filters returned fields; when omitted, all fields are returned.ask_agentPurpose: LLM-powered intelligent search agent for complex, cross-domain, or ambiguous queries that require multi-step reasoning.
Limitations:
web_search for external information.jq for local data queries plus optional web_search. The web search step is not user-controllable.search_source instead for simple keyword matching or structured filtering — it is faster, deterministic, and cheaper.get_access_guidePurpose: Generate detailed access instructions for a specific data source using RAG (Retrieval-Augmented Generation).
Limitations:
source_id returns {"error": "数据源 xxx 不存在"}.top_k range: 1–5 (default: 3).operation parameter. Vague descriptions yield lower-quality matches. Use specific action verbs and entity names (e.g., "查询2024年M2货币供应量数据" rather than "查数据").report_feedbackPurpose: Submit user feedback to the development team when FirstData has a confirmed issue.
Limitations:
feedback_message length: 10–2,000 characters.Examples:
# Example 1: Broken link
feedback_message="链接失效:数据源 china-pbc 的 data_url 返回 404,无法访问数据页面。检索关键词:中国货币供应量"
# Example 2: Outdated content
feedback_message="数据内容过时:数据源 worldbank-open-data 的 update_frequency 标注为 quarterly,但实际已超过 6 个月未更新"When adding or modifying MCP tool descriptions, follow these principles (based on MCP tool description quality research):
Core principle: "Write it right before writing it all" — Functionality accuracy (+11.6% impact) matters ~8× more than Conciseness (+1.5%).
6-dimension checklist (check all before submitting):
FirstData is an open-source project — join us in building the External Facts Context Layer for AI Agents:
© MLT-OSS, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 2 other files (references) in skills/firstdata of MLT-OSS/FirstData.
Open the folder on GitHubat commit e4a6687
Firstdata next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Firstdata this skillMLT-OSS/FirstData | 183 | — | ~3.1k | Automated safety check: Pass | MIT | |
| Read GitHubAgentTeam-TaichuAI/ScienceClaw | 670 | 2 repos | ~638 | Automated safety check: Pass | None | |
| Report Issue Frameworkcyanheads/pubmed-mcp-server | 155 | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | |
| Serply Search MCPsickn33/agentic-awesome-skills | 47k | 1 repos | ~1.5k | Automated safety check: Pass | MIT | |
| Bgpt MCPClawBio/ClawBio | 1.2k | 1 repos | ~3.2k | Automated safety check: Pass | MIT | |
| Setup MedsciAperivue/medsci-skills | 329 | — | ~960 | Automated safety check: Pass | MIT |
AgentTeam-TaichuAI/ScienceClaw
Read and search GitHub repository documentation via gitmcp.io MCP service.
cyanheads/pubmed-mcp-server
File a bug or feature request against @cyanheads/mcp-ts-core when you hit a framework issue.
sickn33/agentic-awesome-skills
Search Google, Bing, Google News and Google Scholar, and read public pages, with the Serply MCP server.
ClawBio/ClawBio
Search scientific papers via the BGPT MCP server and retrieve structured experimental data — methods, results, conclusions, quality scores, and 25+ metadata fields per paper.
Aperivue/medsci-skills
A skill your agent uses when a skill fails for a missing tool or the environment needs checking.
brycewang-stanford/Auto-Empirical-Research-Skills
VS-Enhanced Journal Matcher with Journal Intelligence MCP — Real-time journal data pipeline with checkpoint-based human decisions.
Works with
Find official portals, APIs, and download paths for authoritative primary data sources (governments, international organizations, research institutions, etc.). Firstdata is an agent skill from MLT-OSS/FirstData.).
Firstdata fits situations like: users need to know where to find this data from an official source; which source is more authoritative; how to cite primary data.
Run `npx skills add MLT-OSS/FirstData --skill firstdata -a claude-code`. Or copy the skill folder (skills/firstdata in MLT-OSS/FirstData) into .claude/skills/firstdata in your project. Claude Code loads it when a task matches its description.
Run `npx skills add MLT-OSS/FirstData --skill firstdata -a codex`. Or copy the skill folder (skills/firstdata in MLT-OSS/FirstData) into .agents/skills/firstdata in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add MLT-OSS/FirstData --skill firstdata -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/firstdata, .gemini/skills/firstdata, .github/skills/firstdata and .opencode/skills/firstdata in your project.
Going by SKILL.md and its folder, Firstdata needs the command-line tools its instructions call (npx) and credentials named FIRSTDATA_API_KEY. Our summary lists: Node.js; A credential in FIRSTDATA_API_KEY.
SKILL.md names 2 domains. In commands or code: firstdata.deepminer.com.cn; the agent is likely to contact it when it follows the instructions. As links in the text: arxiv.org. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Firstdata is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.1k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.2k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Firstdata: Read GitHub (AgentTeam-TaichuAI/ScienceClaw, 670 stars), Report Issue Framework (cyanheads/pubmed-mcp-server, 155 stars), Serply Search MCP (sickn33/agentic-awesome-skills, 47k stars) and Bgpt MCP (ClawBio/ClawBio, 1.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
MLT-OSS (a GitHub organization) maintains it in MLT-OSS/FirstData, which has 183 GitHub stars. The repository was last updated on October 3, 2026.
Source: MLT-OSS/FirstData on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.