Agent skill

Manage Reference Data

by Abilityai in Abilityai/cornelius

Upsert and maintain entity notes in a REFERENCE-kind scope (external facts/records), then reindex so they become connection/insight targets.

MITAuto-check: notes

Install Manage Reference Data

skills CLI
$ npx skills add Abilityai/cornelius --skill manage-reference-data -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Abilityai/cornelius manage-reference-data --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Abilityai/cornelius.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/manage-reference-data .claude/skills/manage-reference-data && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
manage-reference-data
GitHub stars
109
Token cost
~1.8k tokens
SKILL.md length
656 words
Files
1
Skills in repo
51
Repo updated
First seen
Licence
MIT

At a glance

Upsert and maintain entity notes in a REFERENCE-kind scope (external facts/records), then reindex so they become connection/insight targets.

  • Works in 5 steps: Resolve the scope and confirm it is… → Privacy (settled: private agent, no gate) → Upsert the entity note… → …
  • SKILL.md covers Step 0 — Resolve the scope and…, Step 1 — Privacy (settled:…, Step 2 — Upsert the entity… and Step 3 — Reindex (so the note…, plus 3 more sections
  • Calls python

What it does

Manage Reference Data is an agent skill from Abilityai/cornelius. Upsert and maintain entity notes in a REFERENCE-kind scope (external facts/records), then reindex so they become connection/insight targets. Generic across reference scopes; Company is the default. The scope's entity taxonomy (type/relationship) lives in that scope's own schema doc (Company's is COMPANY-BRAIN-SCHEMA.md). NEVER crystallizes, lifecycle-classifies, or trains on reference data.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

The repository describes itself as: AI-powered second brain template for Claude Code + Obsidian. The licence is MIT.

Example prompts

  • “s entity taxonomy (type/relationship) lives in that scope”
  • “/manage-reference-data”

Requirements

  • Python 3
  • Pre-approved tools (allowed-tools): Read, Write, Edit, Bash, Glob, Grep, AskUserQuestion

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Resolve the scope and confirm it is registered + reference-kind
  2. Privacy (settled: private agent, no gate)
  3. Upsert the entity note (overwrite-in-place)
  4. Reindex (so the note becomes a connection target)
  5. Mount for discovery (opt-in, per-invocation)

What it can do on your machine

Read from SKILL.md and the folder at commit fd5e9a4. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Write
    • Edit
    • Bash
    • Glob
    • Grep
    • AskUserQuestion

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Manage Reference Data loads about 1.8k tokens when it runs. Until then it costs about 104 tokens; SKILL.md has 656 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~104
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Read, Write, Edit, Bash, Glob, Grep, AskUserQuestion

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Abilityai/cornelius at commit fd5e9a4, republished under its MIT licence (© Abilityai). 656 words, ~1,755 tokens.

Download SKILL.mdSave it as .claude/skills/manage-reference-data/SKILL.md (or your agent's skills folder).
name
manage-reference-data
description
Upsert and maintain entity notes in a REFERENCE-kind scope (external facts/records), then reindex so they become connection/insight targets. Generic across reference scopes; Company is the default. The scope's entity taxonomy (type/relationship) lives in that scope's own schema doc (Company's is COMPANY-BRAIN-SCHEMA.md). NEVER crystallizes, lifecycle-classifies, or trains on reference data.
allowed-tools
Read, Write, Edit, Bash, Glob, Grep, AskUserQuestion
automation
manual
user-invocable
true
argument-hint
[add|update|supersede|list|mount-help] [scope=Company] [entity name]
metadata.version
1.0
metadata.created
2026-06-26
metadata.author
Cornelius

Manage Reference Data

Maintain a reference-kind scope — a non-core Brain/ layer of external facts/records the personal brain discovers bridges to, not a query store (a CRM owns queries). This skill upserts entity notes (overwrite-in-place), stamps the reference contract, and reindexes. It is generic: pass scope=<Folder> for any registered reference scope; Company is the default.

Read first: resources/layered-brains/REFERENCE-SCOPE-SCHEMA.md (the generic kind contract: frontmatter + linking + privacy) and the target scope's own schema doc for its type:/relationship: enum + structure (Company → COMPANY-BRAIN-SCHEMA.md, which also defines the family sub-scopes and the ref-* playbook suite this skill is the seed of). This skill enforces those docs; it does not restate them.

Hard rules (non-negotiable — these are what make it a reference scope):

  • Every note gets provenance: reference. Never originated/endorsed/encountered/ai-inferred.
  • Reference notes are never crystallized, lifecycle-classified, or used as a synthesis-pulse source.
  • Reference scopes are non-core → they never train q-values (the engine learn-gate already guarantees this; do not bypass it).
  • Default read scope is unchanged. Reference data surfaces only under an explicit BRAIN_READ_SCOPE=core,<Scope> mount.

Step 0 — Resolve the scope and confirm it is registered + reference-kind

bash
cd resources/local-brain-search
SCOPE="${1:-Company}"   # or parse scope=<Folder> from args
./venv/bin/python - "$SCOPE" <<'PY'
import sys, memory_config as mc
folder = sys.argv[1]
print("kind         :", mc.scope_kind(folder))
print("is_reference :", mc.is_reference_scope(folder))
print("is_core      :", folder in mc.CORE_FOLDERS)
PY
  • If is_reference is False: the scope is not registered as reference. Stop and either (a) add a ScopeDef(..., SCOPE_KIND_REFERENCE, False, (...)) to SCOPE_REGISTRY in memory_config.py, or (b) register_scope("<Folder>", kind="reference", slugs=(...)). See the schema doc → "Adding a new reference scope". Do not write entity notes into a non-reference folder.

Step 1 — Privacy (settled: private agent, no gate)

Cornelius is a fully private agent on a private remote, so reference data is committed/synced by default — no per-scope gate to clear. Just populate.

The only exception: if this specific scope must be host-local (e.g. third-party data you've agreed not to sync), add Brain/$SCOPE/ to .gitignore first. Default is to commit.

Step 2 — Upsert the entity note (overwrite-in-place)

One note = one entity. Path: Brain/<Scope>/<entity-name>.md for a flat scope; for a family scope, route into the correct child folder per that scope's schema doc (e.g. Brain/Company/people/<name>.md, Company/market/<name>.md — never loose at the family root). Apply the frontmatter contract exactly — type:/relationship: values come from the scope's schema doc:

yaml
---
type: <per the scope's schema doc — Company's enum is in COMPANY-BRAIN-SCHEMA.md>
relationship: <per the scope's schema doc, when the scope uses the axis>
provenance: reference
scope: <Scope>
status: active
as_of: <today YYYY-MM-DD>
updated: <today>
created: <today, or preserve existing>
created_by: <your model name>
updated_by: <your model name>
agent_version: 01.25
---

# <Entity display name>

<concise factual body>

## Relations
- <relation>: [[Other Entity]]

## Change Log
- <today> — <what changed and the source>
  • add: create the file. update: overwrite-in-place (re-stamp as_of:/updated:, preserve created/created_by). Reference data is mutable — overwriting is correct, not destructive.
  • supersede: set the old note status: superseded + superseded_by: "[[New Entity]]"; create the new note status: active. Keep both (history).
  • Relations are [[wiki-links]] (explicit edges, traversed by connection-finder). Use a contract/engagement note as the join for many-to-many (Client × Product × owner × dates × value).
Show full SKILL.md (258 more words)Show less

Step 3 — Reindex (so the note becomes a connection target)

bash
cd resources/local-brain-search && ./run_index.sh        # then, if the daemon is running:
./run_daemon.sh restart                                   # daemon is long-lived: RESTART, never reload

index_brain.py uses content-hash change detection: if nothing changed it skips; if any note changed it does a full rebuild (FAISS has no good incremental append). So an add/overwrite triggers a full reindex of the vault — expect minutes on the full Brain. After it completes, the reference note exists in the index as a non-core node — invisible to default (core) readers, surfaced only when mounted. Restart the daemon afterward so the long-lived process serves the new note (the wrapper's stale-guard only checks index mtime, not whether new scope tokens were added to the code).

Step 4 — Mount for discovery (opt-in, per-invocation)

To let the personal brain find bridges to these entities, run a discovery pass with the scope mounted:

bash
BRAIN_READ_SCOPE=core,<Scope> resources/local-brain-search/run_connections.sh --note "<a core note>"
# or auto-discovery / connection-finder with the same env var

Unset, the scope stays absent from every reader. The mounted (non-pure-core) read trains no q-values — verify data/q_values.json is byte-unchanged after a mounted discovery pass (existing learn-gate; this is the positive safety control).


Acceptance (matches Phase 6 in SCOPE-IMPLEMENTATION-PLAN.md)

  1. BRAIN_READ_SCOPE unset → reference notes absent from recall/search/connections; q_values.json byte-unchanged.
  2. =core,<Scope> → entities surface and link to core notes via wiki-links; the mounted read trains no q-values.
  3. Editing a note in place + reindex re-embeds it and drops the old content (content-hash diff).
  4. Reference notes never appear in --hubs / lifecycle / tension output (non-core, fingerprint-excluded), and are never crystallized.

What this skill must refuse

  • Writing into a non-reference (cognitive/core) folder.
  • Setting any provenance other than reference.
  • Any crystallize / graduate / lifecycle / synthesis-pulse action on a reference note (route those to cognitive scopes only).

© Abilityai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .claude/skills/manage-reference-data of Abilityai/cornelius.

Open the folder on GitHubat commit fd5e9a4

Compare with similar skills

Manage Reference Data next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Manage Reference Data compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Manage Reference Data this skillAbilityai/cornelius109—~1.8kAutomated safety check: NotesMIT
Wiki Maintaineropenclaw/openclaw392k1 repos~462Automated safety check: PassMIT
Obsidian Vault Maintaineropenclaw/openclaw392k2 repos~262Automated safety check: PassMIT
Openclaw PR Maintaineropenclaw/openclaw392k—~2.2kAutomated safety check: PassMIT
External Linksthedaviddias/Front-End-Checklist74k—~773Automated safety check: PassMIT
Open Source Maintainer Assistantslopus/happy24k—~1.9kAutomated safety check: PassMIT

Similar skills

  • Wiki Maintainer

    openclaw/openclaw

    Maintain the OpenClaw memory wiki vault with deterministic pages, managed blocks, and source-backed updates.

    392k GitHub starsUsed in 1 repo~462 tokens
    Knowledge ManagementAuto-check passed
  • Obsidian Vault Maintainer

    openclaw/openclaw

    Maintain an Obsidian-friendly memory wiki vault with wikilinks, frontmatter, and official Obsidian CLI awareness.

    392k GitHub starsUsed in 2 repos~262 tokens
    Knowledge ManagementAuto-check passed
  • Openclaw PR Maintainer

    openclaw/openclaw

    Review, triage, repair, or land OpenClaw issues and pull requests with current-source evidence and the native maintainer workflow.

    392k GitHub stars~2.2k tokensUpdated today
    DevelopmentAuto-check passed
  • External Links

    thedaviddias/Front-End-Checklist

    A skill your agent uses when auditing content pages for citation quality, suggesting authoritative sources to link for factual claims, or reviewing whether a page's external link attributes…

    74k GitHub stars~773 tokensUpdated yesterday
    Research & ScienceAuto-check passed
  • Helps maintain the slopus/happy open source project by triaging issues, drafting closing comments, finding duplicates and checking fixes, with approval before anything is posted.

    24k GitHub stars~1.9k tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Walks through step 2 of adding a syncable entity to the Twenty server: a cache service for flat entity maps and utilities that transform entities and DTOs into universal flat entities.

    58k GitHub stars~2.4k tokensUpdated today
    Backend & APIsAuto-check passed

More from Abilityai/cornelius

All 51 skills in this repo
  • Nano Banana Image Generator

    Abilityai/cornelius

    Generate images using Google's Nano Banana (Gemini 2.5 Flash Image).

    109 GitHub stars~1.2k tokensUpdated 15 days ago
    Auto-check: notes
  • Changelog Protocol

    Abilityai/cornelius

    Protocol for creating dated changelog files after significant agent sessions.

    109 GitHub stars~555 tokensUpdated 15 days ago
    Auto-check passed
  • Create Article

    Abilityai/cornelius

    Create long-form articles from knowledge base insights. An agent skill from Abilityai/cornelius.

    109 GitHub stars~2.1k tokensUpdated 15 days ago
    Auto-check: notes
  • Epistemic Classification

    Abilityai/cornelius

    Framework for distinguishing research findings from hypotheses and speculative synthesis.

    109 GitHub stars~1.5k tokensUpdated 15 days ago
    Auto-check passed
  • Get Youtube Transcript

    Abilityai/cornelius

    Extract the transcript from a YouTube video by URL or video ID.

    109 GitHub stars~525 tokensUpdated 15 days ago
    Auto-check: notes
  • Insight Capture Format

    Abilityai/cornelius

    Standard format for capturing and documenting insights in the knowledge base.

    109 GitHub stars~616 tokensUpdated 15 days ago
    Auto-check passed

Questions about Manage Reference Data

What does Manage Reference Data do?

Upsert and maintain entity notes in a REFERENCE-kind scope (external facts/records), then reindex so they become connection/insight targets. Manage Reference Data is an agent skill from Abilityai/cornelius. Upsert and maintain entity notes in a REFERENCE-kind scope (external facts/records), then reindex so they become connection/insight targets.

How do I install Manage Reference Data in Claude Code?

Run `npx skills add Abilityai/cornelius --skill manage-reference-data -a claude-code`. Or copy the skill folder (.claude/skills/manage-reference-data in Abilityai/cornelius) into .claude/skills/manage-reference-data in your project. Claude Code loads it when a task matches its description.

How do I install Manage Reference Data in Codex?

Run `npx skills add Abilityai/cornelius --skill manage-reference-data -a codex`. Or copy the skill folder (.claude/skills/manage-reference-data in Abilityai/cornelius) into .agents/skills/manage-reference-data in your project. Codex loads it when a task matches its description.

Can I use Manage Reference Data in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Abilityai/cornelius --skill manage-reference-data -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/manage-reference-data, .gemini/skills/manage-reference-data, .github/skills/manage-reference-data and .opencode/skills/manage-reference-data in your project.

What does Manage Reference Data need to run?

Going by SKILL.md and its folder, Manage Reference Data needs the command-line tools its instructions call (python). Our summary lists: Python 3. Its frontmatter pre-approves these tools: Read, Write, Edit, Bash, Glob, Grep, AskUserQuestion.

Does Manage Reference Data access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Manage Reference Data safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Manage Reference Data use?

Manage Reference Data is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Manage Reference Data use?

About 1.8k tokens (SKILL.md is roughly 7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Manage Reference Data?

Skills that share tags, products or a category with Manage Reference Data: Wiki Maintainer (openclaw/openclaw, 392k stars), Obsidian Vault Maintainer (openclaw/openclaw, 392k stars), Openclaw PR Maintainer (openclaw/openclaw, 392k stars) and External Links (thedaviddias/Front-End-Checklist, 74k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Manage Reference Data?

Abilityai (a GitHub organization) maintains it in Abilityai/cornelius, which has 109 GitHub stars. The repository holds 51 skills in this directory. The repository was last updated on September 22, 2026.

Source: Abilityai/cornelius on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.