Add Model
guoqingbao/xinfer
Adapt and port new LLM model architectures to this xinfer project.
A skill your agent uses when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or…
$ npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Mesh-LLM/mesh-llm llama-stage-patch-changes --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Mesh-LLM/mesh-llm.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/llama-stage-patch-changes .claude/skills/llama-stage-patch-changes && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "llama-stage-patch-changes" agent skill from https://github.com/Mesh-LLM/mesh-llm/tree/main/.agents/skills/llama-stage-patch-changes into .claude/skills/llama-stage-patch-changes/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "llama-stage-patch-changes", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Mesh-LLM/mesh-llm/tree/main/.agents/skills/llama-stage-patch-changesType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Mesh-LLM/mesh-llm llama-stage-patch-changes --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Mesh-LLM/mesh-llm.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/llama-stage-patch-changes .agents/skills/llama-stage-patch-changes && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "llama-stage-patch-changes" agent skill from https://github.com/Mesh-LLM/mesh-llm/tree/main/.agents/skills/llama-stage-patch-changes into .agents/skills/llama-stage-patch-changes/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "llama-stage-patch-changes", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Mesh-LLM/mesh-llm llama-stage-patch-changes --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Mesh-LLM/mesh-llm.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/llama-stage-patch-changes .cursor/skills/llama-stage-patch-changes && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "llama-stage-patch-changes" agent skill from https://github.com/Mesh-LLM/mesh-llm/tree/main/.agents/skills/llama-stage-patch-changes into .cursor/skills/llama-stage-patch-changes/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "llama-stage-patch-changes", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Mesh-LLM/mesh-llm.git --path .agents/skills/llama-stage-patch-changes--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Mesh-LLM/mesh-llm llama-stage-patch-changes --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Mesh-LLM/mesh-llm.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/llama-stage-patch-changes .gemini/skills/llama-stage-patch-changes && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "llama-stage-patch-changes" agent skill from https://github.com/Mesh-LLM/mesh-llm/tree/main/.agents/skills/llama-stage-patch-changes into .gemini/skills/llama-stage-patch-changes/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "llama-stage-patch-changes", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Mesh-LLM/mesh-llm llama-stage-patch-changesInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Mesh-LLM/mesh-llm.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/llama-stage-patch-changes .github/skills/llama-stage-patch-changes && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "llama-stage-patch-changes" agent skill from https://github.com/Mesh-LLM/mesh-llm/tree/main/.agents/skills/llama-stage-patch-changes into .github/skills/llama-stage-patch-changes/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "llama-stage-patch-changes", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Mesh-LLM/mesh-llm llama-stage-patch-changes --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Mesh-LLM/mesh-llm.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/llama-stage-patch-changes .opencode/skills/llama-stage-patch-changes && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "llama-stage-patch-changes" agent skill from https://github.com/Mesh-LLM/mesh-llm/tree/main/.agents/skills/llama-stage-patch-changes into .opencode/skills/llama-stage-patch-changes/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "llama-stage-patch-changes", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
llama-stage-patch-changesA skill your agent uses when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or…
Llama Stage Patch Changes is an agent skill from Mesh-LLM/mesh-llm. Use this skill when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or patch queue.
Its SKILL.md is about 2.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in AI & LLM Engineering, covering LLM inference and serving. It works with llama.cpp and Rust. The repository describes itself as: Distributed AI/LLM for the people. Share compute privately or publicly to power your agents and chat. The licence is Apache-2.0.
Read from SKILL.md and the folder at commit 48bf685. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
cargogitpython3From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Llama Stage Patch Changes loads about 2.6k tokens when it runs. Until then it costs about 57 tokens; SKILL.md has 1,178 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from Mesh-LLM/mesh-llm at commit 48bf685, republished under its Apache-2.0 licence (© Mesh-LLM). 1,178 words, ~2,611 tokens.
.claude/skills/llama-stage-patch-changes/SKILL.md (or your agent's skills folder).Use this skill when changing the Skippy staged-runtime ABI carried in
skippy/llama_cpp/patches.
skippy/llama_cpp/patches: top-level core patches first,
model_support/series second, and generated/series last.skippy/llama_cpp/upstream.txt..deps/llama.cpp as the final artifact; regenerate the
patch queue from commits.model_support/ patch. Keep
reusable staged-runtime machinery in the core lane and generated graph
annotations in the generated lane.include/skippy.h is an umbrella only. Put public C ABI declarations in
standalone include/skippy/<capability>.h headers.src/skippy/<capability>.cpp and private C++
declarations in narrowly named src/skippy/*.h headers.snake_case capability names. Keep exported symbols prefixed with
skippy_ and avoid generic helpers, utils, or expanded common modules.src/skippy.cpp is retired. Extend the owning capability module and keep new
implementation files below 1,000 lines.Treat Doxygen-style comments in include/skippy.h and
include/skippy/*.h as the source of truth for the public API reference.
Every public header and exported skippy_* function must have an adjacent
@brief describing what it is used for.
When the public header surface changes, prepare the patched checkout and regenerate the website reference before finishing the change:
scripts/prepare-llama.sh pinned
python3 scripts/generate-skippy-api-doc.py
python3 scripts/generate-skippy-api-doc.py --checkCommit mesh/website/src/docs/pages/skippy-api.md alongside the native queue
change. The generated page must not be hand-edited, and its inventory must
include every public header and exported function in the prepared checkout.
Every pull request that changes the Skippy ABI must include an explicit ABI inventory in the PR description. Do not describe a changed function signature as a newly added function.
The inventory must state, for each change:
Use this compact table in the PR description:
| Status | Symbol/declaration | Public header | Implementation / mirror | Reason | Lockstep update |
|---|---|---|---|---|---|
| Changed / Added / Removed | exact name and signature | include/skippy/<capability>.h | src/skippy/<capability>.cpp; Rust FFI path | behavior enabled | Rust mirror/callers updated; old ABI not supported |
For a changed function signature, call out that it is an ABI change even when
the symbol name is unchanged. List removed declarations explicitly as
“none” when no functions or fields were deleted; this prevents reviewers from
having to infer removals from a patch diff. Keep this inventory synchronized
with the ABI version constants in include/skippy/common.h and the mirrors in
skippy/crates/skippy-ffi/src/lib.rs. Do not add compatibility shims solely to
support an older native runtime; the acceptance criterion is a synchronized
Rust/native build and a clear version mismatch if the pieces are mixed.
Prepare the pinned checkout and current patch queue:
scripts/prepare-llama.sh pinnedFor llama-side editing, work in .deps/llama.cpp or another llama.cpp
checkout where commits can be named and inspected. Base the branch on the
pinned upstream, then carry core stage ABI commits, model-support commits, and
generated family commits in that order.
For an ordinary core capability change, emit one focused mail-format patch
after the current top-level core lane. Do not rewrite unrelated entries or put
the patch after model_support/ or generated shards:
repo_root="$(pwd)"
llama_checkout="${LLAMA_CHECKOUT:-$repo_root/.deps/llama.cpp}"
last_patch="$(find skippy/llama_cpp/patches -maxdepth 1 -type f -name '*.patch' | sort | tail -n 1)"
last_number="${last_patch##*/}"
last_number="${last_number%%-*}"
next_number=$((10#$last_number + 1))
git -C "$llama_checkout" format-patch -1 \
--start-number "$next_number" \
--output-directory "$repo_root/skippy/llama_cpp/patches" HEADFor a deliberate queue-boundary or source-layout change, rebuild the affected series from the pinned upstream instead. Create capability-owned commits in their intended order, place declarations and implementation in their final modules from the first patch that introduces them, and format the complete replacement series with contiguous numbering. Before replacing the durable queue, verify both of these invariants:
# The reconstructed commit series has exactly the intended final tree.
git diff --exit-code <authoritative-final-commit> <reconstructed-series-head>
# No patch defers the structural change to the end of the series.
git log --reverse --oneline <pinned-upstream>..<reconstructed-series-head>Move the old queue to an explicit temporary backup, generate the replacement
into a fresh skippy/llama_cpp/patches directory, and retain the backup
until clean application and native compilation pass. Never keep both series or
duplicate patch numbers in the durable directory.
Validate patch application in a clean checkout:
tmp_root="$(mktemp -d /tmp/mesh-llama.XXXXXX)"
trap 'rm -rf -- "$tmp_root"' EXIT
LLAMA_WORKDIR="$tmp_root/llama.cpp" scripts/prepare-llama.sh pinned
LLAMA_WORKDIR="$tmp_root/llama.cpp" \
MESH_LLM_LLAMA_BUILD_ROOT="$tmp_root/build" \
LLAMA_STAGE_BACKEND=cpu \
LLAMA_STAGE_LINK_MODE=static \
scripts/build-llama.shAdvancing skippy/llama_cpp/upstream.txt can silently invalidate a patch
that depends on upstream's ordering, not just its symbols. The queue still
applies, everything compiles, and the behavior is broken. This happened with
upstream 1269cb1, which moved check_tensor_dims ahead of buft_for_tensor
and left the stage tensor filter running too late; split serving was broken on
main because no test opened a real mid-stage artifact.
So on every re-pin, in addition to the checks above:
git log <old-pin>..<new-pin> -- src/llama-model-loader.* src/llama-model.*
for changes to load order, not just to signatures the patches touch.cargo test -p skippy-package-builder covers this via
mid_stage_artifact_opens_with_the_stage_filter_applied.SKIPPY_CORRECTNESS_MODEL; without it the test prints
skipping mid-stage: SKIPPY_CORRECTNESS_MODEL is not set and passes. Grep the
CI log for mid_stage_artifact_opens_with_the_stage_filter_applied ... ok, or
set the variable locally. A skipped gate reads identically to a pass.Compile each new public header once as C11 and once as C++17 with warnings treated as errors. For implementation moves, run the tests owned by the moved capability in addition to the Rust fallout checks below.
For Rust fallout, run cargo commands serially:
cargo fmt --all --check
cargo check -p mesh-llm
cargo test -p skippy-runtime --lib
cargo test -p skippy-serving --lib
cargo test -p mesh-llm --libPatch files are mail-format artifacts. Do not hand-normalize them in a way that
breaks git am.
© Mesh-LLM, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .agents/skills/llama-stage-patch-changes of Mesh-LLM/mesh-llm.
Open the folder on GitHubat commit 48bf685
Llama Stage Patch Changes next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Llama Stage Patch Changes this skillMesh-LLM/mesh-llm | 3.5k | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | |
| Add Modelguoqingbao/xinfer | 333 | — | ~4.2k | Automated safety check: Notes | MIT | |
| Test Modelguoqingbao/xinfer | 333 | — | ~2.6k | Automated safety check: Pass | MIT | |
| Aider DelegateamElnagdy/delegate-skills | 2.3k | 3 repos | ~3k | Automated safety check: Pass | MIT | |
| Qwen Mtp GgufR6410418/Jackrong-llm-finetuning-guide | 1.7k | — | ~1.7k | Automated safety check: Pass | MIT | |
| Quantizationvllm-project/vllm-omni | 7.1k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 |
guoqingbao/xinfer
Adapt and port new LLM model architectures to this xinfer project.
guoqingbao/xinfer
Test LLM models served by xinfer for correctness, output quality, and performance.
amElnagdy/delegate-skills
Delegate a coding task to Aider (aider) as a background implementer, then review its diff and land it yourself.
R6410418/Jackrong-llm-finetuning-guide
Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release.
vllm-project/vllm-omni
Work on vLLM-Omni quantization for diffusion, autoregressive, omni, or multi-stage models.
Blackwellboy/model-serving-minefield
Diagnose OpenAI-compatible model-serving failures from symptoms, endpoint reports, explicit configuration files, or logs while preserving evidence status and requiring confirm/refute checks.
Mesh-LLM/mesh-llm
A skill your agent uses when validating a MeshLLM release candidate or current HEAD against the last GitHub release, assembling the canonical feature/fix/modification inventory, testing locally…
Mesh-LLM/mesh-llm
A skill your agent uses when running, debugging, interpreting, or documenting mesh-llm benchmark tune model-serving throughput trials, including choosing…
Mesh-LLM/mesh-llm
A skill your agent uses when adding, renaming, removing, validating, or exposing mesh-llm config settings, including built-in settings, plugin config schemas, owner-control apply behavior, CLI…
Mesh-LLM/mesh-llm
A skill your agent uses when connecting agent tools or OpenAI clients to mesh-llm — launching or configuring Goose, Claude Code, OpenCode, Pi, curl, or any OpenAI-compatible client against a local…
Mesh-LLM/mesh-llm
A skill your agent uses when converting Hugging Face SafeTensors checkpoints into split BF16 GGUF model repos with skippy-quantize on Hugging Face Jobs or a local machine, then publishing the…
Mesh-LLM/mesh-llm
A skill your agent uses when creating, monitoring, validating, or documenting low-memory Hugging Face Jobs or local runs that quantize split BF16/FP16 GGUF model repos into custom quant GGUF repos…
Categories
A skill your agent uses when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or…. Llama Stage Patch Changes is an agent skill from Mesh-LLM/mesh-llm.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or patch queue.
Llama Stage Patch Changes fits situations like: changing mesh-llms patched llama.cpp Skippy ABI; model introspection; tensor filtering; activation-frame execution.
Run `npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a claude-code`. Or copy the skill folder (.agents/skills/llama-stage-patch-changes in Mesh-LLM/mesh-llm) into .claude/skills/llama-stage-patch-changes in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a codex`. Or copy the skill folder (.agents/skills/llama-stage-patch-changes in Mesh-LLM/mesh-llm) into .agents/skills/llama-stage-patch-changes in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Mesh-LLM/mesh-llm --skill llama-stage-patch-changes -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/llama-stage-patch-changes, .gemini/skills/llama-stage-patch-changes, .github/skills/llama-stage-patch-changes and .opencode/skills/llama-stage-patch-changes in your project.
Going by SKILL.md and its folder, Llama Stage Patch Changes needs the command-line tools its instructions call (cargo, git and python3). Our summary lists: Python 3.
SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Llama Stage Patch Changes is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.6k tokens (SKILL.md is roughly 10k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Llama Stage Patch Changes: Add Model (guoqingbao/xinfer, 333 stars), Test Model (guoqingbao/xinfer, 333 stars), Aider Delegate (amElnagdy/delegate-skills, 2.3k stars) and Qwen Mtp Gguf (R6410418/Jackrong-llm-finetuning-guide, 1.7k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Mesh-LLM (a GitHub organization) maintains it in Mesh-LLM/mesh-llm, which has 3,485 GitHub stars. The repository holds 25 skills in this directory. The repository was last updated on October 8, 2026.
Source: Mesh-LLM/mesh-llm on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.