Agent skill

Inference Precision Tensor Dump Compare

by ascend-ai-coding in ascend-ai-coding/awesome-ascend-skills

模型层 Tensor 打点与精度对比工具。用于在模型 forward 过程中捕获模型各层中间 tensor,实现 GPU/NPU 精度对比调试。支持 vLLM、SGLang 推理框架。When to use: When you need to debug precision issues between GPU and NPU,or validate layer-wise tensor…

No licenceAuto-check passedAI & LLM Engineering

Install Inference Precision Tensor Dump Compare

skills CLI
$ npx skills add ascend-ai-coding/awesome-ascend-skills --skill inference-precision-tensor-dump-compare -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ascend-ai-coding/awesome-ascend-skills inference-precision-tensor-dump-compare --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ascend-ai-coding/awesome-ascend-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/inference/inference-precision/inference-precision-tensor-dump-compare .claude/skills/inference-precision-tensor-dump-compare && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
inference-precision-tensor-dump-compare
GitHub stars
174
Token cost
~2.4k tokens
SKILL.md length
483 words
Files
10 (incl. scripts, references)
Skills in repo
69
Repo updated
First seen
Licence
None found

At a glance

模型层 Tensor 打点与精度对比工具。用于在模型 forward 过程中捕获模型各层中间 tensor,实现 GPU/NPU 精度对比调试。支持 vLLM、SGLang 推理框架。When to use: When you need to debug precision issues between GPU and NPU,or validate layer-wise tensor…

  • Tasks that involve LLM inference and serving
  • SKILL.md covers ⚠️ 必须环境变量, 框架选择, ⚠️ 工作流程 (必须遵循) and 通用环境变量, plus 7 more sections
  • Runs Python scripts from its folder; calls python

What it does

Inference Precision Tensor Dump Compare is an agent skill from ascend-ai-coding/awesome-ascend-skills. 模型层 Tensor 打点与精度对比工具。用于在模型 forward 过程中捕获模型各层中间 tensor,实现 GPU/NPU 精度对比调试。支持 vLLM、SGLang 推理框架。When to use: When you need to debug precision issues between GPU and NPU,or validate layer-wise tensor outputs during inference.

Its SKILL.md is about 2.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 11 other files, including scripts and reference files (for example `references/checklist.md`, `references/framework-integration.md` and `references/review-guide.md`).

It sits in AI & LLM Engineering, covering LLM inference and serving. It works with vLLM and SGLang. The repository describes itself as: A comprehensive knowledge base for Huawei Ascend NPU development, structured as distributed Agent Skills. https://ascend-ai-coding.github.io/awesome-ascend-skills/.

When your agent uses it

  • Tasks that involve LLM inference and serving

Example prompts

  • “/inference-precision-tensor-dump-compare”

Requirements

  • Python 3

What it can do on your machine

Read from SKILL.md and the folder at commit 62a4ecb. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 4 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Inference Precision Tensor Dump Compare loads about 2.4k tokens when it runs, and up to ~19k if it reads all its reference files. Until then it costs about 65 tokens; SKILL.md has 483 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~65
When it runs · the whole SKILL.md, loaded when a task matches
~2.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~19k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 483 words (~2,416 tokens).

“在模型 forward 过程中打点,捕获中间 tensor,用于 GPU/NPU 精度对比。”

— opening of SKILL.md by ascend-ai-coding
name
inference-precision-tensor-dump-compare
keywords
tensor dump, tensor打点, GPU NPU 对比, precision debugging, forward dump, 层tensor, vLLM, SGLang, 精度调试

Read the full SKILL.md on GitHub

Files

SKILL.md and 9 other files (scripts, references) in skills/inference/inference-precision/inference-precision-tensor-dump-compare of ascend-ai-coding/awesome-ascend-skills.

  • SKILL.md
  • references/checklist.md
  • references/framework-integration.md
  • references/review-guide.md
  • references/tensor-tags-guide.md
  • references/verification-debugging.md
  • scripts/analyze_logs.py
  • scripts/inject_tensor_dump.py
  • scripts/rollback_tensor_dump.py
  • scripts/verify_tags.py

Open the folder on GitHubat commit 62a4ecb

Compare with similar skills

Inference Precision Tensor Dump Compare next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Inference Precision Tensor Dump Compare compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Inference Precision Tensor Dump Compare this skillascend-ai-coding/awesome-ascend-skills174—~2.4kAutomated safety check: PassNone
SageMaker Serving Image Selectionhuggingface/skills11k1 repos~4.6kAutomated safety check: PassApache-2.0
Dstack Prototypingdstackai/dstack2.3k—~1.6kAutomated safety check: PassMPL-2.0
Debug InferenceNVIDIA/OpenShell15k—~1.9kAutomated safety check: PassApache-2.0
Graphsignalgraphsignal/graphsignal257—~6.2kAutomated safety check: PassApache-2.0
One EvalOpenDCAI/One-Eval165—~2.4kAutomated safety check: PassApache-2.0

Similar skills

  • Official

    Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones.

    11k GitHub starsUsed in 1 repo~4.6k tokens
    AI & LLM EngineeringAuto-check passed
  • Dstack Prototyping

    dstackai/dstack

    Use with the dstack skill for model-serving work when the image, serving command, resources, backend/fleet choice, or service behavior is not proven.

    2.3k GitHub stars~1.6k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Debug Inference

    NVIDIA/OpenShell

    Official

    Debug inference clients that use an attached provider and its native endpoint, including hosted APIs and host-local Ollama, vLLM, SGLang, TRT-LLM, LM Studio, or NIM.

    15k GitHub stars~1.9k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Graphsignal

    graphsignal/graphsignal

    Profile AI inference workloads (vLLM, SGLang, TensorRT-LLM, PyTorch, any GPU application) with the Graphsignal profiler and read the results from its local /signals JSON endpoint.

    257 GitHub stars~6.2k tokensUpdated 9 days ago
    AI & LLM EngineeringAuto-check passed
  • One Eval

    OpenDCAI/One-Eval

    驱动 One-Eval 对 API 或本地模型做端到端评测,覆盖纯文本、多模态、代码生成、函数调用和 Agent benchmark。当用户想评测模型在一个或多个 benchmark 上的表现、比较分数、补充 metric,或生成图文评测报告时使用本 skill。

    165 GitHub stars~2.4k tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed
  • LLM Torch Profiler Trace Analysis

    BBuf/AI-Infra-Auto-Driven-SKILLS

    Analyzes Torch Profiler traces from SGLang, vLLM and TensorRT-LLM servers into kernel attribution, overlap and fusion tables.

    900 GitHub stars~2.8k tokensUpdated 2 days ago
    AI & LLM EngineeringAuto-check passed

More from ascend-ai-coding/awesome-ascend-skills

All 69 skills in this repo
  • Ascend Dmi

    ascend-ai-coding/awesome-ascend-skills

    当用户需要对华为昇腾 NPU 进行硬件层面的管理、测试或诊断时使用此 skill。典型场景: - 查看 NPU 卡的状态、温度、利用率 - 测试内存带宽(h2d/d2h/d2d/p2p) - 跑算力/功耗基准测试(TFLOPS、TOPS) - 诊断 NPU 硬件故障或做健康检查 - 对 NPU 卡做压力测试(aicore、内存) - 复位/恢复卡住或异常的 NPU 卡 典型用户问题(即使不提…

    174 GitHub stars~1.7k tokensUpdated today
    Auto-check passed
  • Ascendc

    ascend-ai-coding/awesome-ascend-skills

    End-to-end AscendC custom operator development for Ascend NPU in an ascend-kernel (csrc/ops + build.sh + torchnpu PyTorch custom op) project.

    174 GitHub stars~3.5k tokensUpdated today
    Auto-check passed
  • Atc Model Converter

    ascend-ai-coding/awesome-ascend-skills

    Complete toolkit for Huawei Ascend NPU model conversion and end-to-end inference adaptation.

    174 GitHub stars~4.6k tokensUpdated today
    Auto-check passed
  • External Cannbot Ops Pypto Op Develop

    ascend-ai-coding/awesome-ascend-skills

    当需要编写 PyPTO 算子实现时使用此 skill。基于需求规格、设计方案和参考实现,生成完整可运行的 PyPTO 算子实现与配套测试、文档。Triggers: 实现算子、写 kernel、编写实现、写 impl、算子编码、开始编码、code the op、写 test、生成测试、写实现代码、op develop、kernel 实现。

    174 GitHub stars~2.1k tokensUpdated today
    Auto-check passed
  • External Gitcode Ascend Megatron Change Analyzer

    ascend-ai-coding/awesome-ascend-skills

    Analyze official Megatron-LM commits, PRs, and branch change sets to identify feature evolution, candidate breaking changes, and migration-relevant events.

    174 GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • External Gitcode Ascend Megatron Commit Tracker

    ascend-ai-coding/awesome-ascend-skills

    Track and normalize change requests against the official Megatron-LM repository by branch, PR, commit, commit range, or time window.

    174 GitHub stars~1.1k tokensUpdated today
    Auto-check passed

Works with

Questions about Inference Precision Tensor Dump Compare

What does Inference Precision Tensor Dump Compare do?

模型层 Tensor 打点与精度对比工具。用于在模型 forward 过程中捕获模型各层中间 tensor,实现 GPU/NPU 精度对比调试。支持 vLLM、SGLang 推理框架。When to use: When you need to debug precision issues between GPU and NPU,or validate layer-wise tensor…. Inference Precision Tensor Dump Compare is an agent skill from ascend-ai-coding/awesome-ascend-skills. 模型层 Tensor 打点与精度对比工具。用于在模型 forward 过程中捕获模型各层中间 tensor,实现 GPU/NPU 精度对比调试。支持 vLLM、SGLang 推理框架。When to use: When you need to debug precision issues between GPU and NPU,or validate layer-wise tensor outputs during inference.

When should I use Inference Precision Tensor Dump Compare?

Inference Precision Tensor Dump Compare fits situations like: tasks that involve LLM inference and serving.

How do I install Inference Precision Tensor Dump Compare in Claude Code?

Run `npx skills add ascend-ai-coding/awesome-ascend-skills --skill inference-precision-tensor-dump-compare -a claude-code`. Or copy the skill folder (skills/inference/inference-precision/inference-precision-tensor-dump-compare in ascend-ai-coding/awesome-ascend-skills) into .claude/skills/inference-precision-tensor-dump-compare in your project. Claude Code loads it when a task matches its description.

How do I install Inference Precision Tensor Dump Compare in Codex?

Run `npx skills add ascend-ai-coding/awesome-ascend-skills --skill inference-precision-tensor-dump-compare -a codex`. Or copy the skill folder (skills/inference/inference-precision/inference-precision-tensor-dump-compare in ascend-ai-coding/awesome-ascend-skills) into .agents/skills/inference-precision-tensor-dump-compare in your project. Codex loads it when a task matches its description.

Can I use Inference Precision Tensor Dump Compare in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ascend-ai-coding/awesome-ascend-skills --skill inference-precision-tensor-dump-compare -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/inference-precision-tensor-dump-compare, .gemini/skills/inference-precision-tensor-dump-compare, .github/skills/inference-precision-tensor-dump-compare and .opencode/skills/inference-precision-tensor-dump-compare in your project.

What does Inference Precision Tensor Dump Compare need to run?

Going by SKILL.md and its folder, Inference Precision Tensor Dump Compare needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Inference Precision Tensor Dump Compare access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Inference Precision Tensor Dump Compare safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Inference Precision Tensor Dump Compare use?

No licence was found for Inference Precision Tensor Dump Compare or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Inference Precision Tensor Dump Compare use?

About 2.4k tokens (SKILL.md is roughly 9.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 17k tokens, read only when the agent opens those files.

What are the alternatives to Inference Precision Tensor Dump Compare?

Skills that share tags, products or a category with Inference Precision Tensor Dump Compare: SageMaker Serving Image Selection (huggingface/skills, 11k stars), Dstack Prototyping (dstackai/dstack, 2.3k stars), Debug Inference (NVIDIA/OpenShell, 15k stars) and Graphsignal (graphsignal/graphsignal, 257 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Inference Precision Tensor Dump Compare?

ascend-ai-coding (a GitHub organization) maintains it in ascend-ai-coding/awesome-ascend-skills, which has 174 GitHub stars. The repository holds 69 skills in this directory. The repository was last updated on October 7, 2026.

Source: ascend-ai-coding/awesome-ascend-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.