Deepstream Sop
NVIDIA/skills
A skill your agent uses when building, deploying, evaluating, debugging, or measuring latency for the DeepStream SOP Inference Microservice — a GPU-accelerated FastAPI service that detects whether…
Creates and validates dubbo-go-pixiu LLM gateway conf.yaml. An agent skill from apache/dubbo-go-pixiu.
$ npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install apache/dubbo-go-pixiu pixiu-llm-gateway --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/apache/dubbo-go-pixiu.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/pixiu-llm-gateway .claude/skills/pixiu-llm-gateway && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "pixiu-llm-gateway" agent skill from https://github.com/apache/dubbo-go-pixiu/tree/develop/.agents/skills/pixiu-llm-gateway into .claude/skills/pixiu-llm-gateway/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pixiu-llm-gateway", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/apache/dubbo-go-pixiu/tree/develop/.agents/skills/pixiu-llm-gatewayType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install apache/dubbo-go-pixiu pixiu-llm-gateway --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apache/dubbo-go-pixiu.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/pixiu-llm-gateway .agents/skills/pixiu-llm-gateway && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "pixiu-llm-gateway" agent skill from https://github.com/apache/dubbo-go-pixiu/tree/develop/.agents/skills/pixiu-llm-gateway into .agents/skills/pixiu-llm-gateway/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pixiu-llm-gateway", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install apache/dubbo-go-pixiu pixiu-llm-gateway --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apache/dubbo-go-pixiu.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/pixiu-llm-gateway .cursor/skills/pixiu-llm-gateway && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "pixiu-llm-gateway" agent skill from https://github.com/apache/dubbo-go-pixiu/tree/develop/.agents/skills/pixiu-llm-gateway into .cursor/skills/pixiu-llm-gateway/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pixiu-llm-gateway", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/apache/dubbo-go-pixiu.git --path .agents/skills/pixiu-llm-gateway--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install apache/dubbo-go-pixiu pixiu-llm-gateway --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apache/dubbo-go-pixiu.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/pixiu-llm-gateway .gemini/skills/pixiu-llm-gateway && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "pixiu-llm-gateway" agent skill from https://github.com/apache/dubbo-go-pixiu/tree/develop/.agents/skills/pixiu-llm-gateway into .gemini/skills/pixiu-llm-gateway/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pixiu-llm-gateway", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install apache/dubbo-go-pixiu pixiu-llm-gatewayInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/apache/dubbo-go-pixiu.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/pixiu-llm-gateway .github/skills/pixiu-llm-gateway && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "pixiu-llm-gateway" agent skill from https://github.com/apache/dubbo-go-pixiu/tree/develop/.agents/skills/pixiu-llm-gateway into .github/skills/pixiu-llm-gateway/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pixiu-llm-gateway", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install apache/dubbo-go-pixiu pixiu-llm-gateway --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apache/dubbo-go-pixiu.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/pixiu-llm-gateway .opencode/skills/pixiu-llm-gateway && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "pixiu-llm-gateway" agent skill from https://github.com/apache/dubbo-go-pixiu/tree/develop/.agents/skills/pixiu-llm-gateway into .opencode/skills/pixiu-llm-gateway/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pixiu-llm-gateway", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
pixiu-llm-gatewayCreates and validates dubbo-go-pixiu LLM gateway conf.yaml. An agent skill from apache/dubbo-go-pixiu.
Pixiu LLM Gateway is an agent skill from apache/dubbo-go-pixiu. Creates and validates dubbo-go-pixiu LLM gateway conf.yaml. Use for LLM proxy/tokenizer/kvcache filters, llmmeta, vLLM, LMCache, retry/fallback, or Nacos LLM discovery. Do not use for MCP gateway config.
Its SKILL.md is about 2.5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in AI & LLM Engineering, covering Model routing and gateways and LLM inference and serving. It works with Model Context Protocol, vLLM, Apache Kafka and gRPC. The repository describes itself as: Based on the proxy gateway service of dubbo-go, it solves the problem that the external protocol calls the internal Dubbo cluster. At present, it supports HTTP and… The licence is Apache-2.0.
5 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit ba01888. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md (its code samples are yaml).
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Pixiu LLM Gateway loads about 2.5k tokens when it runs. Until then it costs about 56 tokens; SKILL.md has 559 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from apache/dubbo-go-pixiu at commit ba01888, republished under its Apache-2.0 licence (© apache). 559 words, ~2,526 tokens.
.claude/skills/pixiu-llm-gateway/SKILL.md (or your agent's skills folder).Generate a complete or embeddable LLM gateway conf.yaml that configures Pixiu as a multi-provider HTTP proxy.
Use when:
conf.yaml, including LLM proxy/tokenizer/kvcache, llm_meta, retry/fallback, vLLM/LMCache, or Nacos LLM discovery.Do not use for:
<HARD-GATE>
Do not generate YAML, write code, create files, or take any implementation action until the user has provided all required inputs. This is a first principle.
If any required input is missing, this turn must only ask for the missing fields in the current input group; do not generate examples, defaults, YAML, code, or final output.
Even if configuration information seems inferable, obvious, or implied by context, you must still ask the user to confirm it. Do not proceed until the user confirms it.
</HARD-GATE>
Choose upstream mode (required):
upstream_mode: static or registrylistener (required):
address.socket_address.address: 0.0.0.0address.socket_address.port: 8888route_config (required):
routes[].match.prefix: /v1routes[].route.cluster: llmdgp.filter.llm.proxy (required):
config.scheme: http (use https when the upstream is an HTTPS endpoint)config.timeout: 60sconfig.maxIdleConns: 100config.maxIdleConnsPerHost: 100config.maxConnsPerHost: 100Static upstream (required when upstream_mode: static):
endpoints[].IDendpoints[].socket_address.address or endpoints[].socket_address.domainsendpoints[].socket_address.portendpoints[].llm_meta.providerendpoints[].llm_meta.api_keyendpoints[].llm_meta.retry_policy.configclusters[].name: llmclusters[].lb_policy: RoundRobinendpoints[].llm_meta.retry_policy.name: NoRetryendpoints[].llm_meta.fallback: falseendpoints[].llm_meta.health_check_interval: 5000Registry upstream (required when upstream_mode: registry):
registries.nacos.addressregistries.nacos.groupregistries.nacos.namespaceregistries.nacos.usernameregistries.nacos.passwordadapters[].id: llm-registryadapters[].name: dgp.adapter.llmregistrycenterregistries.nacos.protocol: nacosregistries.nacos.timeout: 5sregistries.nacos.group: DEFAULT_GROUPdgp.filter.llm.tokenizer (optional):
config.log_to_console: falsedgp.filter.ai.kvcache (optional):
config.enabled: true (must be set explicitly when enabling kvcache)config.vllm_endpointconfig.lmcache_endpointconfig.default_modelconfig.token_cacheconfig.cache_strategyconfig.request_timeout: 2sconfig.lookup_routing_timeout: 50msconfig.hot_window: 5mconfig.hot_max_records: 300config.max_idle_conns: 100config.max_idle_conns_per_host: 100config.max_conns_per_host: 100config.retry.max_attempts: 3config.retry.base_backoff: 100msconfig.retry.max_backoff: 2sconfig.circuit_breaker.failure_threshold: 5config.circuit_breaker.recovery_timeout: 10sconfig.circuit_breaker.half_open_max_calls: 2Inputs group at a time. Each time, output only that group's required fields, with a short explanation after each field.pkg/common/constant/key.go.pkg/filter/llm/proxy/filter.go.pkg/filter/llm/tokenizer/tokenizer.go.pkg/filter/ai/kvcache/config.go and pkg/filter/ai/kvcache/handlers.go.pkg/model/llm.go, pkg/model/cluster.go, and pkg/model/base.go.upstream_mode: static uses static_resources.clusters[]; registry uses the LLM registry adapter.dgp.filter.ai.kvcache, tokenizer, and dgp.filter.llm.proxy as needed.socket_address and llm_meta; ensure the LLM cluster does not mix in ordinary HTTP endpoints.dgp.filter.ai.kvcache -> dgp.filter.llm.tokenizer -> dgp.filter.llm.proxy; ignore missing filters within this order chain.scheme is under dgp.filter.llm.proxy.config, not under an endpoint; socket_address.domains contains host names only, such as api.openai.com, not full URLs or paths.instance_id exactly matches the Pixiu endpoint ID.Complete LLM route example (conf.yaml):
static_resources:
listeners:
- name: net/http
protocol_type: HTTP
address:
socket_address:
address: 0.0.0.0
port: 8888
filter_chains:
filters:
- name: dgp.filter.httpconnectionmanager
config:
route_config:
routes:
- match:
prefix: /v1
route:
cluster: llm
http_filters:
- name: dgp.filter.ai.kvcache
config:
enabled: true
vllm_endpoint: "http://127.0.0.1:8000"
lmcache_endpoint: "http://127.0.0.1:9000"
default_model: "Qwen2.5-3B-Instruct"
request_timeout: "2s"
lookup_routing_timeout: "50ms"
hot_window: "5m"
hot_max_records: 300
hot_max_keys: 1000
max_idle_conns: 100
max_idle_conns_per_host: 100
max_conns_per_host: 100
token_cache:
enabled: true
max_size: 1024
ttl: "10m"
cache_strategy:
enable_compression: true
enable_pinning: true
enable_eviction: true
memory_threshold: 0.85
hot_content_threshold: 10
load_threshold: 0.7
pin_instance_id: "vllm-instance-1"
pin_location: "LocalCPUBackend"
compress_instance_id: "vllm-instance-1"
compress_location: "LocalCPUBackend"
compress_method: "zstd"
evict_instance_id: "vllm-instance-1"
circuit_breaker:
failure_threshold: 5
recovery_timeout: "10s"
half_open_max_calls: 2
retry:
max_attempts: 3
base_backoff: "100ms"
max_backoff: "2s"
- name: dgp.filter.llm.tokenizer
config:
log_to_console: false
- name: dgp.filter.llm.proxy
config:
scheme: http
timeout: "60s"
maxIdleConns: 100
maxIdleConnsPerHost: 100
maxConnsPerHost: 100
clusters:
- name: llm
lb_policy: RoundRobin
endpoints:
- ID: vllm-instance-1
socket_address:
address: 127.0.0.1
port: 8000
llm_meta:
provider: vllm
api_key: "<api-key>"
fallback: false
health_check_interval: 5000
retry_policy:
name: ExponentialBackoff
config:
times: 3
initialInterval: "200ms"
maxInterval: "5s"
multiplier: 2.0Nacos LLM registry mode example (conf.yaml):
static_resources:
listeners:
- name: net/http
protocol_type: HTTP
address:
socket_address:
address: 0.0.0.0
port: 8888
filter_chains:
filters:
- name: dgp.filter.httpconnectionmanager
config:
route_config:
routes:
- match:
prefix: /v1
route:
cluster: llm
http_filters:
- name: dgp.filter.llm.proxy
config:
scheme: http
timeout: "60s"
adapters:
- id: llm-nacos
name: dgp.adapter.llmregistrycenter
config:
registries:
nacos:
protocol: nacos
address: "127.0.0.1:8848"
timeout: "5s"
group: DEFAULT_GROUP
namespace: public© apache, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .agents/skills/pixiu-llm-gateway of apache/dubbo-go-pixiu.
Open the folder on GitHubat commit ba01888
Pixiu LLM Gateway next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Pixiu LLM Gateway this skillapache/dubbo-go-pixiu | 568 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | |
| Deepstream SopNVIDIA/skills | 3.5k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | |
| Litellmmagnus919/agent-skills | 113 | — | ~4.2k | Automated safety check: Notes | MIT | |
| Local LLM Routerhoodini/ai-agents-skills | 281 | — | ~20k | Automated safety check: Pass | None | |
| Eks Best Practicesaws-samples/appmod-blueprints | 113 | — | ~5k | Automated safety check: Pass | MIT-0 | |
| Aqua Model Lifecycleoracle/accelerated-data-science | 125 | — | ~1.4k | Automated safety check: Pass | UPL-1.0 |
NVIDIA/skills
A skill your agent uses when building, deploying, evaluating, debugging, or measuring latency for the DeepStream SOP Inference Microservice — a GPU-accelerated FastAPI service that detects whether…
magnus919/agent-skills
Operate, configure, secure, and troubleshoot the LiteLLM AI gateway (proxy) and Python SDK: run the proxy (litellm --config), route to 100+ providers through one OpenAI-compatible API, configure…
hoodini/ai-agents-skills
Route AI coding queries to local LLMs in air-gapped networks.
aws-samples/appmod-blueprints
Advisory guidance for Amazon EKS architecture and configuration decisions — compute strategy, networking, security, reliability, cost, autoscaling, observability, multi-tenancy, and upgrade planning.
oracle/accelerated-data-science
Register, list, get, and manage LLM models in OCI AI Quick Actions (AQUA) using the ADS SDK.
overmind-core/overmind
Deeper backend map of overbae — module layout, celery queue topology, the span-only tracing model and its API surface, capabilities and toml sync, behaviour-keyed scoring, auth and guests, model…
apache/dubbo-go-pixiu
Creates and debugs dubbo-go-pixiu HTTP-to-Dubbo route YAML. An agent skill from apache/dubbo-go-pixiu.
apache/dubbo-go-pixiu
Creates and validates dubbo-go-pixiu MCP gateway conf.yaml. An agent skill from apache/dubbo-go-pixiu.
apache/dubbo-go-pixiu
Creates and debugs dubbo-go-pixiu HTTP/Network filters. An agent skill from apache/dubbo-go-pixiu.
Categories
Creates and validates dubbo-go-pixiu LLM gateway conf.yaml. An agent skill from apache/dubbo-go-pixiu. Pixiu LLM Gateway is an agent skill from apache/dubbo-go-pixiu.yaml.
Pixiu LLM Gateway fits situations like: LLM proxy/tokenizer/kvcache filters; nacos LLM discovery; MCP gateway config.
Run `npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a claude-code`. Or copy the skill folder (.agents/skills/pixiu-llm-gateway in apache/dubbo-go-pixiu) into .claude/skills/pixiu-llm-gateway in your project. Claude Code loads it when a task matches its description.
Run `npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a codex`. Or copy the skill folder (.agents/skills/pixiu-llm-gateway in apache/dubbo-go-pixiu) into .agents/skills/pixiu-llm-gateway in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add apache/dubbo-go-pixiu --skill pixiu-llm-gateway -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/pixiu-llm-gateway, .gemini/skills/pixiu-llm-gateway, .github/skills/pixiu-llm-gateway and .opencode/skills/pixiu-llm-gateway in your project.
SKILL.md names no scripts, command-line tools or credentials: Pixiu LLM Gateway is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Pixiu LLM Gateway is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.5k tokens (SKILL.md is roughly 10k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Pixiu LLM Gateway: Deepstream Sop (NVIDIA/skills, 3.5k stars), Litellm (magnus919/agent-skills, 113 stars), Local LLM Router (hoodini/ai-agents-skills, 281 stars) and Eks Best Practices (aws-samples/appmod-blueprints, 113 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
apache (a GitHub organization) maintains it in apache/dubbo-go-pixiu, which has 568 GitHub stars. The repository holds 4 skills in this directory. The repository was last updated on October 3, 2026.
Source: apache/dubbo-go-pixiu on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.