Search

DeepSeek · LLM inference and serving

22 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~2.3kAutomated safety check: PassNo licence6 days ago
2

Adapt and port new LLM model architectures to this xinfer project.

guoqingbao/xinfer334—~4.2kAutomated safety check: NotesMIT1 mo ago
3

Use only for current stock/share prices, ticker quotes, and financial market movers (gainers, losers, most-traded shares).

zhongkaifu/TensorSharp568—~1kAutomated safety check: PassBSD-3-Clausetoday
4

用于在 cc-router 仓库新增一个 LLM provider(即在 src-tauri/providers/ 下添加 YAML 描述符并完成配套的同步改动)。当用户说「加 provider」「接入 XX 厂商」「新增订阅源」「provider YAML」「让 cc-router 支持 OpenRouter/Together/Groq/Ollama 之类」时必须触发本…

finch-xu/cc-router277—~1.6kAutomated safety check: PassMITyesterday
5

The user wants to connect an LLM or vision provider, already has an API key, asks "can I use OpenAI/Anthropic/Gemini/OpenRouter", wants local Ollama, or needs different cheap and strong models.

oxbshw/watch-skill470—~509Automated safety check: NotesMIT26 days ago
6

Interact with a Meshtastic LoRa mesh network through MESH-API — list nodes, read messages, send texts, and check connection status.

mr-tbot/mesh-api180—~1.8kAutomated safety check: PassGPL-3.02 mo ago
7

A skill your agent uses for web searches and current information lookups, finding sources, fact-checking, researching questions, comparing sources, or summarising web pages.

zhongkaifu/TensorSharp568—~2.3kAutomated safety check: WarnBSD-3-Clausetoday
8

Serves AI models on AMD Instinct GPU hardware using vLLM. An agent skill from amd/skills.

amd/skills408—~4kAutomated safety check: NotesMIT2 days ago
9

Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.2kAutomated safety check: PassMIT3 mo ago
10

Train Mixture of Experts (MoE) models using DeepSpeed or HuggingFace.

Orchestra-Research/AI-Research-SKILLs13k2 repos~3.7kAutomated safety check: PassMIT3 mo ago
11

Breaks LLM torch profiler traces down by forward pass, layer and kernel, with timing tables and Perfetto time ranges for the layers you want to inspect.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~3.9kAutomated safety check: PassNo licence6 days ago
12

A skill your agent uses when a new Ollama Cloud model is announced or available (e.g.

heypinchy/pinchy182—~3.9kAutomated safety check: NotesAGPL-3.020 days ago
13

Audit the DeepSeek V3 MPK demo + builder chain end-to-end and confirm logical equivalence with vLLM's reference implementation.

mirage-project/mirage2.5k—~4.6kAutomated safety check: PassApache-2.03 days ago
14

High-throughput LLM inference on NVIDIA GPUs. An agent skill from Luciole-Studio/Misaka-Agent.

Luciole-Studio/Misaka-Agent1711 repo~1.3kAutomated safety check: PassMIT3 days ago
15

Low-memory file2file quantization for very large safetensors LLMs that cannot be loaded whole.

amd/Quark182—~2.3kAutomated safety check: PassMIT13 days ago
16

Inspect a target model and prepare metadata for Quark PTQ planning.

amd/Quark182—~2kAutomated safety check: PassMIT13 days ago
17

vLLM Ascend plugin for LLM inference serving on Huawei Ascend NPU.

ascend-ai-coding/awesome-ascend-skills174—~2.7kAutomated safety check: PassNo licenceyesterday
18

Track daily PRs and Issues from vllm-project/vllm and vllm-project/vllm-ascend, filter by model (DeepSeek/Qwen/GLM/MiniMax/Kimi) and tech topics (PD disaggregation, MTP, quantization, graph mode…

ascend-ai-coding/awesome-ascend-skills174—~731Automated safety check: PassNo licenceyesterday
19

A skill your agent uses when calling open-weight LLMs on Together AI or Fireworks AI's OpenAI-compatible endpoints — baseurl plus namespaced model id, the cheapest model that clears the bar…

ericrisco/rsc-harness180—~3.3kAutomated safety check: PassMITyesterday
20

昇腾 NPU 平台 vLLM 大模型推理服务一键部署。触发:用户说'部署 模型名'、'NPU 部署模型'、'vllm serve'。流程:SSH检查 → NPU检查 → 配置发现(必须验证) → 用户确认 → 部署 → cron监控 → 验证。约束:(1) 配置必须从官方文档验证,禁止猜测;(2) 后台启动必须用cron监控,禁止手动轮询。支持…

ascend-ai-coding/awesome-ascend-skills174—~1.2kAutomated safety check: PassNo licenceyesterday
21

Select appropriate Ollama models for processing sensitive but legal content.

divinevideo/divine-mobile266—~1.3kAutomated safety check: PassMPL-2.0yesterday
22

A skill your agent uses when choosing an open-weight LLM and clearing it for use — which family and size fit the task, the hardware and the budget, and above all whether the license permits shipping.

ericrisco/rsc-harness180—~4.1kAutomated safety check: PassMITyesterday