Search
DeepSeek · LLM inference and serving
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence. | BBuf/ | 938 | — | ~2.3k | Automated safety check: Pass | No licence | 6 days ago |
| 2 | Adapt and port new LLM model architectures to this xinfer project. | guoqingbao/ | 334 | — | ~4.2k | Automated safety check: Notes | MIT | 1 mo ago |
| 3 | Use only for current stock/share prices, ticker quotes, and financial market movers (gainers, losers, most-traded shares). | zhongkaifu/ | 568 | — | ~1k | Automated safety check: Pass | BSD-3-Clause | today |
| 4 | 用于在 cc-router 仓库新增一个 LLM provider(即在 src-tauri/providers/ 下添加 YAML 描述符并完成配套的同步改动)。当用户说「加 provider」「接入 XX 厂商」「新增订阅源」「provider YAML」「让 cc-router 支持 OpenRouter/Together/Groq/Ollama 之类」时必须触发本… | finch-xu/ | 277 | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 5 | The user wants to connect an LLM or vision provider, already has an API key, asks "can I use OpenAI/Anthropic/Gemini/OpenRouter", wants local Ollama, or needs different cheap and strong models. | oxbshw/ | 470 | — | ~509 | Automated safety check: Notes | MIT | 26 days ago |
| 6 | 6.Mesh API Interact with a Meshtastic LoRa mesh network through MESH-API — list nodes, read messages, send texts, and check connection status. | mr-tbot/ | 180 | — | ~1.8k | Automated safety check: Pass | GPL-3.0 | 2 mo ago |
| 7 | 7.Research A skill your agent uses for web searches and current information lookups, finding sources, fact-checking, researching questions, comparing sources, or summarising web pages. | zhongkaifu/ | 568 | — | ~2.3k | Automated safety check: Warn | BSD-3-Clause | today |
| 8 | Serves AI models on AMD Instinct GPU hardware using vLLM. An agent skill from amd/skills. | amd/ | 408 | — | ~4k | Automated safety check: Notes | MIT | 2 days ago |
| 9 | Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime. | Orchestra-Research/ | 13k | 2 repos | ~2.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 10 | 10.Moe Training Train Mixture of Experts (MoE) models using DeepSpeed or HuggingFace. | Orchestra-Research/ | 13k | 2 repos | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 11 | Breaks LLM torch profiler traces down by forward pass, layer and kernel, with timing tables and Perfetto time ranges for the layers you want to inspect. | BBuf/ | 938 | — | ~3.9k | Automated safety check: Pass | No licence | 6 days ago |
| 12 | A skill your agent uses when a new Ollama Cloud model is announced or available (e.g. | heypinchy/ | 182 | — | ~3.9k | Automated safety check: Notes | AGPL-3.0 | 20 days ago |
| 13 | Audit the DeepSeek V3 MPK demo + builder chain end-to-end and confirm logical equivalence with vLLM's reference implementation. | mirage-project/ | 2.5k | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 14 | 14.Tensorrt LLM High-throughput LLM inference on NVIDIA GPUs. An agent skill from Luciole-Studio/Misaka-Agent. | Luciole-Studio/ | 171 | 1 repo | ~1.3k | Automated safety check: Pass | MIT | 3 days ago |
| 15 | Low-memory file2file quantization for very large safetensors LLMs that cannot be loaded whole. | amd/ | 182 | — | ~2.3k | Automated safety check: Pass | MIT | 13 days ago |
| 16 | Inspect a target model and prepare metadata for Quark PTQ planning. | amd/ | 182 | — | ~2k | Automated safety check: Pass | MIT | 13 days ago |
| 17 | 17.Vllm Ascend vLLM Ascend plugin for LLM inference serving on Huawei Ascend NPU. | ascend-ai-coding/ | 174 | — | ~2.7k | Automated safety check: Pass | No licence | yesterday |
| 18 | Track daily PRs and Issues from vllm-project/vllm and vllm-project/vllm-ascend, filter by model (DeepSeek/Qwen/GLM/MiniMax/Kimi) and tech topics (PD disaggregation, MTP, quantization, graph mode… | ascend-ai-coding/ | 174 | — | ~731 | Automated safety check: Pass | No licence | yesterday |
| 19 | A skill your agent uses when calling open-weight LLMs on Together AI or Fireworks AI's OpenAI-compatible endpoints — baseurl plus namespaced model id, the cheapest model that clears the bar… | ericrisco/ | 180 | — | ~3.3k | Automated safety check: Pass | MIT | yesterday |
| 20 | 昇腾 NPU 平台 vLLM 大模型推理服务一键部署。触发:用户说'部署 模型名'、'NPU 部署模型'、'vllm serve'。流程:SSH检查 → NPU检查 → 配置发现(必须验证) → 用户确认 → 部署 → cron监控 → 验证。约束:(1) 配置必须从官方文档验证,禁止猜测;(2) 后台启动必须用cron监控,禁止手动轮询。支持… | ascend-ai-coding/ | 174 | — | ~1.2k | Automated safety check: Pass | No licence | yesterday |
| 21 | Select appropriate Ollama models for processing sensitive but legal content. | divinevideo/ | 266 | — | ~1.3k | Automated safety check: Pass | MPL-2.0 | yesterday |
| 22 | 22.Open Weights A skill your agent uses when choosing an open-weight LLM and clearing it for use — which family and size fit the task, the hardware and the budget, and above all whether the license permits shipping. | ericrisco/ | 180 | — | ~4.1k | Automated safety check: Pass | MIT | yesterday |