Search

By vipshop

7 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

High-level guide for integrating a new DiT model into cache-dit: Cache (BlockAdapter/ForwardPattern), Context Parallelism, Tensor Parallelism, Text Encoder Parallelism (TE-P), VAE Parallelism…

vipshop/cache-dit1.3k—~11kAutomated safety check: PassApache-2.0today
2

A skill your agent uses when writing, debugging, porting, reviewing, or optimizing CUDA C++ or PTX kernels; investigating CUDA Runtime or Driver API behavior; profiling kernels with Nsight Systems…

vipshop/cache-dit1.3k—~2.3kAutomated safety check: PassApache-2.0today
3

A skill your agent uses when writing, modifying, porting, or optimizing CuTe DSL GPU kernels in Python; reading CuTe DSL API reference material; integrating a CuTe DSL kernel into a project; or…

vipshop/cache-dit1.3k—~2.8kAutomated safety check: PassApache-2.0today
4

A skill your agent uses when writing, debugging, porting, reviewing, or optimizing CUTLASS or CuTe C++ kernels and templates; navigating CUTLASS examples, collectives, epilogues, pipelines, GEMM…

vipshop/cache-dit1.3k—~2.2kAutomated safety check: PassApache-2.0today
5

A skill your agent uses when doing operator migration or kernel migration for CUDA, Triton, or custom ops in cache-dit; porting kernels from nunchaku, deepcompressor, or other repos; designing…

vipshop/cache-dit1.3k—~3.8kAutomated safety check: PassApache-2.0today
6

A skill your agent uses when integrating a new PTQ workflow into cache-dit; designing quantize/load API shape, backend-specific config validation, save/load manifests, benchmark and regression…

vipshop/cache-dit1.3k—~2.8kAutomated safety check: PassApache-2.0today
7

Write optimized Triton GPU kernels for deep learning operations.

vipshop/cache-dit1.3k—~1.1kAutomated safety check: PassApache-2.0today