Search

By BBuf

10 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~2.3kAutomated safety check: PassNo licence4 days ago
2

Reads SGLang or vLLM startup logs to show where GPU memory went and estimates how many concurrent requests fit at common token lengths.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~2.5kAutomated safety check: PassNo licence4 days ago
3

Analyzes Torch Profiler traces from SGLang, vLLM and TensorRT-LLM servers into kernel attribution, overlap and fusion tables.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~2.8kAutomated safety check: PassNo licence4 days ago
4

Looks up public original architecture diagrams for named LLM, vision-language, MoE, diffusion and OCR models and returns the image with its source attribution.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~1.2kAutomated safety check: PassNo licence4 days ago
5

Builds an operator-level compute template for an LLM and estimates FLOPs and MFU for a serving shape, with tensor shapes and parallelism what-if checks.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~4.5kAutomated safety check: PassNo licence4 days ago
6

Reviews SGLang changes the way its maintainers do, drawing on a bundled corpus of public PR review threads and a flowchart of how the diff runs.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~4.6kAutomated safety check: PassNo licence4 days ago
7

Adds verified layer guides such as L0 and L1 and compact GPU lanes to an existing Torch Profiler Chrome trace, changing how it looks but not how it ran.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~2kAutomated safety check: PassNo licence4 days ago
8

Breaks LLM torch profiler traces down by forward pass, layer and kernel, with timing tables and Perfetto time ranges for the layers you want to inspect.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~3.9kAutomated safety check: PassNo licence4 days ago
9

Compares SGLang, vLLM, TensorRT-LLM and TokenSpeed on one model and workload, searching server flags to find the best deployment command within a latency SLA.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~7.5kAutomated safety check: PassNo licence4 days ago
10

A skill your agent uses when an SGLang, vLLM, TensorRT-LLM, or TokenSpeed serving/model optimization task needs prior model-family PR evidence.

BBuf/AI-Infra-Auto-Driven-SKILLS925—~1.5kAutomated safety check: PassNo licence4 days ago