Search
Prometheus · By sickn33
5 skills found.
Category:
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Set up and manage NVIDIA GPU servers for AI workloads. An agent skill from sickn33/agentic-awesome-skills. | sickn33/ | 47k | 2 repos | ~2k | Automated safety check: Notes | MIT | 2 days ago |
| 2 | Build AI-focused SRE incident response practices for LLM outages, degraded quality, runaway cost events, and safety regressions. | sickn33/ | 47k | 2 repos | ~3k | Automated safety check: Pass | MIT | 2 days ago |
| 3 | Set up alerting rules, configure on-call rotations, and manage incident response workflows. | sickn33/ | 47k | 1 repo | ~2.8k | Automated safety check: Pass | MIT | 2 days ago |
| 4 | Auto-scale LLM inference clusters on Kubernetes using KEDA, custom GPU metrics, and horizontal pod autoscaling. | sickn33/ | 47k | 1 repo | ~2.1k | Automated safety check: Pass | MIT | 2 days ago |
| 5 | Set up metrics collection and visualization with Prometheus and Grafana. | sickn33/ | 47k | 1 repo | ~2.7k | Automated safety check: Pass | MIT | 2 days ago |