Search
Playwright · LLM evaluation
5 skills found.
Category:
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Write LLM evaluation spec files with datasets, tasks, and evaluators using the @kbn/evals Playwright fixture. | elastic/ | 21k | — | ~2.3k | Automated safety check: Pass | Unknown | today |
| 2 | Scaffold a new LLM evaluation suite package with Playwright config, evaluate fixture, and package files. | elastic/ | 21k | — | ~1.7k | Automated safety check: Pass | Unknown | today |
| 3 | Write, extend, and debug PXI Playwright E2E tests for Phoenix. | Arize-ai/ | 12k | — | ~2.6k | Automated safety check: Pass | Unknown | yesterday |
| 4 | This skill should be used when a specific quality problem (UX, data, architecture, feature) needs systematic diagnosis and iterative fixing toward a defined target. | jacob-dietle/ | 111 | — | ~5.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 5 | Design and apply QA methodology for software teams: test strategy, regression testing, CI failure triage, test automation, quality gates and metrics, risk-based testing, exploratory testing, test… | magnus919/ | 115 | — | ~3.2k | Automated safety check: Pass | MIT | yesterday |