Topic · AI & LLM Engineering
Best LLM observability skills, page 3
LLM observability skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | Audit documentation gaps across the Phoenix repo by analyzing recent commits to main (default: last 7 days). | Arize-ai/ | 12k | — | ~4.8k | Automated safety check: Pass | Unknown | today |
| 98 | Create a new built-in classification evaluator for Phoenix evals. | Arize-ai/ | 12k | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | today |
| 99 | Generates onboarding code snippets for Phoenix tracing integrations and wires them into the project onboarding UI. | Arize-ai/ | 12k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 100 | 100.Phoenix LLMs Txt Maintain the Phoenix llms.txt documentation index at docs/phoenix/llms.txt — the machine-readable docs map used by AI agents and the px docs fetch CLI. | Arize-ai/ | 12k | — | ~2.3k | Automated safety check: Pass | Unknown | today |
| 101 | Write Playwright E2E tests for the Phoenix AI observability platform. | Arize-ai/ | 12k | — | ~3.6k | Automated safety check: Pass | Unknown | today |
| 102 | Writes and maintains the body of a pull request opened by a coding agent on Arize-ai/phoenix, including a fixed "How this PR was made" section that reports steering, abandoned approaches, and… | Arize-ai/ | 12k | — | ~1.3k | Automated safety check: Pass | Unknown | today |
| 103 | Write, extend, and debug PXI Playwright E2E tests for Phoenix. | Arize-ai/ | 12k | — | ~2.6k | Automated safety check: Pass | Unknown | today |
| 104 | Bump the next release-please version for a Phoenix Python package (arize-phoenix, arize-phoenix-client, arize-phoenix-evals, arize-phoenix-otel) by opening a PR with a Release-As commit footer. | Arize-ai/ | 12k | — | ~708 | Automated safety check: Pass | Apache-2.0 | today |
| 105 | Audit recent changes to Phoenix's user-facing surfaces (Python clients, TypeScript clients, CLI, REST/GraphQL APIs) and patch the three external-facing agent skills — phoenix-tracing, phoenix-cli… | Arize-ai/ | 12k | — | ~5.1k | Automated safety check: Pass | Unknown | today |
| 106 | 106.Phoenix Sqlean Maintaining packages/phoenix-sqlean, the vendored fork of nalgeon/sqlean.py published as arize-phoenix-sqlean. | Arize-ai/ | 12k | — | ~1.7k | Automated safety check: Pass | Unknown | today |
| 107 | TypeScript conventions and patterns for any TypeScript code in the Phoenix monorepo — including js/packages/, js/app/, and any other TS directories. | Arize-ai/ | 12k | — | ~1.5k | Automated safety check: Pass | Unknown | today |
| 108 | Maintain the bundled TypeScript package docs that ship inside Phoenix npm packages. | Arize-ai/ | 12k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | today |
| 109 | Migrate or upgrade TypeScript tooling in the Phoenix monorepo. | Arize-ai/ | 12k | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 110 | 110.Compound Docs Searchable Elixir/Phoenix/Ecto solution documentation system with YAML frontmatter. | oliver-kriska/ | 565 | — | ~547 | Automated safety check: Pass | MIT | 2 days ago |
| 111 | Answers questions about agent spend, token use, traces, events, errors and tool usage by running read-only SQL against a local TMA1 observability store. | tma1-ai/ | 119 | — | ~5.1k | Automated safety check: Notes | Apache-2.0 | 20 days ago |
| 112 | 112.Elixir Pro Write idiomatic Elixir code with OTP patterns, supervision trees, and Phoenix LiveView. | davila7/ | 32k | 8 repos | ~486 | Automated safety check: Pass | MIT | today |
| 113 | Debug LLM applications using the Phoenix CLI. An agent skill from github/awesome-copilot. | github/ | 40k | 2 repos | ~4k | Automated safety check: Pass | Apache-2.0 | today |
| 114 | Development guide for the @arizeai/phoenix-client TypeScript SDK — run and resume experiments, manage OpenTelemetry tracer providers with stack-based attach/detach, and write vitest unit and… | Arize-ai/ | 12k | — | ~370 | Automated safety check: Pass | Apache-2.0 | today |
| 115 | Find out what is going wrong in LLM or agent traffic by reading sampled Phoenix traces, spans, or sessions, writing free-form notes (open coding), then grouping the notes into a few narrow… | Arize-ai/ | 12k | — | ~6.4k | Automated safety check: Pass | Apache-2.0 | today |
| 116 | Create Phoenix release documentation grounded in actual code changes. | Arize-ai/ | 12k | — | ~6.7k | Automated safety check: Pass | Unknown | today |
| 117 | A skill your agent uses when you need to test or evaluate LangGraph/LangChain agents: writing unit or integration tests, generating test scaffolds, mocking LLM/tool behavior, running trajectory… | soba-labs/ | 107 | — | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 118 | Trace, evaluate, and deploy AI agents and LLM applications with LangSmith. | langchain-ai/ | 425 | — | ~935 | Automated safety check: Pass | MIT | today |
| 119 | 119.Deploy Elixir/Phoenix deployment patterns — Dockerfile, fly.toml, runtime.exs, mix release, rel/ overlays. | oliver-kriska/ | 565 | — | ~1.1k | Automated safety check: Pass | MIT | 2 days ago |
| 120 | 120.Evaluators Author or refine a Phoenix evaluator — code or LLM-as-a-judge — that scores a run's output. | Arize-ai/ | 12k | — | ~1.7k | Automated safety check: Pass | Unknown | today |
| 121 | 121.Experiments Run, read, and compare dataset-backed experiments to find evidence that a prompt or pipeline is improving. | Arize-ai/ | 12k | — | ~1.8k | Automated safety check: Pass | Unknown | today |
| 122 | Screenshot a running Phoenix feature and attach images to a GitHub PR. | Arize-ai/ | 12k | — | ~979 | Automated safety check: Pass | Unknown | today |
| 123 | 123.Playground Author, edit, or iterate on prompts in the Phoenix prompt playground, including running experiments over a dataset. | Arize-ai/ | 12k | — | ~3.9k | Automated safety check: Pass | Unknown | today |
| 124 | A skill your agent uses for ServiceRadar Elixir, Phoenix, LiveView, Ash/Ecto, migrations, external downloads, Hex dependencies, formatting, or Dialyzer work. | carverauto/ | 921 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | today |
| 125 | Deploy and operate production agent servers with LangSmith Deployment. | soba-labs/ | 107 | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 126 | Build and run evaluators for AI/LLM applications using Phoenix. | github/ | 40k | 3 repos | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 127 | 127.Phoenix Design Design system conventions for the Phoenix frontend — layout, dialogs, error display, BEM CSS class naming, and CSS design tokens. | Arize-ai/ | 12k | — | ~416 | Automated safety check: Pass | Apache-2.0 | today |
| 128 | Guide for the phoenix-otel TypeScript package — OTel registration, stack-based global provider management, and provider lifecycle. | Arize-ai/ | 12k | — | ~330 | Automated safety check: Pass | Apache-2.0 | today |
| 129 | 129.Phoenix REST API REST API development for Phoenix. An agent skill from Arize-ai/phoenix. | Arize-ai/ | 12k | — | ~286 | Automated safety check: Pass | Unknown | today |
| 130 | 130.Deps Update Bump outdated Hex deps — inventory, snapshot changelogs, update, fix breaks, split reviewable PRs (patches bundled, majors solo). | oliver-kriska/ | 565 | — | ~1.4k | Automated safety check: Pass | MIT | 2 days ago |
| 131 | 131.Datasets Understand what a Phoenix dataset is and reason well about its examples, outputs, splits, and how it feeds evaluators and experiments. | Arize-ai/ | 12k | — | ~1.6k | Automated safety check: Pass | Unknown | today |
| 132 | A skill your agent uses when extending phx.gen.auth — adding registration fields, custom user attributes, extra migrations alongside generated auth, fixture updates. | j-morgan6/ | 166 | — | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 133 | Monitor LLMs and agentic apps: performance, token/cost, response quality, and workflow orchestration. | aspectrr/ | 405 | — | ~858 | Automated safety check: Pass | MIT | 5 mo ago |
| 134 | 134.Phoenix Harbor Configure and interpret the Phoenix plugin for Harbor agent evaluations. | Arize-ai/ | 12k | — | ~3.4k | Automated safety check: Warn | Apache-2.0 | today |
| 135 | Fetch, organize, and analyze LangSmith traces for debugging and evaluation. | soba-labs/ | 107 | — | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 136 | 136.Yuv Reel Covers Generate unified, on-brand Instagram Reel covers for Yuval (YUV.AI Neon Phoenix system) — the signature look is a giant Hebrew headline BEHIND the subject cutout + a punch line IN FRONT (depth… | hoodini/ | 281 | — | ~1.1k | Automated safety check: Pass | No licence | 2 mo ago |
| 137 | Step-by-step guide for adding a new observability/tracing provider to Agent Kernel. | yaalalabs/ | 191 | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | today |
| 138 | A skill your agent uses when deciding who may do what — ownership checks, policy modules, scoped queries, role-based access in LiveViews and controllers. | j-morgan6/ | 166 | — | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 139 | Answer questions about LLM and agentic-application behavior from data already ingested into Elastic: latency and error rate, token and cost utilization, response quality and guardrail events, and… | elastic/ | 592 | — | ~4.9k | Automated safety check: Pass | Apache-2.0 | today |
| 140 | A skill your agent uses when building WebSocket features with Phoenix Channels — socket auth, join authorization, handlein/push/broadcast, Presence. | j-morgan6/ | 166 | — | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 141 | Adds Arize AX tracing to an LLM application for the first time. | github/ | 40k | — | ~6.2k | Automated safety check: Notes | MIT | today |
| 142 | INVOKE THIS SKILL when building, testing, or deploying Managed Deep Agents in LangSmith. | langchain-ai/ | 1.3k | — | ~8.7k | Automated safety check: Notes | MIT | 2 days ago |
| 143 | 143.Phoenix JSON API A skill your agent uses when building JSON API endpoints — :api pipeline, FallbackController, error rendering, pagination, versioning, Bearer auth. | j-morgan6/ | 166 | — | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 144 | Monitor and optimize LLM costs using Langfuse analytics and dashboards. | jeremylongshore/ | 2.8k | 1 repo | ~2.4k | Automated safety check: Pass | MIT | today |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents563
- Deep learning415
- Embeddings386
- LLM inference and serving372
- Prompt engineering360
- Retrieval-augmented generation358
- Fine-tuning309
- LLM evaluation308
- Speech recognition and synthesis308
- Structured output and tool calling276
- LLM cost and token optimization259
- LLM API integration255
- Model routing and gateways255
- LLM guardrails221
- Computer vision203
- Model hubs and datasets180
- GPU and accelerator computing176
- Diffusion and image models166
- Natural language processing131
- Reinforcement learning66
- AI interpretability23