evaluators
Arize-ai/phoenixAuthor or refine a Phoenix evaluator — code or LLM-as-a-judge — that scores a run's output. Trigger when the user wants to create a new evaluator, improve an…
Scores out of 100 · grade A
2026-08-17Works
40% of the score100/100
- Loads cleanly: valid frontmatter, required fields present, no dangling references.
Maintained
25% of the score94/100
- no commits in the last 12 weeks
- no license file
Adopted
20% of the score55/100
- 11,075 stars on the source repo.
Documented
15% of the score55/100
- No usage example or code block.
- 928-word body.
Install
npx skills add Arize-ai/phoenix/evaluatorsWhat the check found
| Finding | What it means |
|---|---|
| No license | The repository ships no license file, so the reuse terms are unclear. |
What it says it does
Author or refine a Phoenix evaluator — code or LLM-as-a-judge — that scores a run's output. Trigger when the user wants to create a new evaluator, improve an existing one's logic or rubric, choose labels, or decide what to measure on a dataset or experiment. Do NOT trigger on: (1) manual prompt drafting (use `playground`), (2) running or comparing experiments themselves (use `experiments`), (3) cross-trace failure diagnosis with no evaluator in scope (use `debug-trace`).
Also in Arize-ai/phoenix
| Artifact | Score | What the check found | Type | Reach | Last commit |
|---|---|---|---|---|---|
| phoenix-cliArize-ai/phoenix | No license | Skill | 1,295 installs | today | |
| phoenix-githubArize-ai/phoenix | No license | Skill | 11k stars | today | |
| pxi-eval-datasetArize-ai/phoenix | No license | Skill | 11k stars | today | |
| gh-stackArize-ai/phoenix | No license | Skill | 11k stars | today | |
| phoenix-docs-gap-auditArize-ai/phoenix | No license | Skill | 11k stars | today | |
| phoenix-evals-new-metricArize-ai/phoenix | No license | Skill | 11k stars | today |
Other ai & agents skills
Browse all| Artifact | Score | What the check found | Category | Reach | Last commit |
|---|---|---|---|---|---|
| mcp-builderanthropics/skills | No license | AI & agents | 104,949 installs | today | |
| agent-developmentanthropics/claude-plugins-official | clean | AI & agents | 5,901 installs | today | |
| plugin-settingsanthropics/claude-plugins-official | clean | AI & agents | 5,465 installs | today | |
| skill-creatoranthropics/claude-plugins-official | clean | AI & agents | 5,836 installs | today | |
| microsoft-foundrymicrosoft/azure-skills | clean | AI & agents | 544,819 installs | today | |
| claude-apianthropics/skills | No license | AI & agents | 59,236 installs | today |
Put this measurement in your README
A badge carrying how many listings this index holds from the repository and how many pass every static structural check. It reads from this index every time somebody loads your page, so it changes when the measurement changes and there is nothing to keep up to date. Free, no account, and the value is not something you or we can set by hand.
[](https://skillworks.kynth.studio/?q=Arize-ai%2Fphoenix)Would rather not hotlink us? Every badge is also served in shields.io’s endpoint schema, so shields renders the image and your readers never talk to our domain:
Published by Toolproof, the masthead over this index and eight others. The method behind the number is at toolproof.kynth.studio/methodology, and the whole thing is readable as JSON with no key at /api.
