meta-eval
agentscope-ai/OpenJudgeUse when the user wants to build an evaluation system for an LLM/agent application but doesn't know where to start — they have traces, prompts, RAG pipelines,…
Scores out of 100 · grade B+
2026-08-13Works
40% of the score92/100
- Frontmatter name `meta-eval` does not match its directory `00-meta-eval`.
Maintained
25% of the score92/100
- no commits in the last 12 weeks
Adopted
20% of the score40/100
- 787 stars on the source repo.
Documented
15% of the score81/100
- 1,445 words with worked examples.
Install
npx skills add agentscope-ai/OpenJudge/meta-evalWhat the check found
| Finding | What it means |
|---|---|
| Name mismatch | The frontmatter name and the directory name disagree. |
What it says it does
Use when the user wants to build an evaluation system for an LLM/agent application but doesn't know where to start — they have traces, prompts, RAG pipelines, or nothing at all. Also use when the user mentions evaluation, eval, benchmarking, testing LLM quality, measuring agent performance, assessing RAG accuracy, or wants to compare prompts/models. This skill is the entry router: it asks diagnostic questions then recommends which sub-skill (local workflow) to use next.
Also in agentscope-ai/OpenJudge
| Artifact | Score | What the check found | Type | Reach | Last commit |
|---|---|---|---|---|---|
| rl-rewardagentscope-ai/OpenJudge | clean | Skill | 787 stars | 10 days ago | |
| openjudgeagentscope-ai/OpenJudge | clean | Skill | 787 stars | 10 days ago | |
| paper-reviewagentscope-ai/OpenJudge | clean | Skill | 787 stars | 10 days ago | |
| metric-designagentscope-ai/OpenJudge | Name mismatch | Skill | 787 stars | 10 days ago | |
| bootstrapagentscope-ai/OpenJudge | Name mismatch | Skill | 787 stars | 10 days ago | |
| find-skills-comboagentscope-ai/OpenJudge | Missing files | Skill | 787 stars | 10 days ago |
Other ai & agents skills
Browse all| Artifact | Score | What the check found | Category | Reach | Last commit |
|---|---|---|---|---|---|
| mcp-builderanthropics/skills | No license | AI & agents | 104,949 installs | today | |
| agent-developmentanthropics/claude-plugins-official | clean | AI & agents | 5,901 installs | today | |
| plugin-settingsanthropics/claude-plugins-official | clean | AI & agents | 5,465 installs | today | |
| skill-creatoranthropics/claude-plugins-official | clean | AI & agents | 5,836 installs | today | |
| microsoft-foundrymicrosoft/azure-skills | clean | AI & agents | 544,819 installs | today | |
| claude-apianthropics/skills | No license | AI & agents | 59,236 installs | today |
Put this measurement in your README
A badge carrying how many listings this index holds from the repository and how many pass every static structural check. It reads from this index every time somebody loads your page, so it changes when the measurement changes and there is nothing to keep up to date. Free, no account, and the value is not something you or we can set by hand.
[](https://skillworks.kynth.studio/?q=agentscope-ai%2FOpenJudge)Would rather not hotlink us? Every badge is also served in shields.io’s endpoint schema, so shields renders the image and your readers never talk to our domain:
Published by Toolproof, the masthead over this index and eight others. The method behind the number is at toolproof.kynth.studio/methodology, and the whole thing is readable as JSON with no key at /api.
