Eval Judge
LLM judge for plugin quality assessment. Scores skills on triggering accuracy, orchestration fitness, output quality, and scope calibration using anchored rubrics.
wshobson/agentsv10 stars · 0 forks · 0 makes≈703 tokens
LLM judge for plugin quality assessment. Scores skills on triggering accuracy, orchestration fitness, output quality, and scope calibration using anchored rubrics.
wshobson/agentsv10 stars · 0 forks · 0 makes≈703 tokens
Community votes for v1. Votes count once an account is 7 days old, or has an approved example or a published make.
No votes yet.
Sign in to vote on which models this works with.
Real outputs people got from this prompt.
No examples yet. Ran this prompt? Share what you got.