Incantory
Sign in
AgentMITNot scanned

Eval Judge

LLM judge for plugin quality assessment. Scores skills on triggering accuracy, orchestration fitness, output quality, and scope calibration using anchored rubrics.

wshobson/agentsv10 stars · 0 forks · 0 makes≈703 tokens

Scores by version

No eval runs reported yet. Incantory never runs prompts: owners report results from their own CI.

Datasets

No datasets.

All runs and datasets