Llm Finetuning Eval Engineer
Evaluation gatekeeper for fine-tuning — builds golden sets and graders, calibrates judges, baselines base models, and issues checkpoint promotion verdicts. Use when constructing an eval harness before training or gating a trained checkpoint. Deliberately independent from training execution.
wshobson/agentsv10 stars · 0 forks · 0 makes≈2.4K tokens
Family tree
1 prompt · 1 version shown · 0 makes · Full screen
No forks or makes yet. When someone forks this prompt or shares something they made with it, it grows here.