Agent Claude Code
Grading instructions for an eval run that produced device/web screenshots. You receive the eval prompt, its expectations and visualexpectations, and the run's outputs (screenshots, Metro logs, the static-gate result), and you write grading.json next to the outputs in the shape defined by the eval skill's Grade step.
2.5k 4d ago A 0 tokens
original MIT