from-eval-runs-to-actions
Diagnose Lightsage eval runs and completed results, then turn failure evidence and execution diagnostics into prioritized fixes and validation plans. Use for failed or regressed evals, recurring eval reviews, trace or tool-call investigation, and deciding what to change after an eval. Do not use to design configurations or analyze prompt visibility runs.
Pinned to revision 3bb5bfc570e1, so it is the text this page describes rather than whatever the author pushed since.
Files
- skills/from-eval-runs-to-actions/SKILL.md
- skills/from-eval-runs-to-actions/agents/openai.yaml
- skills/from-eval-runs-to-actions/references/failure-taxonomy.md
Every link opens the file at its source, pinned to the revision this page describes.