google-agents-cli-eval
This skill should be used when the user wants to "run an evaluation", "evaluate my ADK agent", "write an eval dataset", "analyze eval failures", "compare eval results", "optimize agent", or needs guidance on the Agent Platform eval methodology and the Quality Flywheel. Covers eval metrics, dataset schema, LLM-as-judge scoring, and common failure causes. Do NOT use for API code patterns (use google-agents-cli-adk-code), deployment (use google-agents-cli-deploy), or project scaffolding (use google-agents-cli-scaffold).
- Version
- 1.4.2
Pinned to revision ef7808f33fc3, so it is the text this page describes rather than whatever the author pushed since.
Files
- skills/google-agents-cli-eval/SKILL.md
- skills/google-agents-cli-eval/references/advanced-commands.md
- skills/google-agents-cli-eval/references/builtin-tools-eval.md
- skills/google-agents-cli-eval/references/dataset_schema.md
- skills/google-agents-cli-eval/references/metrics-guide.md
- skills/google-agents-cli-eval/references/multimodal-eval.md
- skills/google-agents-cli-eval/references/user-simulation.md
Every link opens the file at its source, pinned to the revision this page describes.