Clean
google-agents-cli-eval
This skill should be used when the user wants to "run an evaluation", "evaluate my ADK agent", "write an eval dataset", "analyze eval failures", "compare eval results", "optimize agent", or needs guidance on the Agent Platform eval methodology and the Quality Flywheel. Covers eval metrics, dataset schema, LLM-as-judge scoring, and common failure causes. Do NOT use for API code patterns (use google-agents-cli-adk-code), deployment (use google-agents-cli-deploy), or project scaffolding (use google-agents-cli-scaffold).
Publisher google
Sourceskills.sh
TypeDomain expertise
Popularity55,284
UpdatedJul 19, 2026, 10:01 AM
Manifest Score96
Lineage Score96
Safety Score100
Ecosystem Rank2864
Publisher google
Sourceskills.sh
TypeDomain expertise
Popularity55,284
UpdatedJul 19, 2026, 10:01 AM
Manifest Score96
Lineage Score96
Safety Score100
Ecosystem Rank2864
Security Findings
Flagged by the hosting marketplace
the skills.sh security audit flagged this skill (status=warn, risk=medium) — flagged by Snyk
Versions
4- 96Clean
Scanned Jul 19, 2026, 10:01 AM
- 100Clean
Scanned Jun 1, 2026, 05:51 AM
- 100Clean
Scanned May 6, 2026, 06:38 PM
- Clean
Scanned Apr 28, 2026, 04:52 PM