Clean

    google-agents-cli-eval

    This skill should be used when the user wants to "run an evaluation", "evaluate my ADK agent", "write an eval dataset", "analyze eval failures", "compare eval results", "optimize agent", or needs guidance on the Agent Platform eval methodology and the Quality Flywheel. Covers eval metrics, dataset schema, LLM-as-judge scoring, and common failure causes. Do NOT use for API code patterns (use google-agents-cli-adk-code), deployment (use google-agents-cli-deploy), or project scaffolding (use google-agents-cli-scaffold).

    Publisher google
    Sourceskills.sh
    TypeDomain expertise
    Popularity55,284
    UpdatedJul 19, 2026, 10:01 AM
    Manifest Score96
    Lineage Score96
    Safety Score100
    Ecosystem Rank2864

    Security Findings

    Flagged by the hosting marketplace

    the skills.sh security audit flagged this skill (status=warn, risk=medium) — flagged by Snyk

    Versions

    4
    • 96
      Clean

      Scanned Jul 19, 2026, 10:01 AM

    • 100
      Clean

      Scanned Jun 1, 2026, 05:51 AM

    • 100
      Clean

      Scanned May 6, 2026, 06:38 PM

    • Clean

      Scanned Apr 28, 2026, 04:52 PM