Skip to content

Explore by keyword

evaluation

9 plugins use this word, from 8 publishers.

  • aws-agentsbyaws

    Build, deploy, and operate AI agents on AWS. Skills for scaffolding agents with Amazon Bedrock AgentCore (Strands,…

  • agent-attack-replaybydasobral

    Build and replay controlled agent attack scenarios with measurable evidence.

  • arize-axbygithub

    Arize AX platform skills for LLM observability, evaluation, and optimization. Includes trace export, instrumentation,…

  • google-agents-clibygoogle

    Scaffold, develop, evaluate, and deploy AI agents with Google ADK. Bundles skills for the agent development lifecycle.

  • pi-llm-verifierbyoliveralerubio

    Read-only candidate verification with explicit discrete-judge and true-logprob modes.

  • orqbyorq-ai

    Agent skills for building, deploying, evaluating, and monitoring LLM pipelines on the orq.ai platform.

  • admit-benchbyrefiant-inc

    Admissibility gates for agents that act on physical plants. Check a proposed action against evidence, authority,…

  • google-agents-clibytechx-academy

    Scaffold, develop, evaluate, and deploy AI agents with Google ADK. Bundles skills for the agent development lifecycle.

  • phoenixbygithub

    Phoenix AI observability skills for LLM application debugging, evaluation, and tracing. Includes CLI debugging tools,…