次のような場合に使用: エージェント(AIが自律的に処理を進行させるシステム)の品質や処理速度の測定・改善を行う際。評価ツールの設定、常時監視、CI/CD パイプライン(ソフトウェア開発の自動化)での品質チェック、可視化、コスト最適化などに対応します。 以下のキーワード入力で自動的に起動します: 「エージェントを評価したい」「評価ツールを追加したい」「品質を測定したい」「品質チェックを設定したい」「評価を実行したい」「エージェントが遅い」「なぜ遅いのか」「応答時間を短くしたい」「監視機能を設定したい」「CloudWatch ダッシュボード」「エージェントの実行コストはいくらか」「コスト最適化」「ログが表示されない」「ログが見つからない」「トレース情報が見つからない」「評価に失敗した」「評価エラー」「開発環境のトレース」「ローカル環境のトレース」「agentcore 開発トレース」「CloudWatch へのトレース送信」 **注:** エラーやシステムの停止原因を調査する場合は、このスキルではなく agents-debug を使用してください。処理は遅いが正常に動作するルートはこちらで対応し、動作していないルートはデバッグスキルで調査します。
Use when measuring or improving agent quality and performance — set up evaluators, online monitoring, CI/CD quality gates, observability, or cost optimization. Triggers on: "evaluate my agent", "add evaluator", "measure quality", "quality gate", "run evals", "agent too slow", "why is it slow", "reduce latency", "set up observability", "CloudWatch dashboard", "how much does my agent cost", "cost optimization", "logs not showing up", "logs missing", "spans not found", "eval failing", "eval error", "dev traces", "local traces", "agentcore dev traces", "traces to CloudWatch". Not for debugging errors or crashes — use agents-debug. Slow but correct routes here; broken routes to debug.
評価・モニタリング・オブザーバビリティを通じて、AgentCore agentの品質を測定・改善します。
使用しない場合:
agents-debug を使用agents-harden を使用$ARGUMENTS に指定できるもの:
agentcore --version を実行してください。このSkillにはv0.9.0以降が必要です。
agentcore/agentcore.json を読み込み、既存のEvaluator・オンライン評価設定・Agentのセットアップ内容を確認します。
agentcore/agentcore.json が見つからない場合:
「このSkillにはAgentCoreプロジェクトが必要です。
agents-get-startedを使用してプロジェクトを作成してください。」
| 開発者の意図 | アクション |
|---|---|
| 品質の測定、Evaluatorの追加、評価の実行、CI/CDゲート、オンラインモニタリング | references/evals.md を読み込み、そのワークフローに従う |
| オブザーバビリティ・CloudWatch・X-Ray・ログ・メトリクス・ダッシュボードの設定 | references/observability.md を読み込み、そのワークフローに従う |
| AgentCoreのコストの把握またはコスト削減 | references/cost.md を読み込む |
| 両方 — 「Agentを理解して改善したい」 | オブザーバビリティの設定から始め、その後Evalを追加する |
リファレンスファイルに完全な手順が記載されています。ステップに沿って進めてください。
agents-harden を提案するagents-debug を提案するagents-build を提案するワークフローによって異なります。具体的な出力内容については、読み込んだリファレンスを参照してください。
Measure and improve your AgentCore agent's quality through evaluation, monitoring, and observability.
Do NOT use for:
agents-debugagents-harden$ARGUMENTS can be:
Run agentcore --version. This skill requires v0.9.0 or later.
Read agentcore/agentcore.json to understand existing evaluators, online eval configs, and agent setup.
If agentcore/agentcore.json is not found:
"This skill requires an AgentCore project. Use
agents-get-startedto create one."
| Developer intent | Action |
|---|---|
| Measure quality, add evaluator, run eval, CI/CD gate, online monitoring | Load references/evals.md and follow its workflow |
| Set up observability, CloudWatch, X-Ray, logs, metrics, dashboards | Load references/observability.md and follow its workflow |
| Understand or reduce AgentCore costs | Load references/cost.md |
| Both — "I want to understand and improve my agent" | Start with observability setup, then add evals |
The reference file contains the full procedure. Follow it step by step.
agents-harden for production readinessagents-debug for root cause analysisagents-buildDepends on the workflow — see the loaded reference for specific outputs.
原文・著作権は Anthropic および各プラグイン作者に帰属します。日本語訳は Claude API による自動翻訳です。