MCP HubMCP Hub
SKILL·1E5980

retro-meta

Necmttn
更新日 27 days ago
3 閲覧
104
12
104
GitHubで表示
その他ai

について

レトロメタスキルは、過去の振り返りと現在の設定を深く遡及分析し、既存のパイプラインでは見落とされがちな改善点を特定します。特定のフレーズやコマンドによってトリガーされると、高度な思考能力を持つ外部AIエージェントを用いて推論を推進します。標準的な振り返りの後、提案が乏しい場合や、ヒューリスティックベースの提案を超えた広範な探索が必要な際にご利用ください。

クイックインストール

Claude Code

推奨
メイン
npx skills add Necmttn/ax -a claude-code
プラグインコマンド代替
/plugin add https://github.com/Necmttn/ax
Git クローン代替
git clone https://github.com/Necmttn/ax.git ~/.claude/skills/retro-meta

このコマンドをClaude Codeにコピー&ペーストしてスキルをインストールします

ドキュメント

ax:retro-meta - deep retro of retros

The companion to ax:retro. Where ax:retro walks the heuristic-derived proposals one by one, retro-meta asks: what improvements does the existing pipeline NOT yet see?

The external agent (this one, Claude Code or Codex with high thinking) drives the reasoning. The CLI just produces structured evidence and takes user-approved plans back.

When to fire

Explicit triggers only:

  • "let's do a deep retro" / "retro of retros"
  • "investigate my ax setup" / "what should I fix in my setup"
  • "review proposals the heuristic missed"
  • /ax:retro-meta slash command
  • After ax:retro finishes if the user wants broader exploration

Do NOT auto-trigger on generic "look at recent work".

Prerequisites

  • ax (axctl) is on PATH and the local SurrealDB is reachable. If ax doctor fails, stop and tell the user scripts/db-start.sh.
  • At least 3 retros in the last 30 days. Below that, evidence is too thin for a meta pass - recommend ax:retro first.

Workflow

Step 1 - Snapshot

ax retro meta --json --since=30 > /tmp/ax-meta.json

Read /tmp/ax-meta.json. The keys you care about:

  • experiment_status[] - read this FIRST (see Step 2). Each entry: experiment_id, proposal_dedupe_sig, proposal_title, proposal_form, artifact_path, days_since_accepted, opportunities_count, addressed_count, address_ratio, latest_checkpoint{kind,suggested,observed_at}, locked_verdict. Pending verdicts (locked_verdict=null) come first.
  • retros[] - raw tried/worked/failed/next per session.
  • patterns.tool_failures - sorted by total_count desc.
  • patterns.corrections - total + max-per-session + session_count.
  • patterns.friction_kinds - recurring kinds across sessions.
  • current_state.skills - what's already installed (do NOT propose duplicates).
  • current_state.open_proposals - existing heuristic proposals.
  • current_state.accepted_experiments - accepted but verdict-pending.
  • current_state.claude_md_user / claude_md_project - guidance file paths (null if absent).
  • investigation_prompts[] - the prompts you must walk.

Step 2 - Vet existing experiments FIRST

Walk experiment_status in order. For each entry with locked_verdict=null:

a. If latest_checkpoint.suggested is ignored or regressed: investigate why (read the artifact_path, sample the matching opportunities), then run ax improve verdict --set=<v> <proposal_dedupe_sig> to lock the call. b. If latest_checkpoint.suggested is adopted AND days_since_accepted > 30: lock it as adopted so it stops cluttering the open list: ax improve verdict --set=adopted <proposal_dedupe_sig>. c. If latest_checkpoint is null OR suggested is partial: leave open. Note in the final summary that it's still gathering signal.

A rule of thumb mirrored from investigation_prompts: if address_ratio < 0.1 after t+30, default to locking as ignored unless the artifact has an obvious "not yet exercised" reason.

Step 3 - Walk the investigation prompts (high thinking)

For EACH prompt in investigation_prompts:

  1. Inspect referenced state with Read / Glob / Grep:
    • skill files in ~/.claude/skills/ and ~/.agents/skills/
    • claude_md_user if non-null
    • claude_md_project if non-null
  2. Reason about a candidate improvement. Use a high thinking budget - the point is to see what the heuristic missed.
  3. If you identify a real improvement (NOT a duplicate of an existing skill or open_proposal): a. Draft a plan doc to ~/.claude/plans/<YYYY-MM-DD>-<slug>.md, 30–100 lines. Sections: Problem, Evidence (cite retro ids), Proposed change, Success signal. b. Show the user a 4–6 line summary. c. Ask explicitly: "Register this as an accepted experiment? (y/n)" d. ONLY on yes:
    ax retro plan \
      --slug=<kebab-slug> \
      --form=skill|hook|guidance|automation \
      --title="<short title>" \
      --hypothesis="<one sentence>" \
      --plan-path=~/.claude/plans/<file>.md \
      --evidence-retros=<retro:id1,retro:id2> \
      --confidence=low|medium|high
    
  4. If the prompt resolves to "no change needed" or "duplicate of existing", say so out loud and move on.

Step 4 - Optional: hand off to scaffolder

For each plan you registered, you may run:

ax improve accept --with-agent <dedupe_sig>

This spawns the internal scaffolding agent to draft an artifact (SKILL.md, hook script, etc) from the plan. Skip if the plan is already self-sufficient.

Step 5 - Summary

Print one paragraph:

  • N plans registered, M of those scaffolded
  • V verdicts locked (with kind, e.g. "2× ignored, 1× adopted")
  • K open_proposals reviewed (and their disposition)
  • Any prompts that resolved to "nothing here"
  • Suggested next retro window

Anti-patterns

  • NEVER register a plan without an explicit user yes per plan. The human is the final filter.
  • NEVER auto-accept all open_proposals - the heuristic surfaces them but the deep pass exists precisely to triage them by reasoning, not by frequency rank.
  • NEVER write directly to ~/.claude/skills/. Use ax retro plan + ax improve accept --with-agent.
  • NEVER skip Step 2's duplicate check. Proposing a Pre-Bash guard when one is already accepted just wastes the user's time.
  • Don't trust frequency alone. A frequency=1 retro can still be load-bearing if it represents a category Claude can't get right.
  • NEVER propose a new improvement that overlaps a pending experiment. Vet that one first - lock its verdict or escalate before piling on more proposals in the same area. The retrospective loop is incomplete if old experiments stay in limbo.

CLI reference

# Snapshot only (no side effects)
ax retro meta --since=30 [--limit-retros=50] [--pretty]

# Register a user-approved plan as accepted proposal + experiment
ax retro plan \
  --slug=<kebab> \
  --form=skill|hook|guidance|automation \
  --title="<title>" \
  --hypothesis="<hyp>" \
  --plan-path=<path-to-plan.md> \
  [--evidence-retros=retro:a,retro:b] \
  [--artifact-path=<path>] \
  [--confidence=low|medium|high] \
  [--frequency=<N>] \
  [--json]

# Optionally hand off scaffolding to the internal agent
ax improve accept --with-agent <dedupe_sig>

# Lock the verdict on a previously-accepted experiment
ax improve verdict --set=adopted|ignored|regressed|partial|no_longer_needed <dedupe_sig>

Output of ax retro meta defaults to JSON because the reader is you, not a human.

GitHub リポジトリ

Necmttn/ax
パス: skills/retro-meta
0
agent-memoryagent-observabilityai-agentsbunclaude-codecodex
FAQ

よくある質問

retro-meta Skillとは何ですか?

retro-meta はNecmttn が作成した Claude Skillです。Skillは、Claudeが必要に応じて読み込む指示とリソースをまとめ、追加の指示なしで retro-meta に関連するタスクを実行できるようにします。

retro-meta をインストールするには?

このページのインストールコマンドを使用してください。retro-meta をプラグインとして Claude Code に追加するか、リポジトリを skills ディレクトリにクローンし、Claudeを再起動してSkillを読み込みます。

retro-meta はどのカテゴリに属しますか?

retro-meta は その他 カテゴリに属します。

retro-meta は無料で利用できますか?

はい。retro-meta は AIMCP に掲載されており、無料でインストールできます。

関連スキル

llamaguard
その他

LlamaGuardは、暴力やヘイトスピーチなど6つの安全性カテゴリーにおいて、LLMの入力と出力をモデレートするMetaの70-80億パラメータモデルです。94〜95%の精度を提供し、vLLM、Hugging Face、Amazon SageMakerを使用してデプロイ可能です。このスキルを使用して、AIアプリケーションにコンテンツフィルタリングと安全策を簡単に統合できます。

スキルを見る
cost-optimization
その他

このClaudeスキルは、リソースの適正サイジング、タグ付け戦略、支出分析を通じて、開発者がクラウドコストを最適化することを支援します。AWS、Azure、GCPにわたるクラウド支出の削減とコストガバナンスの実施のためのフレームワークを提供します。インフラコストの分析、リソースの適正サイジング、または予算制約への対応が必要な際にご利用ください。

スキルを見る
sports-betting-analyzer
その他

このClaudeスキルは、スポーツベッティング市場(スプレッド、オーバー/アンダー、プロップベットなど)を分析し、過去の傾向や状況統計を検証することでバリューベットを特定します。教育目的のための実践的な提案を構造化されたマークダウン形式で出力します。開発者はスポーツベッティング分析ツールとして本機能を活用できますが、娯楽および教育目的に限定されている点に留意してください。

スキルを見る
quantizing-models-bitsandbytes
その他

このスキルは、bitsandbytesを使用してLLMを8ビットまたは4ビット精度に量子化し、精度の低下を最小限に抑えつつ50〜75%のメモリ削減を実現します。限られたGPUメモリでより大規模なモデルを実行したり、推論を高速化するのに理想的で、INT8、NF4、FP4などのフォーマットをサポートしています。HuggingFace Transformersと統合され、QLoRAトレーニングや8ビットオプティマイザーを可能にします。

スキルを見る