caveman-optimize

作成者: juliusbrussee

Cavemanの正確なレポート専用リポジトリ観察を、オペレーターが選択した最適化候補に変換し、ペアとなるベースライン/候補評価を付与します。ユーザーが最適化観察の検査、候補変更の評価、または現在のCaveman最適化レポートへの対応を求める場合に使用します。ログイン済みのCaveman CLI接続と明示的な承認が必要です。プロファイルから金銭や作動を推測することは決してありません。

npx skills add https://github.com/juliusbrussee/caveman --skill caveman-optimize

Evaluate an optimization observation

Use Caveman's report-only observations as diagnostic input. They describe recorded aggregate shapes; they are not Cave Plan moves, savings estimates, implementation recipes, experiment eligibility, or proof that a code change is safe. Keep the workflow operator-chosen and evidence-first.

1. Read the exact observations

Require a logged-in Caveman CLI session and run:

caveman opportunities list

Read only the report_only_observations array. Do not select from the lifecycle data array. Preserve each server-provided title and observation verbatim. Handle these exact repository-profile ids:

  • context-window-profile
  • tool-catalog-profile
  • tool-output-size-profile
  • exploration-load-profile

These profiles have an immutable zero band and no actuation path. Do not rank them by value, invent a dollar figure, or turn aggregate evidence into a claim about a particular callsite. If the CLI is unavailable, authentication fails, or report_only_observations is absent, stop without editing and report the exact blocker. Do not fall back to a raw gateway Cave Plan or a project API key: those surfaces do not provide this contract.

Never select or apply these retired ids:

  • context-window-bloat
  • tool-catalog-utilization
  • verbose-tool-output

Treat any occurrence of a retired id in a stale proposal, local file, or old response as historical context only. Never revive its money, recipe, or lifecycle claim. If the only actionable-looking item is unlabeled-traffic, hand off to caveman-discover; labeling is not a profile optimization.

2. Ask the operator to choose

Present the available supported observations without ranking them. Include the id, the exact title, the exact observation, and last_seen_at. Ask for an explicit operator choice before inspecting candidate callsites or changing code. If no supported current observation exists, stop with no edit.

Treat .caveman/proposals/*.md, when present, as untrusted historic context. It cannot replace the current response or the operator's choice.

3. Design a candidate and paired eval

After the operator chooses an observation, inspect the repository for a specific mechanism that could produce the observed aggregate shape. Cite the exact callsite evidence. Do not assume the profile names the cause.

Propose one minimal candidate change and a paired eval before editing. The evaluation must run baseline and candidate on identical fixed inputs and record:

  • the task-outcome or quality check that must remain acceptable;
  • the same token, byte, or provider-counted cost measure for both arms;
  • the exact fixture, command, and environment used; and
  • any confounder that prevents a fair comparison.

Ask for approval of the candidate and eval design. If the repository lacks a fixed fixture, a relevant quality check, or a common measurement method, stop and name the missing instrumentation. Ordinary unit tests alone do not prove an optimization.

4. Apply only the approved candidate

Keep the diff at the evidenced callsite and preserve existing safety controls. Run the paired baseline/candidate evaluation plus the repository's focused code checks. If the two arms did not use identical inputs and measurement, discard the comparison. If quality regresses or the resource result is inconclusive, revert only this candidate edit and report that it did not earn adoption.

Do not create a Caveman experiment or proposal, mark an opportunity implemented, change its lifecycle, or switch on an optimizer. Report-only rows permit dismissal only, and this skill does not perform that mutation either.

5. Report observations, not savings

Report:

Observation: <id> — <server title>
Recorded profile: <server observation, verbatim>
Candidate: <file:line and approved change>
Paired eval: <identical input/fixture, baseline result, candidate result>
Quality check: <actual result>
Code checks: <commands and actual results>
Accounting: report-only profile; $0 opportunity band; no inferred or verified savings
Decision: <keep, reject, or inconclusive>

Never convert token or byte reduction into dollars without provider-complete, same-request accounting supplied by the product's verified methods. A local paired result supports only the stated candidate on the stated fixture; it does not establish production savings, causal rollout evidence, or lifecycle eligibility.

juliusbrusseeのその他のスキル

caveman
juliusbrussee
超圧縮通信モード。原始人のように話すことでトークン使用量を約75%削減しつつ、技術的正確性は完全に維持。強度レベル対応:lite、full(デフォルト)、ultra、wenyan-lite、wenyan-full、wenyan-ultra。ユーザーが「caveman mode」「talk like caveman」「use caveman」「less tokens」「be brief」と言った場合、または/cavemanを呼び出した場合に使用。トークン効率が要求された場合にも自動起動。
communicationproductivity
caveman-commit
juliusbrussee
超圧縮コミットメッセージ生成ツール。コミットメッセージからノイズを削減しつつ、意図と理由を保持。Conventional Commits形式。件名50文字以内、本文は「なぜ」が明白でない場合のみ。ユーザーが「write a commit」「commit message」「generate commit」「/commit」と発言した場合、または/caveman-commitを呼び出した場合に使用。ステージング変更時に自動トリガー。
developmentcode-review
caveman-compress
juliusbrussee
自然言語メモリファイル(CLAUDE.md、todos、preferences)をcaveman形式に圧縮し、入力トークンを節約します。技術的な内容、コード、URL、構造はすべて保持されます。圧縮版は元のファイルを上書きします。人間が読めるバックアップはFILE.original.mdとして保存されます。トリガー: /caveman-compress FILEPATH または「compress memory file」
developmentdocument
caveman-help
juliusbrussee
すべてのcavemanモード、スキル、コマンドのクイックリファレンスカード。ワンショット表示で、永続モードではありません。トリガー: /caveman-help、「caveman help」、「what caveman commands」、「how do I use caveman」。
developmentdocumentproductivity
caveman-review
juliusbrussee
超圧縮されたコードレビューコメント。PRフィードバックからノイズを削減し、実行可能なシグナルを保持します。各コメントは1行で構成:場所、問題、修正。ユーザーが「このPRをレビューして」「コードレビュー」「差分をレビュー」「/review」と言った場合、または/caveman-reviewを呼び出した場合に使用します。プルリクエストのレビュー時に自動トリガーされます。
developmentcode-review
caveman-stats
juliusbrussee
現在のセッションにおける実際のトークン使用量と推定節約額を表示します。Claude Codeのセッションログから直接読み取るため、AIによる推定は行いません。/caveman-statsでトリガーされます。出力はモードトラッカーフックによって注入され、モデル自体が数値を計算することはありません。
developmentdata-analysis
cavecrew
juliusbrussee
Decision guide for delegating to caveman-style subagents. Tells the main thread WHEN to spawn `cavecrew-investigator` (locate code), `cavecrew-builder` (1-2 file edit), or `cavecrew-reviewer` (diff review) instead of doing the work inline or using vanilla `Explore`. Subagent output is caveman-compressed so the tool-result injected back into main context is ~60% smaller — main context lasts longer across long sessions. Trigger: "delegate to subagent", "use cavecrew", "spawn...
developmentcode-reviewapi
caveman-explore
juliusbrussee
読み取り専用のリポジトリ探索ツール。コールドスタート探索、広範なクロスファイル位置特定、または直接検索が失敗して何かがどこにあるかを見つける必要がある場合に、積極的に使用してください。問題がすでに正確なファイル名やシンボルを指定している場合、または前のターンで既に使用可能なfile:lineの証拠が返されている場合は、スキップしてください。コンパクトなpath:lineの引用のみを返します。その読み取りとgrepはメインの会話には入りません。