Canary WatchSAFE
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
Overview
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
e9ca581a4f44OBSERVED · 2026-09-19What it tells the agent
The instruction file, verbatim from the audited commit — this is the text the model reads, and the surface the audit's instruction layer examines. Quoted here so you can judge it without cloning anything.
--- name: canary-watch description: このスキルを使用して、デプロイメント、マージ、または依存関係アップグレード後にデプロイされたURLの回帰を監視します。 origin: ECC --- # カナリアウォッチ — デプロイ後の監視 ## 使用時期 - 本番またはステージングへのデプロイ後 - 危険なPRをマージした後 - 修正が実際に修正されたことを確認したい場合 - ローンチウィンドウ中の継続的監視 - 依存関係アップグレード後 ## 動作方法 デプロイされたURLの回帰を監視します。停止されるか監視ウィンドウが期限切れになるまで、ループで実行されます。 ### 監視内容 ``` 1. HTTPステータス — ページは200を返していますか? 2. コンソールエラー — 以前なかった新しいエラーはありますか? 3. ネットワークの障害 — 失敗したAPIコール、5xx応答? 4. パフォーマンス — LCP/CLS/INPの回帰対ベースライン? 5. コンテンツ — 主要な要素は消えましたか?(h1、nav、footer、CTA) 6. API健康 — 重要なエンドポイントはSLA内で応答していますか? ``` ### 監視モード **クイックチェック**(デフォルト):シングルパス、レポート結果 ``` /canary-watch https://myapp.com ``` **継続監視**:N分ごとにM時間チェック ``` /canary-watch https://myapp.com --interval 5m --duration 2h ``` **差分モード**:ステージング対本番を比較 ``` /canary-watch --compare https://staging.myapp.com https://myapp.com ``` ### 警告しきい値 ```yaml critical: # 即座の警告 - HTTPステータス != 200 - コンソールエラー数 > 5(新しいエラーのみ) - LCP > 4s - APIエンドポイントは5xxを返す warning: # レポートで報告 - LCP ベースラインから > 500ms増加 - CLS > 0.1 - 新しいコンソール警告 - レスポンス時間 > 2xベースライン info: # ログのみ - マイナーパフォーマンス分散 - 新しいネットワークリクエスト(サードパーティスクリプトが追加された?) ``` ### 通知 重大なしきい値を超えたとき: - デスクトップ通知(macOS/Linux) - オプション:Slack/Discord Webhook - `~/.claude/canary-watch.log`にログ ## 出力 ```markdown ## Canary Report — myapp.com — 2026-03-23 03:15 PST ### Status - ✓ HTTP 200 - ✓ No critical errors - ✓ LCP within SLA (1.8s) ### Diffs from Baseline - CLS: 0.08 (↓ 0.02) - Response: 245ms (↑ 12ms, OK) - Network: 42 requests (↑ 3, investigate third-party?) ``` ## 統合 - `/benchmark`とペアリングしてパフォーマンス比較 - `/browser-qa`とペアリングして完全なUIテスト - CI/CDパイプラインに組み込んでオートメーション監視
Trust audit
SAFEgrade B · trust 89/100 Nothing in the source contradicts what it says it does. Grade A is reserved for packages that have also passed the behavioural sandbox.
| Layer | What it checks | Result |
|---|---|---|
| L0 | Provenance & inventory | PASS |
| L1 | Static analysis of the code | NA |
| L2 | Instruction surface (what it tells the agent) | PASS |
| L3 | Class-specific surface | PASS |
| L4 | Behavioural (sandbox) | SKIPPED |
What the source does
- Filesystem
- none-observed
- Network
- none-observed
- Shell
- none-observed
- Dependencies
- pinned
- Secrets in source
- none-found
Findings (0)
No findings outside the package's declared scope.
Gates applied: no_behavioural_pass.
e9ca581a4f44full audit observations/trust-audit/skill/affaan-m__canary-watch.json · Report an issue / request a re-scanAudit history
Every audit this skill has had.
| Date | Source | Verdict | Grade | Score | Change |
|---|---|---|---|---|---|
| 2026-09-19 | e9ca581a4f44 | SAFE | B | 89 | first audit |
Questions
What does the Canary Watch skill do?
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
Is Canary Watch safe to install?
The audit found nothing in the source that contradicts what it says it does, and graded it B (89/100). Grade A is held back for packages that have also passed a sandboxed behavioural run, which is why a clean skill reads B.
What can Canary Watch access on my machine?
The audit observed no filesystem, network or shell use at all in its source.
How current is this page?
The grade is for one exact copy of the source (e9ca581a4f44), read on 2026-09-19. The repository is watched, and a new audit runs when it changes — this is the first audit.