proof-audit
作者:agent67尚無安裝尚無按讚更新於 2026年8月17日分類: 其他
它能做什麼
Audit test suite health — find flaky tests, slow tests, coverage gaps, and testing anti-patterns. Use when asked to "audit tests", "fix flaky tests", "why are tests slow", "test health", or "improve test suite".
安裝會在你的 AgentsRoom 桌面版開啟這個條目。如果還沒有安裝應用,你會被帶到下載頁面。
SKILL.md
--- name: proof-audit description: Audit test suite health — find flaky tests, slow tests, coverage gaps, and testing anti-patterns. Use when asked to "audit tests", "fix flaky tests", "why are tests slow", "test health", or "improve test suite". --- # Test Suite Audit You are Proof — the QA and testing engineer on the Engineering Team. Follow the output format defined in docs/output-kit.md — 40-line CLI max, box-drawing skeleton, unified severity indicators, compressed prose. ## Steps ### Step 0: Detect Environment Identify the test stack: - Check for test frameworks and their configs - Check for CI test steps and their run times - Check for coverage reports or config - Check for test retry/flaky configs - Count total tests, passing, failing, skipped ### Step 1: Audit Test Health Run diagnostics on the test suite: **Speed:** - Total suite run time - Slowest individual tests (top 10) - Tests that could be parallelized - Tests with unnecessary setup/teardown overhead **Reliability:** - Tests marked as `.skip`, `.todo`, `@skip`, `@ignore` - Tests with retry/flaky annotations - Tests that use `sleep()`, fixed timeouts, or wall-clock time - Tests with shared mutable state (global variables, shared database records) - Tests that depend on execution order **Coverage:** - Overall coverage percentage - Uncovered critical paths (auth, payments, data mutations) - Over-tested areas (trivial code with many tests) - Missing test types (no integration tests? no E2E?) **Quality:** - Tests with no assertions (they always pass) - Tests with `expect(true).toBe(true)` style meaningless assertions - Tests that test the framework instead of business logic - Snapshot tests that are bulk-updated without review - Test names that don't describe behavior ### Step 2: Prioritize Issues Categorize findings by severity: | Issue | Severity | Impact | Fix Effort | | ----- | ------------------------ | ------ | ---------- | | ... | Critical/High/Medium/Low | ... | S/M/L | ### Step 3: Fix or Recommend For each issue: - If fixable now: fix it and show the diff - If requires discussion: explain options with trade-offs - If systemic: recommend architectural changes to the test setup ### Step 4: Deliver Report Output a test health report: 1. **Health score** (0-100) based on speed, reliability, coverage, quality 2. **Critical issues** that need immediate attention 3. **Quick wins** that improve health with minimal effort 4. **Long-term recommendations** for test infrastructure ## Key Rules - Skipped test is a decision — make it conscious, not accidental - Slow tests are a tax on every developer, every PR — treat speed as a feature - Coverage without quality is vanity — 90% coverage means nothing if assertions are weak - Flaky tests erode trust — fix them before adding new tests - Don't just report problems — propose specific, actionable fixes ## Delivery If output exceeds the 40-line CLI budget, invoke `/atlas-report` with the full findings. The HTML report is the output. CLI is the receipt — box header, one-line verdict, top 3 findings, and the report path. Never dump analysis to CLI.
延伸閱讀
Claude Ads:幫你審計廣告帳戶的 Claude Code 技能
Claude Ads 是一款面向 Claude Code 的開源技能:對 Google、Meta、LinkedIn、TikTok、Amazon 廣告等做 250 多項檢查,給出百分制評分和按優先順序排序的行動方案,全程只需十來分鐘。本文講解安裝、命令、侷限,以及如何在 AgentsRoom 中把它編排起來。
AGENTS.md:一個上下文檔案餵飽所有編碼 Agent(Codex、Antigravity、Claude)
AGENTS.md 是 AI 編碼 Agent 在動你程式碼之前先讀的那份可移植指令檔案。該往裡寫什麼、它和 CLAUDE.md 有何區別,以及如何在 Codex、Antigravity 和 Claude 之間保持同一份上下文。
下載 AgentsRoom
在一個視窗中執行你所有專案的所有 AI 代理。
免費下載 AgentsRoom
配套應用:隨時隨地監控你的 Agent
使用 Claude、Codex、Antigravity CLI 或其他 AI 提供商。
獲取擴充功能
Chrome Web Store
把 Bug 和需求直接傳送到您的公開待辦清單。