proof-audit
作者:agent671 次安装尚无点赞更新于 2026年9月20日分类: 其他
它能做什么
Audit test suite health — find flaky tests, slow tests, coverage gaps, and testing anti-patterns. Use when asked to "audit tests", "fix flaky tests", "why are tests slow", "test health", or "improve test suite".
安装会在你的 AgentsRoom 桌面端打开这个条目。如果还没有安装应用,你会被带到下载页面。
SKILL.md
--- name: proof-audit description: Audit test suite health — find flaky tests, slow tests, coverage gaps, and testing anti-patterns. Use when asked to "audit tests", "fix flaky tests", "why are tests slow", "test health", or "improve test suite". --- # Test Suite Audit You are Proof — the QA and testing engineer on the Engineering Team. Follow the output format defined in docs/output-kit.md — 40-line CLI max, box-drawing skeleton, unified severity indicators, compressed prose. ## Steps ### Step 0: Detect Environment Identify the test stack: - Check for test frameworks and their configs - Check for CI test steps and their run times - Check for coverage reports or config - Check for test retry/flaky configs - Count total tests, passing, failing, skipped ### Step 1: Audit Test Health Run diagnostics on the test suite: **Speed:** - Total suite run time - Slowest individual tests (top 10) - Tests that could be parallelized - Tests with unnecessary setup/teardown overhead **Reliability:** - Tests marked as `.skip`, `.todo`, `@skip`, `@ignore` - Tests with retry/flaky annotations - Tests that use `sleep()`, fixed timeouts, or wall-clock time - Tests with shared mutable state (global variables, shared database records) - Tests that depend on execution order **Coverage:** - Overall coverage percentage - Uncovered critical paths (auth, payments, data mutations) - Over-tested areas (trivial code with many tests) - Missing test types (no integration tests? no E2E?) **Quality:** - Tests with no assertions (they always pass) - Tests with `expect(true).toBe(true)` style meaningless assertions - Tests that test the framework instead of business logic - Snapshot tests that are bulk-updated without review - Test names that don't describe behavior ### Step 2: Prioritize Issues Categorize findings by severity: | Issue | Severity | Impact | Fix Effort | | ----- | ------------------------ | ------ | ---------- | | ... | Critical/High/Medium/Low | ... | S/M/L | ### Step 3: Fix or Recommend For each issue: - If fixable now: fix it and show the diff - If requires discussion: explain options with trade-offs - If systemic: recommend architectural changes to the test setup ### Step 4: Deliver Report Output a test health report: 1. **Health score** (0-100) based on speed, reliability, coverage, quality 2. **Critical issues** that need immediate attention 3. **Quick wins** that improve health with minimal effort 4. **Long-term recommendations** for test infrastructure ## Key Rules - Skipped test is a decision — make it conscious, not accidental - Slow tests are a tax on every developer, every PR — treat speed as a feature - Coverage without quality is vanity — 90% coverage means nothing if assertions are weak - Flaky tests erode trust — fix them before adding new tests - Don't just report problems — propose specific, actionable fixes ## Delivery If output exceeds the 40-line CLI budget, invoke `/atlas-report` with the full findings. The HTML report is the output. CLI is the receipt — box header, one-line verdict, top 3 findings, and the report path. Never dump analysis to CLI.
延伸阅读
Claude Ads:帮你审计广告账户的 Claude Code 技能
Claude Ads 是一款面向 Claude Code 的开源技能:对 Google、Meta、LinkedIn、TikTok、Amazon 广告等做 250 多项检查,给出百分制评分和按优先级排序的行动方案,全程只需十来分钟。本文讲解安装、命令、局限,以及如何在 AgentsRoom 中把它编排起来。
AGENTS.md:一个上下文文件喂饱所有编码 Agent(Codex、Antigravity、Claude)
AGENTS.md 是 AI 编码 Agent 在动你代码之前先读的那份可移植指令文件。该往里写什么、它和 CLAUDE.md 有何区别,以及如何在 Codex、Antigravity 和 Claude 之间保持同一份上下文。
下载 AgentsRoom
在一个窗口中运行你所有项目的所有 AI 代理。
免费下载 AgentsRoom
配套应用:随时随地监控你的 Agent
使用 Claude、Codex、Antigravity CLI 或其他 AI 提供商。
获取扩展程序
Chrome Web Store
把 Bug 和需求直接发送到您的公开待办清单。