本文へ移動
cccskills
無料GitHub で公開

agent-integration-testing

Use when the user requests integration testing, feature validation, or test plan execution

インストール方法を見る

含まれるファイル(2)

  • SKILL.md3.5 KB
  • README.md433 B

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

Agent Integration Testing

Overview

This skill guides the creation and autonomous execution of verifiable integration test specifications. It ensures that tests are actionable by agents, properly documented, and systematically executed by subagents to validate features or fix failures.

Core Process

  1. Investigate Codebase Area

    • By default, investigate the entire codebase to understand the context.
    • If the user specifies a feature area, use glob and grep to narrow the investigation.
  2. Write Test Specification

    • Create a test spec file at ./tests/<name>.md.
    • Use <name> = "integration" unless the user specifies a particular feature area.
    • The document MUST include a Prerequisites section at the top detailing any setup needed before tests become runnable (e.g., environment variables, database seeding, background services).
  3. Define Verifiable Tests

    • Each test must be written in plain English.
    • Include clear steps to reproduce.
    • Include a set of expectations.
    • CRITICAL: Every expectation must be strictly verifiable by an agent using available tools (e.g., shell commands, HTTP requests, reading file outputs). If a test cannot be verified by an agent, it is invalid and must be rewritten or removed.
  4. Execute Tests via Subagents

    • Spawn subagents (using the Task tool or @mention subagent system) to run each individual test.
    • The subagent must follow the prerequisites, execute the steps, and validate the outcomes against the expectations.
    • Collect the results (Pass/Fail and logs) from the subagents.
  5. Fix Failures (Optional)

    • If the user explicitly specifies that failures should be fixed, spawn another subagent (e.g., the SWE or BUILDER agent) to investigate and fix any noted failures.

Quick Reference

ActionPattern / Command
Test File Location./tests/<name>.md (default: integration.md)
PrerequisitesMust be documented at the top of the test file
Test FormatPlain English, Repro Steps, Verifiable Expectations
ExecutionSpawn one subagent per test or test suite
FixingSpawn SWE/BUILDER subagent if requested by user

Red Flags - STOP and Start Over

  • Unverifiable Tests: "Verify the UI looks nice" or "Check if the animation is smooth." (Agents cannot verify visual aesthetics without specific tools). Fix: Rewrite to check DOM elements, network responses, or file states.
  • Missing Prerequisites: Subagents failing because the server wasn't started. Fix: Ensure the prerequisite section explicitly defines the commands to start dependencies.
  • Executing Tests Manually: Running tests in the main conversation thread instead of spawning subagents. Fix: Dispatch parallel subagents for isolated execution.

Example Test Specification (./tests/auth-integration.md)

# Auth Integration Tests

## Prerequisites
- Start the test database: `docker compose up -d db`
- Run migrations: `npm run migrate`
- Start the server in background: `npm run start:test &`

## Test 1: User Registration
**Steps:**
1. Send a POST request to `/api/register` with payload `{"email": "test@example.com", "password": "pass"}`.

**Expectations:**
1. The HTTP response status must be `201 Created`.
2. A subsequent query to the database using `sqlite3 test.db "SELECT email FROM users WHERE email='test@example.com';"` must return the email.

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

anneal

無料

Use when the user wants to systematically fix AI code slop — duplicated logic, over-engineering, silent error swallowing, convention drift, cargo-cult patterns, and other LLM-introduced architectural decay — over a specified duration

日本語の概要は準備中です。原文の説明を表示しています。

av/skills192026年10月9日 更新

Produce a researched long-form article from a topic prompt via an orchestrated pipeline - research agent (first-person sources, working-definition gate), narrative-architecture outline, writer/cold-reviewer loop with an explicit ACCEPT/REVISE verdict contract, then a catalog-deslop pass with a regression gate. The orchestrator dispatches subagents only; the writer never judges its own draft. Use when the user says "article factory", "write an article about X", "run the article pipeline", or asks for a researched long-form piece produced end-to-end. For essays and micro posts in the user's own voice without a research stage, use the prose skill instead.

日本語の概要は準備中です。原文の説明を表示しています。

av/skills192026年10月9日 更新

Runs autonomous keep/discard experiments on a codebase to optimize a single metric for a fixed duration, in the style of karpathy/autoresearch. Use when the user says "autoresearch" (optionally with a focus, e.g. "autoresearch the optimizer"), asks to run experiments on a repo overnight, to hill-climb or optimize a metric autonomously, or points at a repo with a karpathy-style program.md.

日本語の概要は準備中です。原文の説明を表示しています。

av/skills192026年10月9日 更新

Create custom modules for [Harbor Boost](https://github.com/av/harbor/tree/main/boost), an optimizing LLM proxy. Use when building Python modules that intercept/transform LLM chat completions—reasoning chains, prompt injection, structured outputs, artifacts, or custom workflows. Triggers on requests to create Boost modules, extend LLM behavior via proxy, or implement chat completion middleware.

日本語の概要は準備中です。原文の説明を表示しています。

av/skills192026年10月9日 更新

bugbash

無料

Systematically explore and test any software project (CLI, API, Backend, Library, etc.) to find bugs, usability issues, and edge cases. Produces a structured report with full reproduction evidence (exact commands, inputs, logs, and tracebacks) for every issue.

日本語の概要は準備中です。原文の説明を表示しています。

av/skills192026年10月9日 更新

bughunt

無料

Fully autonomous bug hunting pipeline — discover bugs in a scoped area using parallel subagents, independently triage each finding, fix confirmed issues with subagents, then audit all fixes against repo constraints and target platforms. Runs end-to-end without user interaction.

日本語の概要は準備中です。原文の説明を表示しています。

av/skills192026年10月9日 更新

av のスキルをすべて見る

このスキルの問題を報告する