Patterns and techniques for evaluating and improving AI agent outputs.
日本語の概要は準備中です。原文の説明を表示しています。
Comprehensive testing framework for skills
インストール方法を見るインストールする前に、エージェントに与えられる指示の中身を確認できます。
Validate skill functionality with automated tests.
skill/
├── SKILL.md
└── test-cases/
├── basic.yml
├── edge-cases.yml
└── integration.yml
name: Basic Functionality
description: Test core skill behavior
setup:
- mkdir -p /tmp/test-project
- cd /tmp/test-project
tests:
- name: Should load successfully
input: invoke skill
expected:
status: success
output_contains: 'Skill loaded'
- name: Should reject invalid input
input: invoke skill with "invalid"
expected:
status: error
error_contains: 'Invalid input'
teardown:
- rm -rf /tmp/test-project
dotnet-harness:test <skill> - Run all testsdotnet-harness:test <skill> --filter basic - Filter testsdotnet-harness:test --all - Test all skillsdotnet-harness:test --watch - Watch modestatus: success|erroroutput_contains: "string"output_matches: /regex/file_exists: "path"no_errors - Check stderr empty✓ skill-name
✓ basic functionality (12ms)
✓ edge cases (8ms)
✗ integration test (failed)
Expected: "success"
Got: "error: missing dependency"
Results: 2 passed, 1 failed
- name: Test Skills
run: |
dotnet-harness:test --all --format junit > results.xml
- name: Upload Results
uses: actions/upload-artifact@v4
with:
name: test-results
path: results.xml
Test fails: Check skill dependencies Timeout: Increase timeout in test config Flaky tests: Use deterministic inputs
Primary approach: Use Serena symbol operations for efficient code navigation:
serena_find_symbol instead of text searchserena_get_symbols_overview for file organizationserena_find_referencing_symbols for impact analysisserena_replace_symbol_body for clean modificationsWhen to use Serena vs traditional tools:
Example workflow:
# Instead of:
Read: src/Services/OrderService.cs
Grep: "public void ProcessOrder"
# Use:
serena_find_symbol: "OrderService/ProcessOrder"
serena_get_symbols_overview: "src/Services/OrderService.cs"
まだレビューはありません。使ってみた感想をお寄せください。
概要と使いどころ
Patterns and techniques for evaluating and improving AI agent outputs.
日本語の概要は準備中です。原文の説明を表示しています。
Comprehensive AI prompt engineering safety review and improvement prompt. Analyzes prompts for safety, bias, security vulnerabilities, and effectiveness while providing detailed improvement recommendations.
日本語の概要は準備中です。原文の説明を表示しています。
Use when user requests research requiring multiple sources, comprehensive analysis, or synthesis across topics - technical research, domain knowledge gathering, market analysis, or learning about complex subjects
日本語の概要は準備中です。原文の説明を表示しています。
AI-powered wiki generation for code repositories with commands, agents, and skills
日本語の概要は準備中です。原文の説明を表示しています。
Use when building .NET 10 or C# 14 applications; when using minimal APIs, modular monolith patterns, or feature folders; when implementing HTTP resilience, Options pattern, Channels, or validation; when seeing outdated patterns like old extension method syntax
日本語の概要は準備中です。原文の説明を表示しています。
Implements accessible .NET UI. SemanticProperties, ARIA, AutomationPeer, testing per platform.
日本語の概要は準備中です。原文の説明を表示しています。