Designs a comprehensive testing approach to validate detection effectiveness. Produces positive and negative test procedures, an evaluation metrics framework, an Atomic Red Team test plan, and an iterative tuning strategy. Use after detection logic is designed to plan how to verify the detection fires correctly and doesn't produce excessive false positives.
Inputs:
Workflow steps:
Design positive validation tests - Create malicious behavior simulation plan:
Design negative validation tests - Create benign behavior test cases:
Plan historical data analysis - If production data available:
Define evaluation metrics - Establish measurement criteria:
Create test execution matrix - Document test scenarios:
| Test ID | Type | Procedure | Expected Result | Pass/Fail Criteria |
Plan iterative tuning - Define refinement process:
Outputs: Validation plan document containing:
References:
Search for places (restaurants, cafes, etc.) via Google Places API proxy on localhost.
Interact with GitHub using the `gh` CLI. Use `gh issue`, `gh pr`, `gh run`, and `gh api` for issues, PRs, CI runs, and advanced queries.
Create or update AgentSkills. Use when designing, structuring, or packaging skills with scripts, references, and assets.
Start voice calls via the OpenClaw voice-call plugin.
Notion API for creating and managing pages, databases, and blocks.
Gemini CLI for one-shot Q&A, summaries, and generation.
Category:developer