quality AI Agent Skills

Browse 67 skills related to quality

Verification & Quality Assurance

17.8k
ruvnetruvnet

Comprehensive truth scoring, code quality verification, and automatic rollback system with 0.95 accuracy threshold for ensuring high-quality agent outputs and codebase reliability.

verificationtruth-scoringquality+3
141 days ago

Testing Anti-Patterns

451
rohitg00rohitg00

Common testing mistakes to avoid for reliable, maintainable tests

testinganti-patternsquality+1
141 days ago

Red-Green-Refactor

451
rohitg00rohitg00

Test-driven development cycle for writing reliable, well-designed code

tddtestingquality+1
141 days ago

Structured Code Review

451
rohitg00rohitg00

Multi-stage code review process for thorough and constructive feedback

collaborationcode-reviewquality+1
141 days ago

Verification Gates

451
rohitg00rohitg00

Define checkpoints to validate work before proceeding to next phase

planningverificationquality+1
141 days ago

Test Patterns

451
rohitg00rohitg00

Effective patterns for writing maintainable, reliable tests

testingpatternsquality+1
141 days ago

test-coverage

241
MadAppGangMadAppGang

On-demand test coverage analysis. Use when identifying untested code, finding test gaps, measuring coverage metrics, or improving test quality. Trigger keywords - "test coverage", "coverage report", "untested code", "test gaps", "missing tests", "coverage metrics".

devtestingcoverage+2
141 days ago

audit

241
MadAppGangMadAppGang

On-demand security and code quality audit. Use when checking for vulnerabilities, security issues, code smells, or compliance problems. Trigger keywords - "audit", "security check", "vulnerability scan", "code quality", "compliance", "security audit".

devauditsecurity+2
141 days ago

brutal-honesty-review

215
proffesor-for-testingproffesor-for-testing

Unvarnished technical criticism combining Linus Torvalds' precision, Gordon Ramsay's standards, and James Bach's BS-detection. Use when code/tests need harsh reality checks, certification schemes smell fishy, or technical decisions lack rigor. No sugar-coating, just surgical truth about what's broken and why.

code-reviewhonestycritical-thinking+2
141 days ago

qe-holistic-testing-pact

215
proffesor-for-testingproffesor-for-testing

Apply the Holistic Testing Model evolved with PACT (Proactive, Autonomous, Collaborative, Targeted) principles. Use when designing comprehensive test strategies for Classical, AI-assisted, Agent based, or Agentic Systems building quality into the team, or implementing whole-team quality practices.

holisticpactquality+5
141 days ago

code-review-quality

215
proffesor-for-testingproffesor-for-testing

Conduct context-driven code reviews focusing on quality, testability, and maintainability. Use when reviewing code, providing feedback, or establishing review practices.

code-reviewfeedbackquality+3
141 days ago

holistic-testing-pact

215
proffesor-for-testingproffesor-for-testing

Apply the Holistic Testing Model evolved with PACT (Proactive, Autonomous, Collaborative, Targeted) principles. Use when designing comprehensive test strategies for classical, AI-assisted, agent-based, or agentic systems to build quality into the team or implement whole-team quality practices.

holisticpactquality+5
141 days ago

qe-brutal-honesty-review

215
proffesor-for-testingproffesor-for-testing

Unvarnished technical criticism combining Linus Torvalds' precision, Gordon Ramsay's standards, and James Bach's BS-detection. Use when code/tests need harsh reality checks, certification schemes smell fishy, or technical decisions lack rigor. No sugar-coating, just surgical truth about what's broken and why.

code-reviewhonestycritical-thinking+2
141 days ago

Verification & Quality Assurance

215
proffesor-for-testingproffesor-for-testing

Comprehensive truth scoring, code quality verification, and automatic rollback system with 0.95 accuracy threshold for ensuring high-quality agent outputs and codebase reliability.

verificationtruth-scoringquality+3
141 days ago

qe-bug-reporting-excellence

215
proffesor-for-testingproffesor-for-testing

Write high-quality bug reports that get fixed quickly. Use when reporting bugs, training teams on bug reporting, or establishing bug report standards.

bugsreportingcommunication+2
141 days ago

qe-verification-quality

215
proffesor-for-testingproffesor-for-testing

Comprehensive truth scoring, code quality verification, and automatic rollback system with 0.95 accuracy threshold for ensuring high-quality agent outputs and codebase reliability.

verificationtruth-scoringquality+3
141 days ago

test-idea-rewriting

215
proffesor-for-testingproffesor-for-testing

Transform passive "Verify X" test descriptions into active, observable test actions. Use when test ideas lack specificity, use vague language, or fail quality validation. Converts to action-verb format for clearer, more testable descriptions.

test-ideasrewritingaction-verbs+2
141 days ago

bug-reporting-excellence

215
proffesor-for-testingproffesor-for-testing

Write high-quality bug reports that get fixed quickly. Use when reporting bugs, training teams on bug reporting, or establishing bug report standards.

bugsreportingcommunication+2
141 days ago

qe-code-review-quality

215
proffesor-for-testingproffesor-for-testing

Conduct context-driven code reviews focusing on quality, testability, and maintainability. Use when reviewing code, providing feedback, or establishing review practices.

code-reviewfeedbackquality+3
141 days ago

slop-detector

196
atholaathola

Detect and flag AI-generated content markers in documentation and prose. Use when reviewing documentation for AI markers, cleaning up LLM-generated content, or auditing prose quality. Do not use when generating new content (use doc-generator) or learning writing styles (use style-learner).

ai-detectionslopwriting+3
141 days ago

verification-quality-assurance

188
aiskillstoreaiskillstore

Comprehensive truth scoring, code quality verification, and automatic rollback system with 0.95 accuracy threshold for ensuring high-quality agent outputs and codebase reliability.

verificationtruth-scoringquality+3
141 days ago

code-review

188
aiskillstoreaiskillstore

Battle-tested code review practices optimizing for codebase health and team velocity

code-reviewprquality+3
141 days ago

code-review

127
femtofemto

Performs thorough code reviews with focus on best practices, security, performance, and maintainability. Use this skill when reviewing pull requests, auditing code quality, or getting feedback on implementations.

code-reviewsecurityperformance+2
141 days ago

golden-dataset

97
yonatangrossyonatangross

Golden dataset lifecycle patterns for curation, versioning, quality validation, and CI integration. Use when building evaluation datasets, managing dataset versions, validating quality scores, or integrating golden tests into pipelines.

golden-datasetevaluationdataset-curation+3
141 days ago

code-review-playbook

97
yonatangrossyonatangross

Use this skill when conducting or improving code reviews. Provides structured review processes, conventional comments patterns, language-specific checklists, and feedback templates. Use when reviewing PRs or standardizing review practices.

code-reviewqualitycollaboration+1
141 days ago

quality-gates

97
yonatangrossyonatangross

Use when assessing task complexity, before starting complex tasks, when stuck after multiple attempts, or reviewing code against best practices. Provides quality-gates scoring (1-5), escalation workflows, and pattern library management.

qualitycomplexityplanning+6
141 days ago

assess

97
yonatangrossyonatangross

Assesses and rates quality 0-10 with pros/cons analysis. Use when evaluating code, designs, or approaches.

assessmentevaluationquality+3
141 days ago

audit-skills

97
yonatangrossyonatangross

Audits all OrchestKit skills for quality, completeness, and compliance with authoring standards. Use when checking skill health, before releases, or after bulk skill edits to surface SKILL.md files that are too long, have missing frontmatter, lack rules/references, or are unregistered in manifests.

auditqualityskills+1
141 days ago

review-pr

97
yonatangrossyonatangross

PR review with parallel specialized agents. Use when reviewing pull requests or code.

code-reviewpull-requestquality+2
141 days ago

Advanced Mutation Testing

71
PramodDuttaPramodDutta

Advanced mutation testing using Stryker, PIT, and mutmut to measure test suite quality by introducing controlled mutations and tracking kill rates.

mutation-testingstrykerpit+2
141 days ago

copy-editor

57
atrislabsatrislabs

Detects AI writing patterns and fixes them. Use when reviewing any written output, including docs, READMEs, messages, PRDs. Based on Wikipedia's AI Cleanup patterns. Triggers on "copy edit", "review writing", "humanize", "deslopper", "ai patterns", "make it sound human".

copy-editorwritinganti-slop+1
141 days ago

dag-output-validator

43
curiositechcuriositech

Validates agent outputs against expected schemas and quality criteria. Ensures outputs meet structural requirements and content standards. Activate on 'validate output', 'output validation', 'schema validation', 'check output', 'output quality'. NOT for confidence scoring (use dag-confidence-scorer) or hallucination detection (use dag-hallucination-detector).

dagqualityvalidation+2
141 days ago

code-review-checklist

43
curiositechcuriositech

Generates comprehensive, context-aware code review checklists tailored to the specific codebase, programming language, and team standards. Analyzes PR diffs and suggests what reviewers should focus on.

code-reviewqualitychecklist+2
141 days ago

skill-coach

43
curiositechcuriositech

Guides creation of high-quality Agent Skills with domain expertise, anti-pattern detection, and progressive disclosure best practices. Activate on keywords: create skill, review skill, skill quality, skill best practices, skill anti-patterns, improve skill, skill audit. NOT for general coding advice, slash commands, MCP development, or non-skill Claude Code features.

skillsqualityanti-patterns+2
141 days ago

dag-confidence-scorer

43
curiositechcuriositech

Assigns confidence scores to agent outputs based on multiple factors including source quality, consistency, and reasoning depth. Produces calibrated confidence estimates. Activate on 'confidence score', 'how confident', 'certainty level', 'output confidence', 'reliability score'. NOT for validation (use dag-output-validator) or hallucination detection (use dag-hallucination-detector).

dagqualityconfidence+2
141 days ago

dag-hallucination-detector

43
curiositechcuriositech

Detects fabricated content, false citations, and unverifiable claims in agent outputs. Uses source verification and consistency checking. Activate on 'detect hallucination', 'fact check', 'verify claims', 'check accuracy', 'find fabrications'. NOT for validation (use dag-output-validator) or confidence scoring (use dag-confidence-scorer).

dagqualityhallucination+2
141 days ago

dag-iteration-detector

43
curiositechcuriositech

Identifies when task outputs require iteration based on quality signals, unmet requirements, or explicit feedback. Triggers appropriate re-execution strategies. Activate on 'needs iteration', 'retry needed', 'not good enough', 'try again', 'refine output'. NOT for feedback generation (use dag-feedback-synthesizer) or convergence tracking (use dag-convergence-monitor).

dagfeedbackiteration+2
141 days ago

skill-logger

43
curiositechcuriositech

Logs and scores skill usage quality, tracking output effectiveness, user satisfaction signals, and improvement opportunities. Expert in skill analytics, quality metrics, feedback loops, and continuous improvement. Activate on "skill logging", "skill quality", "skill analytics", "skill scoring", "skill performance", "skill metrics", "track skill usage", "skill improvement". NOT for creating skills (use agent-creator), skill documentation (use skill-coach), or runtime debugging (use debugger skills).

logginganalyticsmetrics+2
141 days ago

code-analyzer

35
cladamcladam

Code quality analysis skill for tbdflow, covering smells, maintainability, and refactoring guidance.

analysisqualityrefactoring+1
141 days ago

code-review

30
SimHackerSimHacker

Systematic code analysis with evidence collection

moollmdevelopmentquality+2
141 days ago

e2e-test

20
DNYoussefDNYoussef

End-to-end testing workflow for validating complete user journeys through web applications using claude-in-chrome MCP. Specializes in test assertions, suite organization, evidence collection, and pass/fail reporting.

testinge2eautomation+3
141 days ago

code-review

18
LangConfigLangConfig

Systematic code review guidance covering best practices, security, performance, and maintainability. Use when reviewing code, checking PRs, or analyzing code quality.

code-reviewbest-practicessecurity+2
141 days ago

codex-review-workflow

17
daffy0208daffy0208

Automated code review workflow using OpenAI Codex CLI. Implements iterative fix-and-review cycles until code passes validation or reaches iteration limit. Use when building features requiring automated code validation, security checks, or quality assurance through Codex CLI.

automationcode-reviewquality+1
141 days ago

Skill Validator

17
daffy0208daffy0208

Validates that a skill or MCP implementation matches its manifest by running Codex-powered semantic comparisons across descriptions, preconditions, effects, and API surface.

codexvalidationquality
141 days ago

quality-audit

11
NickCrewNickCrew

Meta-skill for auditing and validating skill quality. Use when reviewing skills for consistency, completeness, accuracy, and adherence to standards. Provides structured rubrics, scoring frameworks, and actionable recommendations.

metaqualityvalidation+2
141 days ago

code-review-playbook

8
ArieGoldkinArieGoldkin

Use this skill when conducting or improving code reviews. Provides structured review processes, conventional comment patterns, language-specific checklists, and feedback templates. Ensures consistent, constructive, and thorough code reviews across teams.

code-reviewqualitycollaboration+1
141 days ago

skill-md-validator

7
louloulinlouloulin

Validates SKILL.md files for format, completeness, and best practices

validationdocumentationquality+1
141 days ago

rulebook-quality-gates

7
hivellmhivellm

Automated quality checks and enforcement for code commits. Use when validating code quality, running pre-commit checks, ensuring test coverage, or enforcing coding standards before commits and pushes.

qualitytestingcoverage+3
141 days ago