AI agents as force multipliers for quality work. Core skill for all 19 QE agents using PACT principles.
Quick Agent Selection:
qe-test-generatorqe-coverage-analyzerqe-quality-gateqe-security-scannerqe-performance-testerqe-fleet-commanderCritical Success Factors:
| Principle | Agent Behavior | Human Role |
|---|---|---|
| Proactive | Analyze pre-merge, predict risk | Set guardrails |
| Autonomous | Execute tests, fix flaky tests | Review critical |
| Collaborative | Multi-agent coordination | Provide context |
| Targeted | Risk-based prioritization | Define risk areas |
| Structured | Governance, observability, explainable decisions (measure confidence, not trust) | Audit behavior, set policy |
| Category | Agents | Primary Use |
|---|---|---|
| Core Testing (5) | test-generator, test-executor, coverage-analyzer, quality-gate, quality-analyzer | Daily testing |
| Performance/Security (2) | performance-tester, security-scanner | Non-functional |
| Strategic (3) | requirements-validator, production-intelligence, fleet-commander | Planning |
| Advanced (4) | regression-risk-analyzer, test-data-architect, api-contract-validator, flaky-test-hunter | Specialized |
| Visual/Chaos (2) | visual-tester, chaos-engineer | Edge cases |
| Deployment (1) | deployment-readiness | Release |
| Analysis (1) | code-complexity | Maintainability |
Hierarchical: fleet-commander ā [generators] ā [executors] ā quality-gate
Mesh: test-gen ā coverage ā quality (peer decisions)
Sequential: risk-analyzer ā test-gen ā executor ā coverage ā gate
ā 10x deployment frequency with same/better quality ā Coverage gaps detected in real-time ā Bugs caught pre-production ā Agents acting without human oversight on critical decisions ā Deploying all 19 agents at once (start with 1-2)
| Stage | Approach | Limitation |
|---|---|---|
| Traditional | Manual everything | Human bottleneck |
| Automation | Scripts + fixed scenarios | Needs orchestration |
| Agentic | AI agents + human judgment | Requires trust-building |
Core Premise: Agents amplify human expertise for 10x scale.
1. Intelligent Test Generation
// Agent analyzes code change, generates targeted tests
const tests = await qeTestGenerator.generate(prDiff);
// ā Happy path, edge cases, error handling tests
2. Pattern Detection - Scan logs, find anomalies, correlate errors
3. Adaptive Strategy - Adjust test focus based on risk signals
4. Root Cause Analysis - Link failures to code changes, suggest fixes
aqe/test-plan/* - Test planning decisions
aqe/coverage/* - Coverage analysis results
aqe/quality/* - Quality metrics and gates
aqe/learning/* - Patterns and Q-values
aqe/coordination/* - Cross-agent state
CRITICAL: Always use aqe memory store with persist: true for learnings.
1. Store data to persistent memory:
// Store test plan decisions (persisted to .agentic-qe/memory.db)
aqe memory store \
--key "aqe/test-plan/pr-123" \
--namespace "aqe/test-plan" \
--value '{...}' \
--json
2. Retrieve prior learnings before task:
// Query patterns before starting test generation
const priorData = await aqe memory get --key "aqe/learning/patterns/test-generation/*" --namespace "aqe/learning" --json
// Use patterns to guide current task
if (priorData.success) {
console.log(`Loaded ${priorData.patterns.length} prior patterns`);
}
3. Store coverage analysis results:
aqe memory store \
--key "aqe/coverage/auth-module" \
--namespace "aqe/coverage" \
--value '{...}' \
--json
For coordinated multi-agent tasks, use the STATUS ā PROGRESS ā COMPLETE pattern:
// PHASE 1: STATUS - Task starting
aqe memory store \
--key "aqe/coordination/task-123/status" \
--namespace "aqe/coordination" \
--value '{...}' \
--json
// PHASE 2: PROGRESS - Intermediate updates
aqe memory store \
--key "aqe/coordination/task-123/progress" \
--namespace "aqe/coordination" \
--value '{...}' \
--json
// PHASE 3: COMPLETE - Task finished
aqe memory store \
--key "aqe/coordination/task-123/complete" \
--namespace "aqe/coordination" \
--value '{...}' \
--json
| Event | Trigger | Subscribers |
|---|---|---|
test:generated |
New tests created | executor, coverage |
coverage:gap |
Gap detected | test-generator |
quality:decision |
Gate evaluated | fleet-commander |
security:finding |
Vulnerability found | quality-gate |
// 1. Risk analysis
const risks = await Task("Analyze PR", prDiff, "qe-regression-risk-analyzer");
// 2. Generate tests for risks
const tests = await Task("Generate tests", risks, "qe-test-generator");
// 3. Execute + analyze
const results = await Task("Run tests", tests, "qe-test-executor");
const coverage = await Task("Check coverage", results, "qe-coverage-analyzer");
// 4. Quality decision
const decision = await Task("Evaluate", {results, coverage}, "qe-quality-gate");
// ā GO/NO-GO with rationale
| Phase | Duration | Goal | Agent(s) |
|---|---|---|---|
| Experiment | Weeks 1-4 | Validate one use case | 1 agent |
| Integrate | Months 2-3 | CI/CD pipeline | 3-4 agents |
| Scale | Months 4-6 | Multiple use cases | 8+ agents |
| Evolve | Ongoing | Continuous learning | Full fleet |
# Week 1: Deploy single agent
aqe agent spawn qe-test-generator
# Weeks 2-3: Generate tests for 10 PRs
# Track: bugs found, test quality, review time
# Week 4: Measure impact
aqe agent metrics qe-test-generator
# ā Tests: 150, Bugs: 12, Time saved: 8h
| Do | Don't |
|---|---|
| Start with one agent, one use case | Deploy all 18 at once |
| Build feedback loops early | Deploy and forget |
| Human reviews agent output | Auto-merge without review |
| Measure bugs caught, time saved | Track vanity metrics (test count) |
| Build trust gradually | Give full autonomy immediately |
Month 1: Agent suggests ā Human decides
Month 2: Agent acts ā Human reviews after
Month 3: Agent autonomous on low-risk
Month 4: Agent handles critical with oversight
coordination:
topology: hierarchical
commander: qe-fleet-commander
memory_namespace: aqe/coordination
blackboard_topic: qe-fleet
preload_skills:
- agentic-quality-engineering # Always (this skill)
- risk-based-testing # For prioritization
- quality-metrics # For measurement
agent_assignments:
qe-test-generator: [api-testing-patterns, tdd-london-chicago]
qe-coverage-analyzer: [quality-metrics, risk-based-testing]
qe-security-scanner: [security-testing, risk-based-testing]
qe-performance-tester: [performance-testing]
holistic-testing-pact - PACTS principles deep diverisk-based-testing - Prioritize agent focusquality-metrics - Measure agent effectivenessapi-testing-patterns, security-testing, performance-testing - Specialized testing.claude/agents/aqe agent --helpaqe fleet statusSuccess Metric: Deploy 10x more frequently with same or better quality through intelligent agent collaboration.