🔨 Phase 3 Playbook — Build & Iterate agent

> Duration: 2-12 weeks (varies by scope) | Agents: 15-30+ | Gate Keeper: Agents Orchestrator

by msitarzewski·MIT license·★ 154,023 Stars on the repo·GitHub ↗

Files of 🔨 Phase 3 Playbook — Build & Iterate

msitarzewski/main1 file
phase-3-build.md
Show the full text287 lines

🔨 Phase 3 Playbook — Build & Iterate

Duration: 2-12 weeks (varies by scope) | Agents: 15-30+ | Gate Keeper: Agents Orchestrator


Objective

Implement all features through continuous Dev↔QA loops. Every task is validated before the next begins. This is where the bulk of the work happens — and where NEXUS's orchestration delivers the most value.

Pre-Conditions

  • Phase 2 Quality Gate passed (foundation verified)
  • Sprint Prioritizer backlog available with RICE scores
  • CI/CD pipeline operational
  • Design system and component library ready
  • API scaffold with auth system ready

The Dev↔QA Loop — Core Mechanic

The Agents Orchestrator manages every task through this cycle:

FOR EACH task IN sprint_backlog (ordered by RICE score):

  1. ASSIGN task to appropriate Developer Agent (see assignment matrix)
  2. Developer IMPLEMENTS task
  3. Evidence Collector TESTS task
     - Visual screenshots (desktop, tablet, mobile)
     - Functional verification against acceptance criteria
     - Brand consistency check
  4. IF verdict == PASS:
       Mark task complete
       Move to next task
     ELIF verdict == FAIL AND attempts < 3:
       Send QA feedback to Developer
       Developer FIXES specific issues
       Return to step 3
     ELIF attempts >= 3:
       ESCALATE to Agents Orchestrator
       Orchestrator decides: reassign, decompose, defer, or accept
  5. UPDATE pipeline status report

Agent Assignment Matrix

Primary Developer Assignment
Task Category Primary Agent Backup Agent QA Agent
React/Vue/Angular UI Frontend Developer Rapid Prototyper Evidence Collector
REST/GraphQL API Backend Architect Senior Developer API Tester
Database operations Backend Architect — API Tester
Mobile (iOS/Android) Mobile App Builder — Evidence Collector
ML model/pipeline AI Engineer — Test Results Analyzer
CI/CD/Infrastructure DevOps Automator Infrastructure Maintainer Performance Benchmarker
Premium/complex feature Senior Developer Backend Architect Evidence Collector
Quick prototype/POC Rapid Prototyper Frontend Developer Evidence Collector
WebXR/immersive XR Immersive Developer — Evidence Collector
visionOS visionOS Spatial Engineer macOS Spatial/Metal Engineer Evidence Collector
Cockpit controls XR Cockpit Interaction Specialist XR Interface Architect Evidence Collector
CLI/terminal tools Terminal Integration Specialist — API Tester
Code intelligence LSP/Index Engineer — Test Results Analyzer
Performance optimization Performance Benchmarker Infrastructure Maintainer Performance Benchmarker
Specialist Support (activated as needed)
Specialist When to Activate Trigger
UI Designer Component needs visual refinement Developer requests design guidance
Whimsy Injector Feature needs delight/personality UX review identifies opportunity
Visual Storyteller Visual narrative content needed Content requires visual assets
Brand Guardian Brand consistency concern QA finds brand deviation
XR Interface Architect Spatial interaction design needed XR feature requires UX guidance
Analytics Reporter Deep data analysis needed Feature requires analytics integration

Parallel Build Tracks

For NEXUS-Full deployments, four tracks run simultaneously:

Track A: Core Product Development
Managed by: Agents Orchestrator (Dev↔QA loop)
Agents: Frontend Developer, Backend Architect, AI Engineer,
        Mobile App Builder, Senior Developer
QA: Evidence Collector, API Tester, Test Results Analyzer

Sprint cadence: 2-week sprints
Daily: Task implementation + QA validation
End of sprint: Sprint review + retrospective
Track B: Growth & Marketing Preparation
Managed by: Project Shepherd
Agents: Growth Hacker, Content Creator, Social Media Strategist,
        App Store Optimizer

Sprint cadence: Aligned with Track A milestones
Activities:
- Growth Hacker → Design viral loops and referral mechanics
- Content Creator → Build launch content pipeline
- Social Media Strategist → Plan cross-platform campaign
- App Store Optimizer → Prepare store listing (if mobile)
Track C: Quality & Operations
Managed by: Agents Orchestrator
Agents: Evidence Collector, API Tester, Performance Benchmarker,
        Workflow Optimizer, Experiment Tracker

Continuous activities:
- Evidence Collector → Screenshot QA for every task
- API Tester → Endpoint validation for every API task
- Performance Benchmarker → Periodic load testing
- Workflow Optimizer → Process improvement identification
- Experiment Tracker → A/B test setup for validated features
Track D: Brand & Experience Polish
Managed by: Brand Guardian
Agents: UI Designer, Brand Guardian, Visual Storyteller,
        Whimsy Injector

Triggered activities:
- UI Designer → Component refinement when QA identifies visual issues
- Brand Guardian → Periodic brand consistency audit
- Visual Storyteller → Visual narrative assets as features complete
- Whimsy Injector → Micro-interactions and delight moments

Sprint Execution Template

Sprint Planning (Day 1)
Sprint Prioritizer activates:
1. Review backlog with updated RICE scores
2. Select tasks for sprint based on team velocity
3. Assign tasks to developer agents
4. Identify dependencies and ordering
5. Set sprint goal and success criteria

Output: Sprint Plan with task assignments
Daily Execution (Day 2 to Day N-1)
Agents Orchestrator manages:
1. Current task status check
2. Dev↔QA loop execution
3. Blocker identification and resolution
4. Progress tracking and reporting

Status report format:
- Tasks completed today: [list]
- Tasks in QA: [list]
- Tasks in development: [list]
- Blocked tasks: [list with reason]
- QA pass rate: [X/Y]
Sprint Review (Day N)
Project Shepherd facilitates:
1. Demo completed features
2. Review QA evidence for each task
3. Collect stakeholder feedback
4. Update backlog based on learnings

Participants: All active agents + stakeholders
Output: Sprint Review Summary
Sprint Retrospective
Workflow Optimizer facilitates:
1. What went well?
2. What could improve?
3. What will we change next sprint?
4. Process efficiency metrics

Output: Retrospective Action Items

Orchestrator Decision Logic

Task Failure Handling
WHEN task fails QA:
  IF attempt == 1:
    → Send specific QA feedback to developer
    → Developer fixes ONLY the identified issues
    → Re-submit for QA
    
  IF attempt == 2:
    → Send accumulated QA feedback
    → Consider: Is the developer agent the right fit?
    → Developer fixes with additional context
    → Re-submit for QA
    
  IF attempt == 3:
    → ESCALATE
    → Options:
      a) Reassign to different developer agent
      b) Decompose task into smaller sub-tasks
      c) Revise approach/architecture
      d) Accept with known limitations (document)
      e) Defer to future sprint
    → Document decision and rationale
Parallel Task Management
WHEN multiple tasks have no dependencies:
  → Assign to different developer agents simultaneously
  → Each runs independent Dev↔QA loop
  → Orchestrator tracks all loops concurrently
  → Merge completed tasks in dependency order

WHEN task has dependencies:
  → Wait for dependency to pass QA
  → Then assign dependent task
  → Include dependency context in handoff

Quality Gate Checklist

# Criterion Evidence Source Status
1 All sprint tasks pass QA (100% completion) Evidence Collector screenshots per task ☐
2 All API endpoints validated API Tester regression report ☐
3 Performance baselines met (P95 < 200ms) Performance Benchmarker report ☐
4 Brand consistency verified (95%+ adherence) Brand Guardian audit ☐
5 No critical bugs (zero P0/P1 open) Test Results Analyzer summary ☐
6 All acceptance criteria met Task-by-task verification ☐
7 Code review completed for all PRs Git history evidence ☐

Gate Decision

Gate Keeper: Agents Orchestrator

  • PASS: Feature-complete application → Phase 4 activation
  • CONTINUE: More sprints needed → Continue Phase 3
  • ESCALATE: Systemic issues → Studio Producer intervention

Handoff to Phase 4

## Phase 3 → Phase 4 Handoff Package

### For Reality Checker:
- Complete application (all features implemented)
- All QA evidence from Dev↔QA loops
- API Tester regression results
- Performance Benchmarker baseline data
- Brand Guardian consistency audit
- Known issues list (if any accepted limitations)

### For Legal Compliance Checker:
- Data handling implementation details
- Privacy policy implementation
- Consent management implementation
- Security measures implemented

### For Performance Benchmarker:
- Application URLs for load testing
- Expected traffic patterns
- Performance budgets from architecture

### For Infrastructure Maintainer:
- Production environment requirements
- Scaling configuration needs
- Monitoring alert thresholds

Phase 3 is complete when all sprint tasks pass QA, all API endpoints are validated, performance baselines are met, and no critical bugs remain open.

1# 🔨 Phase 3 Playbook — Build & Iterate
2 
3> **Duration**: 2-12 weeks (varies by scope) | **Agents**: 15-30+ | **Gate Keeper**: Agents Orchestrator
4 
5---
6 
7## Objective
8 
9Implement all features through continuous Dev↔QA loops. Every task is validated before the next begins. This is where the bulk of the work happens — and where NEXUS's orchestration delivers the most value.
10 
11## Pre-Conditions
12 
13- [ ] Phase 2 Quality Gate passed (foundation verified)
14- [ ] Sprint Prioritizer backlog available with RICE scores
15- [ ] CI/CD pipeline operational
16- [ ] Design system and component library ready
17- [ ] API scaffold with auth system ready
18 
19## The Dev↔QA Loop — Core Mechanic
20 
21The Agents Orchestrator manages every task through this cycle:
22 
23```
24FOR EACH task IN sprint_backlog (ordered by RICE score):
25 
26 1. ASSIGN task to appropriate Developer Agent (see assignment matrix)
27 2. Developer IMPLEMENTS task
28 3. Evidence Collector TESTS task
29 - Visual screenshots (desktop, tablet, mobile)
30 - Functional verification against acceptance criteria
31 - Brand consistency check
32 4. IF verdict == PASS:
33 Mark task complete
34 Move to next task
35 ELIF verdict == FAIL AND attempts < 3:
36 Send QA feedback to Developer
37 Developer FIXES specific issues
38 Return to step 3
39 ELIF attempts >= 3:
40 ESCALATE to Agents Orchestrator
41 Orchestrator decides: reassign, decompose, defer, or accept
42 5. UPDATE pipeline status report
43```
44 
45## Agent Assignment Matrix
46 
47### Primary Developer Assignment
48 
49| Task Category | Primary Agent | Backup Agent | QA Agent |
50|--------------|--------------|-------------|----------|
51| **React/Vue/Angular UI** | Frontend Developer | Rapid Prototyper | Evidence Collector |
52| **REST/GraphQL API** | Backend Architect | Senior Developer | API Tester |
53| **Database operations** | Backend Architect | — | API Tester |
54| **Mobile (iOS/Android)** | Mobile App Builder | — | Evidence Collector |
55| **ML model/pipeline** | AI Engineer | — | Test Results Analyzer |
56| **CI/CD/Infrastructure** | DevOps Automator | Infrastructure Maintainer | Performance Benchmarker |
57| **Premium/complex feature** | Senior Developer | Backend Architect | Evidence Collector |
58| **Quick prototype/POC** | Rapid Prototyper | Frontend Developer | Evidence Collector |
59| **WebXR/immersive** | XR Immersive Developer | — | Evidence Collector |
60| **visionOS** | visionOS Spatial Engineer | macOS Spatial/Metal Engineer | Evidence Collector |
61| **Cockpit controls** | XR Cockpit Interaction Specialist | XR Interface Architect | Evidence Collector |
62| **CLI/terminal tools** | Terminal Integration Specialist | — | API Tester |
63| **Code intelligence** | LSP/Index Engineer | — | Test Results Analyzer |
64| **Performance optimization** | Performance Benchmarker | Infrastructure Maintainer | Performance Benchmarker |
65 
66### Specialist Support (activated as needed)
67 
68| Specialist | When to Activate | Trigger |
69|-----------|-----------------|---------|
70| UI Designer | Component needs visual refinement | Developer requests design guidance |
71| Whimsy Injector | Feature needs delight/personality | UX review identifies opportunity |
72| Visual Storyteller | Visual narrative content needed | Content requires visual assets |
73| Brand Guardian | Brand consistency concern | QA finds brand deviation |
74| XR Interface Architect | Spatial interaction design needed | XR feature requires UX guidance |
75| Analytics Reporter | Deep data analysis needed | Feature requires analytics integration |
76 
77## Parallel Build Tracks
78 
79For NEXUS-Full deployments, four tracks run simultaneously:
80 
81### Track A: Core Product Development
82```
83Managed by: Agents Orchestrator (Dev↔QA loop)
84Agents: Frontend Developer, Backend Architect, AI Engineer,
85 Mobile App Builder, Senior Developer
86QA: Evidence Collector, API Tester, Test Results Analyzer
87 
88Sprint cadence: 2-week sprints
89Daily: Task implementation + QA validation
90End of sprint: Sprint review + retrospective
91```
92 
93### Track B: Growth & Marketing Preparation
94```
95Managed by: Project Shepherd
96Agents: Growth Hacker, Content Creator, Social Media Strategist,
97 App Store Optimizer
98 
99Sprint cadence: Aligned with Track A milestones
100Activities:
101- Growth Hacker → Design viral loops and referral mechanics
102- Content Creator → Build launch content pipeline
103- Social Media Strategist → Plan cross-platform campaign
104- App Store Optimizer → Prepare store listing (if mobile)
105```
106 
107### Track C: Quality & Operations
108```
109Managed by: Agents Orchestrator
110Agents: Evidence Collector, API Tester, Performance Benchmarker,
111 Workflow Optimizer, Experiment Tracker
112 
113Continuous activities:
114- Evidence Collector → Screenshot QA for every task
115- API Tester → Endpoint validation for every API task
116- Performance Benchmarker → Periodic load testing
117- Workflow Optimizer → Process improvement identification
118- Experiment Tracker → A/B test setup for validated features
119```
120 
121### Track D: Brand & Experience Polish
122```
123Managed by: Brand Guardian
124Agents: UI Designer, Brand Guardian, Visual Storyteller,
125 Whimsy Injector
126 
127Triggered activities:
128- UI Designer → Component refinement when QA identifies visual issues
129- Brand Guardian → Periodic brand consistency audit
130- Visual Storyteller → Visual narrative assets as features complete
131- Whimsy Injector → Micro-interactions and delight moments
132```
133 
134## Sprint Execution Template
135 
136### Sprint Planning (Day 1)
137 
138```
139Sprint Prioritizer activates:
1401. Review backlog with updated RICE scores
1412. Select tasks for sprint based on team velocity
1423. Assign tasks to developer agents
1434. Identify dependencies and ordering
1445. Set sprint goal and success criteria
145 
146Output: Sprint Plan with task assignments
147```
148 
149### Daily Execution (Day 2 to Day N-1)
150 
151```
152Agents Orchestrator manages:
1531. Current task status check
1542. Dev↔QA loop execution
1553. Blocker identification and resolution
1564. Progress tracking and reporting
157 
158Status report format:
159- Tasks completed today: [list]
160- Tasks in QA: [list]
161- Tasks in development: [list]
162- Blocked tasks: [list with reason]
163- QA pass rate: [X/Y]
164```
165 
166### Sprint Review (Day N)
167 
168```
169Project Shepherd facilitates:
1701. Demo completed features
1712. Review QA evidence for each task
1723. Collect stakeholder feedback
1734. Update backlog based on learnings
174 
175Participants: All active agents + stakeholders
176Output: Sprint Review Summary
177```
178 
179### Sprint Retrospective
180 
181```
182Workflow Optimizer facilitates:
1831. What went well?
1842. What could improve?
1853. What will we change next sprint?
1864. Process efficiency metrics
187 
188Output: Retrospective Action Items
189```
190 
191## Orchestrator Decision Logic
192 
193### Task Failure Handling
194 
195```
196WHEN task fails QA:
197 IF attempt == 1:
198 → Send specific QA feedback to developer
199 → Developer fixes ONLY the identified issues
200 → Re-submit for QA
201 
202 IF attempt == 2:
203 → Send accumulated QA feedback
204 → Consider: Is the developer agent the right fit?
205 → Developer fixes with additional context
206 → Re-submit for QA
207 
208 IF attempt == 3:
209 → ESCALATE
210 → Options:
211 a) Reassign to different developer agent
212 b) Decompose task into smaller sub-tasks
213 c) Revise approach/architecture
214 d) Accept with known limitations (document)
215 e) Defer to future sprint
216 → Document decision and rationale
217```
218 
219### Parallel Task Management
220 
221```
222WHEN multiple tasks have no dependencies:
223 → Assign to different developer agents simultaneously
224 → Each runs independent Dev↔QA loop
225 → Orchestrator tracks all loops concurrently
226 → Merge completed tasks in dependency order
227 
228WHEN task has dependencies:
229 → Wait for dependency to pass QA
230 → Then assign dependent task
231 → Include dependency context in handoff
232```
233 
234## Quality Gate Checklist
235 
236| # | Criterion | Evidence Source | Status |
237|---|-----------|----------------|--------|
238| 1 | All sprint tasks pass QA (100% completion) | Evidence Collector screenshots per task | ☐ |
239| 2 | All API endpoints validated | API Tester regression report | ☐ |
240| 3 | Performance baselines met (P95 < 200ms) | Performance Benchmarker report | ☐ |
241| 4 | Brand consistency verified (95%+ adherence) | Brand Guardian audit | ☐ |
242| 5 | No critical bugs (zero P0/P1 open) | Test Results Analyzer summary | ☐ |
243| 6 | All acceptance criteria met | Task-by-task verification | ☐ |
244| 7 | Code review completed for all PRs | Git history evidence | ☐ |
245 
246## Gate Decision
247 
248**Gate Keeper**: Agents Orchestrator
249 
250- **PASS**: Feature-complete application → Phase 4 activation
251- **CONTINUE**: More sprints needed → Continue Phase 3
252- **ESCALATE**: Systemic issues → Studio Producer intervention
253 
254## Handoff to Phase 4
255 
256```markdown
257## Phase 3 → Phase 4 Handoff Package
258 
259### For Reality Checker:
260- Complete application (all features implemented)
261- All QA evidence from Dev↔QA loops
262- API Tester regression results
263- Performance Benchmarker baseline data
264- Brand Guardian consistency audit
265- Known issues list (if any accepted limitations)
266 
267### For Legal Compliance Checker:
268- Data handling implementation details
269- Privacy policy implementation
270- Consent management implementation
271- Security measures implemented
272 
273### For Performance Benchmarker:
274- Application URLs for load testing
275- Expected traffic patterns
276- Performance budgets from architecture
277 
278### For Infrastructure Maintainer:
279- Production environment requirements
280- Scaling configuration needs
281- Monitoring alert thresholds
282```
283 
284---
285 
286*Phase 3 is complete when all sprint tasks pass QA, all API endpoints are validated, performance baselines are met, and no critical bugs remain open.*
287 

Discussion

Alternatives