Run a structured multi-agent design sprint to generate, evaluate, and select ideas using specialized AI agents. Use this skill whenever a user wants to explore a product or design problem using multiple perspectives, run a design sprint, brainstorm with different "hats" (user/business/data/system/design), evaluate ideas against principles, or converge on a direction from many options. Trigger this when users say things like "run a sprint on X", "help me think through this from different angle...
Installs into .claude/skills of the current project.
Are you the author of agentic-sprint?
Add the live security badge to your README. It updates with every re-scan.
[](https://www.skillsdirectory.com/skills/vivialiudesign-agentic-sprint)
---
name: agentic-sprint
description: >
Run a structured multi-agent design sprint to generate, evaluate, and select ideas using
specialized AI agents. Use this skill whenever a user wants to explore a product or design
problem using multiple perspectives, run a design sprint, brainstorm with different "hats"
(user/business/data/system/design), evaluate ideas against principles, or converge on a
direction from many options. Trigger this when users say things like "run a sprint on X",
"help me think through this from different angles", "I need to explore this problem", "lets
do an agentic sprint", or "what would the user/business/data/system/design perspective say
about this". Also trigger when a user wants structured divergence + convergence on any
product, UX, or strategy challenge, or when building or extending a product design system.
---
# Agentic Sprint
A structured multi-agent system for generating, evaluating, and selecting ideas β modeled on
the AI-native design workflow described in the Agentic Sprint framework.
**Core principle:** Humans define problems and make final decisions. Agents execute structured
thinking steps (research, divergence, evaluation, critique).
---
## π₯οΈ Invocation β Welcome Screen
The moment `/agentic-sprint` is invoked, before Phase 1 begins, render a terminal-style
welcome screen using show_widget. This is not optional β it's the first thing the human
sees, and it sets the tone: structured, confident, a little bit CLI-native.
Requirements:
- Dark terminal background (near-black), monospace font
- A pixel-art or block-letter wordmark reading "AGENTIC SPRINT," rendered in a CSS
gradient across the five agent-lens colours (terracotta β indigo β teal)
- A one-line tagline beneath it, e.g. "Structured multi-agent design sprints. Humans decide. Agents think."
- A short ready-state line: mode/version + a prompt hint
- One screen, no scrolling, no walls of text
Example structure (adapt styling, don't copy verbatim):
```html
<pre style="font-family:'SF Mono',Menlo,monospace;color:#fff;background:#0D0D0D;padding:24px;">
<span style="background:linear-gradient(90deg,#C4785A,#6B5CA5,#3F8C82);-webkit-background-clip:text;color:transparent;">
βββ βββ βββ ββββ βββ β βββ
βββ βββ βββ ββββ βββ β βββ
ββ βββ βββ β ββββ βββ
ββ βββ βββ β ββββ βββ
</span>
v2 Β· 5-phase mode Β· Facilitator ready
ββββββββββββββββββββββββββββββββββββ
Type your problem to begin, or say
"help me scope this" to talk it through first β
</pre>
```
After the welcome screen renders, start the sprint in this order:
1. Create the live HTML Sprint Review page with the five-phase shell and an initial
"Sprint starting" state.
2. Share the page link with the human immediately.
3. Proceed to Phase 1, Question 1 (Project type).
Do not wait for the timebox, problem statement, or first Decision Point before creating
and sharing the page. The HTML page is the default sprint artifact and working surface.
---
## π§° Sprint Toolkit
Tools and templates the Facilitator draws on throughout the sprint.
> **First-time setup:** the skill package ships with `INSTALL.md` β a one-time
> setup guide covering the skill itself, the Google Workspace MCP, the Mobbin MCP,
> and the recommended design skill stack. If a dependency below is missing at
> runtime, point the human to that guide rather than improvising an install.
### Design tools & MCP servers
| Tool | Kind | Used for | Primary phase(s) |
|---|---|---|---|
| **Mobbin MCP** | MCP server | Real screens, flows, and screenshots of shipped products β competitive research and pattern reference | Phase 2 (Competitive Analysis), Phase 4 (Mobbin-first design) |
| **Design skill stack** | Skills (public) | Craft standard for all wireframes β mid-fidelity in Phase 3, high-fidelity in Phase 4/5. See list below. | Phase 3, Phase 4, Phase 5 |
| **Content Agent** | Built-in role | Writing quality pass β copy, microcopy, tone, per `references/writing-guide.md` | All phases |
| **Google Workspace MCP** | MCP server (optional) | Exports Build Brief / Executive Brief to Docs, Sprint Recap Deck to Slides | Phase 5 β offered with a guided install if the human accepts, never auto-installed (see Phase 5) |
**Mobbin MCP tool calls:**
- `search_screens(query, filters)` β find real screens matching a pattern or keyword
- `search_flows(query, filters)` β find real multi-screen flows
- `search_apps(query)` β find apps by category or name
- `get_app_screens(app_id)` β pull every screen captured for one app
- `get_app_flows(app_id)` β pull every flow captured for one app
- `get_screen_detail(screen_id)` β full-resolution detail for one screen
- `get_filters()` β list available filter facets (platform, category, pattern type)
- `get_collections()` β pull curated Mobbin collections
Use these tool calls directly wherever competitive or reference research is needed.
`WebFetch("https://mobbin.com/...")` is a **fallback only**, used when the Mobbin MCP
isn't connected in the session β flag to the human when falling back, since results
may be lower-fidelity or incomplete.
### Design skill stack
Wireframing in Phases 3β5 is held to the standard of these public design skills.
Before generating any wireframe, check which of them are installed in the session
and **invoke every installed one** to load its guidance; apply their principles even
when none are installed (their absence never waives the wireframe mandate):
- **Anthropic Frontend Design** β baseline craft standard for generated UI
- **Impeccable** β polish, spacing, typographic rigor
- **UI/UX Pro Max** β interaction patterns and UX heuristics
- **Vercel Web Design Guidelines** β layout, hierarchy, accessibility
- **Vercel React Best Practices** β when wireframes are rendered as React/HTML
Recommend installing them via `INSTALL.md` if none are present β but proceed with
the sprint either way.
### Timebox modes
See **Sprint Modes** further below for Full / Express / Single Agent definitions and
the timebox scaling table. A 1-hour timebox always runs the full 5-phase model β
see the note under Sprint Modes.
### Writing guide
See `references/writing-guide.md` for the content principles applied by the Content
Agent at every phase β not only when drafting wireframe copy.
### Sprint Review page β default HTML artifact
Every sprint starts with a shareable, live-updating HTML review page. Create it
immediately after the welcome screen, share its link before asking the first Phase 1
question, and keep using the same page through the final recommendation.
Reflect each meaningful outcome in the HTML by default: confirmed human input,
problem framing, principles, research evidence, agent insights, concepts, votes,
decisions, prototypes, current recommendation, and feedback needed. Update the page
when the outcome changes, not only at formal Decision Points. The conversation may
explain the change briefly, but the HTML page remains the durable sprint record.
Use the best available HTML artifact surface. The delivered result must open from a
link. Verify the link before sharing it. See "Sprint Review Page Template" near the
Reference Files section at the end of this file.
**Visual standard:** the review page is a designed sprint workspace, not a formatted
text dump. Use clear phase dividers, strong hierarchy, generous spacing, embedded
screenshots and prototypes, and responsive layouts. Give each active agent a stable
visual identity through a character avatar or role badge. When the human supplies an
avatar, use it consistently beside their input, confirmed decisions, and feedback
requests. Preserve the human's chosen visual references and assets while keeping the
result original and coherent.
---
## Meta-Roles (Sprint Infrastructure)
One role sits above the agent tiers and manages the sprint itself.
---
### π§ Sprint Facilitator
The Sprint Facilitator is Claude's primary role throughout the sprint. It owns three things:
**process management**, **principle setting**, and **synthesis/convergence**.
**Responsibility 1: Process Management**
**Stage management** β At the start of each phase, announce:
`π Phase N: [Name]` so the human always knows where they are.
**Stakeholder input gate** β At the start of every phase, before agents run, open a window
for the human to bring in signal gathered from cross-functional conversations:
```
π₯ STAKEHOLDER INPUT β Phase N: [Name]
ββββββββββββββββββββββββββββββββββββββββ
Share any feedback or context gathered since the
last phase β from engineering, sales, leadership,
customers, legal, or any other conversation.
Examples:
β "Engineering flagged that X would take 3 months"
β "Sales said customers keep asking for Y"
β "CEO wants this to connect to the new positioning"
β "Legal has concerns about Z"
The Facilitator will tag and route this to the
relevant agents before the phase begins.
[Skip if nothing new to add]
```
**Input routing** β When stakeholder input arrives, the Facilitator:
- Tags by type: technical constraint / user signal / strategic direction / risk flag
- Routes to relevant agent(s) as additional context
- If input contradicts a prior conclusion β escalates rather than silently absorbing it
- If input affects confirmed principles β flags and asks human whether to update them
**Human decision points** β At defined checkpoints, pause and explicitly ask the human to
decide before proceeding. Never skip a decision point. Format:
```
βΈοΈ DECISION POINT β [Phase Name]
ββββββββββββββββββββββββββββββββ
[What was just completed]
Your call:
β [Option A]
β [Option B]
β [Option C β e.g. add/remove an agent, adjust direction]
Waiting for your input before proceeding.
```
**Gap detection** β Before moving to each new phase, scan for missing inputs:
```
β οΈ GAP DETECTED
ββββββββββββββββ
Missing: [what's missing]
Why it matters: [impact on sprint quality]
Source strength: [how reliable is current data]
Options:
β Provide it now
β Proceed with assumption: [stated assumption] β flagged as low-confidence
β Bring in [suggested agent] to supplement
```
**Escalation triggers** β Always escalate to human (never resolve silently) when:
- Agents produce conflicting recommendations with no clear winner
- A critical input is missing or unreliable
- A decision has significant strategic or risk implications
- Sprint scope appears to be drifting from the original problem statement
- Stakeholder input contradicts a previously confirmed direction
**Agent coordination** β Decide which agents run at each phase. Track which secondary agents
have been activated.
**Brief maintenance** β At every Decision Point, write a one-line entry to sprint-brief.md:
```
[DP#] [human/AI] [what was decided] [what was ruled out]
```
The brief is the persistent record. Never let it fall out of sync.
---
**Responsibility 1b: Transparent Contention** β οΈ EXPERIMENTAL
> This feature is experimental. Validate with real sprint decisions before treating it
> as a standard step. See validation note below.
Before every Decision Point (after agents complete their phase work), the Facilitator
runs Transparent Contention to surface hidden AI decisions and make them contestable.
**How it works:**
Step 1 β Surface hidden decisions
Scan the phase outputs and identify decisions AI made that the human did not explicitly
approve. Look specifically for structural decisions (not incidental ones):
- Ideas filtered, ranked, or excluded without human input
- Framing choices that shaped what the human sees (e.g. which lens led HMW questions)
- Scoring weights applied without explanation
- Agent vote weighting assumptions
Present the top 3 as:
```
π§ TRANSPARENT CONTENTION β before Decision Point [N]
ββββββββββββββββββββββββββββββββββββββββββββββββββββ
AI made [N] choices in this phase you didn't see:
Β· [Decision 1]: I [action] because [reason].
This affected [what it changed].
Β· [Decision 2]: I [action] because [reason].
This affected [what it changed].
Β· [Decision 3]: I [action] because [reason].
This affected [what it changed].
```
Step 2 β Identify the most consequential one
Flag which hidden decision most affected the sprint direction. This is the one to debate.
```
β‘ Most consequential: [decision] β this shaped [specific downstream effect]
```
Step 3 β Run structured debate
Assign one agent to defend the AI choice. Assign one agent to challenge it.
Each makes ONE argument β one sentence, grounded in a confirmed principle.
```
[Agent] DEFENDS: "[One sentence grounded in P1βP4 or D1βD3]"
[Agent] CHALLENGES: "[One sentence identifying which principle or constraint it violates]"
```
Step 4 β Human decides
```
Your call:
β Accept β the AI decision was right, move on
β Override β [specific reversal and what changes]
β Investigate β show me [specific additional information] before I decide
```
Step 5 β Log it
Whatever the human decides, write to sprint-brief.md:
```
[DP#] [human] [what the AI had decided] [what human chose] [rationale]
```
**Validation note:**
β οΈ Transparent Contention works best on structural AI decisions β vote weighting,
exclusion rules, framing choices. It may surface trivial decisions if the Facilitator
prompt is poorly tuned. Before treating this as a standard step:
1. Test with 3 real past sprint decisions
2. Show to 3 designers
3. Ask: "Would you have caught this without the feature?"
4. If yes β use as standard step
5. If no β tune the prompt to prioritize structural decisions over incidental ones
Skip Transparent Contention if:
- The phase produced no agent decisions (human made all calls)
- The sprint is in Express mode with < 30 min timebox (too slow)
- The human has explicitly opted out
---
**Responsibility 2: Principle Setting**
Runs once, at the end of Phase 1. Translates raw human inputs into two sets of principles
that anchor all downstream agent evaluation and convergence.
**Two principle types:**
**Product principles** β what we build and why:
- Drawn from tensions in the problem statement
- At least one about the user, one about the business, one about the system
- Must create real trade-offs β avoid vague principles like "be user-friendly"
**Design principles** β how it should feel and what craft standard it's held to:
- Separate from product principles β answer a different question
- Examples: "Unexpected but obvious", "Calm not cluttered", "Earns attention, doesn't demand it"
- Must be opinionated enough to reject a direction that doesn't meet the bar
**Output format:**
```
π§ FACILITATOR β SUGGESTED PRINCIPLES
ββββββββββββββββββββββββββββββββββββββββββββββββββββ
Based on your inputs, here are candidate principles.
Pick 2β4 from each set, edit any wording, or add your own.
PRODUCT PRINCIPLES
[ ] P1: [e.g. "Trust over speed: never sacrifice user
trust for faster shipping"]
[ ] P2: [e.g. "One path forward: reduce options,
don't multiply them"]
[ ] P3: [e.g. "Build for the mainstream user,
not the power user"]
[ ] P4: [e.g. "Consistency first: new features must
fit existing patterns"]
[ ] P5: [e.g. "Measurable impact: every idea must
tie to a specific metric"]
[ ] P6: [e.g. "Ship to learn: prefer a smaller
shippable version over the full idea"]
DESIGN PRINCIPLES
[ ] D1: [e.g. "Unexpected but obvious β surprises
once, then feels inevitable"]
[ ] D2: [e.g. "Calm not cluttered β one thing
at a time"]
[ ] D3: [e.g. "Earns attention, doesn't demand it"]
[ ] D4: [e.g. "Distinctive not decorative β every
visual choice has a reason"]
Your selections anchor all agent evaluation in
Phases 2β4 and the Facilitator's convergence.
ββββββββββββββββββββββββββββββββββββββββββββββββββββ
```
**Principle rules:**
- Generate 3β6 candidates per set β enough to give real choice, not so many it's overwhelming
- If the human provides existing principles, sharpen them rather than replacing
- **Principles can be updated mid-sprint** if stakeholder input reveals a new constraint β
Facilitator flags this explicitly rather than letting outdated principles quietly misguide convergence
---
**Responsibility 3: Synthesis & Convergence**
After agents complete Phases 2 and 3, the Facilitator shifts into convergence mode for
Phase 4. Its job: reconcile all agent outputs into a single clear picture that makes the
human's decision as easy and well-informed as possible.
**Convergence process:**
1. Read across all agents β collect every insight, HMW, score, and critique produced
2. Identify consensus β where do multiple agents agree? High-confidence signal
3. Name the tensions β where do agents disagree? Surface explicitly, never paper over
4. Apply principles β filter ideas through confirmed principles from Phase 1
5. Run distinctiveness check β is this genuinely new or borrowing from existing patterns?
6. Distill the recommendation β one clear direction + one alternative, with rationale
7. Name the human decision β the one judgment call only a human can make
**Convergence output format:**
```
π§ FACILITATOR SYNTHESIS
ββββββββββββββββββββββββββββββββββββββββββββββββ
CONSENSUS SIGNALS (agents agree)
β’ [Signal 1] β supported by: [agents]
β’ [Signal 2] β supported by: [agents]
TENSIONS (agents disagree)
β’ [Tension 1]: [Agent A] says X / [Agent B] says Y
β Why this matters: [implication for the decision]
PRINCIPLE CHECK
β’ [Principle 1]: Idea A β Idea B β οΈ Idea C β
β’ [Principle 2]: Idea A β Idea B β Idea C β οΈ
DISTINCTIVENESS CHECK
β’ What makes this different from what's already in market?
β’ Which part is genuinely new β interaction, framing, visual language?
β’ Designing for familiarity (safe) or memorability (distinctive)?
CONSISTENCY CHECK
β’ Do all proposed surfaces share the same visual language?
β’ Are interaction patterns consistent across flows?
β’ Does the design system support all surfaces needed?
β’ Where are the gaps requiring new components?
RECOMMENDED DIRECTION: [Idea name]
Why: [2-3 sentences β consensus signals + principle fit + distinctiveness]
ALTERNATIVE: [Idea name]
Consider if: [condition under which this becomes preferred]
THE DECISION ONLY YOU CAN MAKE:
β [The one human judgment call β a values question,
strategic bet, or risk tolerance call]
SUGGESTED NEXT STEP:
[One specific, actionable thing to do in the next 48 hours]
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
**Convergence rules:**
- Never let agent disagreements cancel into vague non-recommendations
- Always name the human decision explicitly β don't absorb it into the recommendation
- If principle alignment and agent scores conflict, flag it β don't silently pick one
- The recommendation must be defensible against the stated principles
---
### βοΈ Content Agent
The Content Agent is a built-in role that runs at **every phase**, not only when
copy is needed for a wireframe. Where the Sprint Facilitator owns process and the
five Tier-1 agents own perspective, the Content Agent owns **how everything is
said** β the actual words a human or end user will read. Its principles live in
`references/writing-guide.md`.
**Per-phase responsibility:**
| Phase | What the Content Agent does |
|---|---|
| Define | Sharpens the problem statement, principle wording, and success metrics for clarity β no jargon, no vague adjectives |
| Discover | Reviews insight and HMW phrasing β cuts hedging, keeps claims specific and sourced |
| Diverge | Writes the real copy that appears in every Phase 3 wireframe β never placeholder text |
| Converge | Writes the user story and the copy for every key screen in the flow |
| Build | Runs a full writing pass on all five Phase 5 outputs before they're shown to the human |
Run the Content Agent pass explicitly whenever user-facing copy is produced β not
just informally "write it well." Apply the checklist in
`references/writing-guide.md` before the copy is shown to the human.
**Optional enhancement β external UX-writing skill.** If the open-source
[`ux-writing-skill`](https://github.com/content-designer/ux-writing-skill) (MIT)
is installed in the session, invoke it for the Content Agent pass β it adds a
UX-copy pattern library (buttons, errors, empty states, forms, notifications),
voice/tone guidance, and a scored usability checklist. When it isn't installed,
fall back to the inline principles in `references/writing-guide.md` β those remain
the baseline and the sprint never depends on the external skill.
---
## Agent Tiers
> Core agents shape **what to build and how it should feel.**
> Secondary agents shape **whether and how to ship it.**
### Tier 1 β Core Agents (Always Active)
| Agent | Cross-functional role | Focus |
|---|---|---|
| π§βπ» **User Agent** | UX Researcher | Pain points, needs, mental models |
| πΌ **Business Agent** | Product Manager | Value, revenue, impact, strategy |
| π **Data Agent** | Data Analyst / Growth | Patterns, trends, signals, evidence |
| ποΈ **System Agent** | Engineer / Architect | Gaps, duplication, consistency, architecture |
| π¨ **Design Agent** | Designer / Design Lead | Quality, distinctiveness, craft, design taste |
### Tier 2 β Secondary Agents (Situational)
| Agent | Cross-functional role | Activate when... |
|---|---|---|
| π£ **Marketing Agent** | Marketing / GTM | New market entry, launch, positioning |
| βοΈ **Legal Agent** | Legal / Compliance | Sensitive data, regulated industry, privacy |
| π§ **Customer Success Agent** | Customer Success / Support | Post-launch iteration, field signal |
| π° **Finance Agent** | Finance | Build vs. buy, pricing, margin, headcount |
| π― **Executive Agent** | Executive / Stakeholder | Strategic pivot, board narrative, vision alignment |
| π¬ **Quality Agent** | QA / Quality Engineer | Reliability-critical features, edge cases, failure modes |
**Secondary agent trigger signals:**
| Signal in problem statement | Suggest |
|---|---|
| "launch", "positioning", "go-to-market" | π£ Marketing Agent |
| "GDPR", "HIPAA", "privacy", "compliance" | βοΈ Legal Agent |
| "complaints", "churn", "post-launch", "support" | π§ Customer Success Agent |
| "cost", "pricing", "budget", "build vs. buy" | π° Finance Agent |
| "strategy", "vision", "board", "pivot" | π― Executive Agent |
| "reliability", "edge cases", "failure", "critical" | π¬ Quality Agent |
---
## Source Tracing (All Agents)
Every insight, recommendation, and HMW must cite its source. This ensures quality and
reliability of input data β humans can trace back, sanity check, and judge signal strength.
**Per-agent output format:**
```
π§βπ» USER AGENT
ββββββββββββββββ
Key Insights:
β’ [Insight]
β³ Source: [e.g. User interview β Sarah K., March 2025]
β’ [Insight]
β³ Source: [e.g. Support tickets β 47 mentions in Q1 2025]
Pain Points:
β’ [Pain point]
β³ Source: [e.g. Usability study β Task 3 failure rate 68%]
HMW Questions:
β’ HMW [question]?
β³ Derived from: [which insight(s) above]
```
**When source is unknown or assumed:**
```
β’ [Insight]
β³ Source: β οΈ Assumed β no direct source available.
Recommend validating before converging.
```
**Typical sources by agent:**
| Agent | Typical sources |
|---|---|
| π§βπ» User | User interviews, usability studies, support tickets, NPS verbatims, session recordings |
| πΌ Business | OKRs, roadmap docs, revenue data, competitive analysis, stakeholder interviews |
| π Data | Analytics dashboards, A/B tests, funnel data, retention cohorts, surveys |
| ποΈ System | Tech debt logs, architecture docs, Jira/Linear, incident reports, code reviews |
| π¨ Design | Design audits, competitor UX analysis, taste references, design principle violations |
**Source quality levels** β Facilitator flags these in gap detection:
- π’ **Strong** β direct evidence, recent, high sample size
- π‘ **Moderate** β indirect or small sample, still directionally useful
- π΄ **Weak** β assumed, anecdotal, or outdated β flag before converging
---
## 5-Phase Sprint Model
### π Phase 1 β Define (Human-led)
**Goal:** Set the foundation in as few steps as possible. Two required questions,
one optional. No agents run until this is complete.
π₯ Stakeholder input gate β open first.
---
**Question 1 β Project type (required)**
The single most important question. The answer configures the entire sprint β
which agents lead, how deep each phase goes, what outputs to emphasize.
```
π§ What kind of project is this?
β π Vision
Exploring a future direction or possibility space.
Breadth and provocation over precision.
β π± 0β1 Exploration
Building something new from scratch.
User truth and market fit before anything else.
β π XFN Alignment
Complex stakeholder problem. Need shared
principles before any direction is set.
β π Redesign
Existing product or system. Constraints
and audit before new ideas.
Not sure? Ask: "Is this about discovering what
to build β or how to build something already committed to?"
```
**Consequence shown immediately after selection** β before the human moves to Q2,
the Facilitator confirms what changes:
```
You selected [type] β
Agents leading: [names]
Phase emphasis: [what gets more depth]
Primary outputs: [which Phase 5 outputs]
[Any additional steps added e.g. audit for Redesign]
```
This teaches the human what they configured before they continue β not after.
**What changes based on project type:**
π **Vision**
- Agents leading: Design Β· Business Β· Marketing Β· Executive
- Phase 2: broad market and trend scan, less depth on constraints
- Phase 3 (Diverge): agents explicitly push past the obvious β one "provocative" idea
required per agent alongside the safe direction
- Phase 5 emphasis: Executive Brief + Sprint Recap Deck for leadership alignment
π± **0β1 Exploration**
- Agents leading: User Β· Data Β· Business
- Phase 2: deepened β User Agent prioritizes primary research gaps, Data Agent
runs market sizing pass. Sprint does not move to Phase 3 until user is understood.
- Phase 3 (Diverge): controlled mode by default β ideas per prioritized problem space
- Phase 5 emphasis: Flow + Key Screens + Build Brief
π **XFN Alignment**
- Agents leading: Executive Β· Business Β· Customer Success
- Phase 1 sensemaking expands: Facilitator maps conflicting stakeholder mental
models before principles are set. Principles become the negotiated output of
Phase 1, not just the designer's preferences.
- Phase 4: Facilitator synthesis explicitly maps which ideas have XFN support
and which will face resistance
- Phase 5 emphasis: Sprint Recap Deck for stakeholders is the primary output
π **Redesign**
- Agents leading: System Β· User Β· Design
- Audit step inserted between Phase 1 and Phase 2: System Agent inventories
what exists, what works, and what constraints are non-negotiable before
any new ideas are generated
- Phase 3 (Diverge): System Agent runs a "what to keep" pass alongside idea generation
- Phase 5 emphasis: Design System + Flow (new vs retained components mapped)
---
**Question 2 β Timebox (required, can say no)**
```
π§ Do you want to timebox this sprint?
Recommended by mode:
Full sprint: 1 hour (compressed) Β· 3 days (standard) Β· 1 week (deep)
Express sprint: under 30 minutes
Single phase: 20β60 min per phase
β Yes β pick a mode or give me your deadline
β No β I'll move at my own pace
```
If yes: Facilitator announces target at each phase start and flags if running long.
Never cuts off β always the human's call to wrap or continue.
Agent output depth scales automatically to available time β **a 1-hour timebox
still runs all 5 phases**; see Sprint Modes below for how depth scales instead of
phases being dropped.
---
**Agent roster (optional β skip to move straight to inputs)**
```
π§ Want to customize the agent roster?
Your project type pre-selects the recommended agents.
You can activate secondary agents, deactivate any core
agent, or add a custom agent with its own lens.
β Show me the roster and let me adjust
β Skip β use the recommended agents
```
If shown: human can activate Tier 2 agents, deactivate Tier 1, or add a custom
agent by describing its role and lens. Facilitator generates the reasoning pattern.
Custom agent examples: Sales Β· Localization Β· Accessibility Β· Security Β· Strategy.
Deactivating a core agent triggers a consequence flag before confirming.
---
**Step 1a β Sprint Review page (mandatory)**
The moment the timebox is confirmed β before Goal, Problem statement, or Constraints
are gathered below β the Facilitator creates a shareable **Vibe page** (built via the
Artifact tool) that tracks the sprint live from here through the final recommendation.
Share the link with the human immediately:
```
π Your Sprint Review page is live: [link]
It updates at every Decision Point β share it with
stakeholders any time during the sprint, not just at the end.
```
See "Sprint Review Page Template" near the end of this file for the page structure.
This is a live shareable link, not a local file β never a static sprint-review.html
the human has to find and open manually.
---
Gather from the human:
- **Goal**: What are we trying to achieve?
- **Problem statement**: What problem, for whom? (can be rough at this stage)
- **Constraints**: What do we know already?
- **Success metrics** (required): How will we know this worked? Confirmed alongside
principles at Decision Point 1 β not an afterthought bolted on later.
Gap-check: if goal, problem statement, or success metrics are missing entirely, flag before continuing.
---
**Step 2 β Problem sensemaking (Facilitator-led)**
Before principles are set, the Facilitator helps the human sharpen the problem statement.
Raw inputs are rarely sprint-ready β they are often symptoms, too broad, assumption-laden,
or missing a specific user. This step does that work explicitly.
The Facilitator checks for four patterns and surfaces each one it finds:
**Symptom vs root cause**
"AI workflow is chaotic" is a symptom.
"Designers have no judgment framework for AI-generated output" is a root cause.
β Facilitator surfaces the distinction and asks: which level are we solving at?
**Too broad**
If the problem contains more than one distinct user + one distinct pain point, it needs
to be split or scoped. Multiple problems in one sprint = convergence on nothing.
β Facilitator identifies overlapping problems and asks the human to pick one for this sprint.
**Assumption embedded**
"We need a better AI tool" is a solution disguised as a problem.
β Facilitator surfaces the assumption and reframes: "What would the right solution actually
need to do for the user?"
**Missing who**
A problem without a specific user is a feature request, not a design problem.
β Facilitator prompts: "Who specifically feels this most acutely?"
Sensemaking output format:
```
π§ FACILITATOR β PROBLEM SENSEMAKING
ββββββββββββββββββββββββββββββββββββββββββββββββββββ
Raw input: [what the human provided]
Patterns detected:
β οΈ [Pattern type]: [what was found]
β [Reframe or clarifying question]
β οΈ [Pattern type]: [what was found]
β [Reframe or clarifying question]
Sharpened problem statement:
"[Specific user] struggles with [specific pain]
in the context of [specific situation],
which leads to [specific consequence]."
Confirm this before we set principles?
ββββββββββββββββββββββββββββββββββββββββββββββββββββ
```
Human confirms the sharpened problem statement before moving to Step 3.
If the human wants to adjust β iterate until confirmed.
**Why this matters:**
A fuzzy problem statement produces fuzzy principles, fuzzy HMWs, and fuzzy ideas.
The sensemaking step is the single highest-leverage moment in the entire sprint β
5 minutes here saves 2 hours of misdirected exploration in Phase 3.
---
**Step 3 β Principle Setting (Facilitator-led)**
Now that the problem statement is confirmed, Facilitator runs **Responsibility 2:
Principle Setting** β generates 3β6 product principles and 3β6 design principles
drawn from the tensions in the confirmed problem statement, and confirms the
success metrics gathered above alongside them β principles and metrics are
confirmed together, in the same step, not principles-then-metrics-later.
Human picks 2β4 from each set, edits wording, or adds their own. Success metrics
are confirmed or refined at the same time.
---
**Step 4 β Secondary agent recommendation**
Scan confirmed problem statement for secondary agent trigger signals.
Recommend relevant secondary agents for human to approve.
**Step 5 β Render Phase 1 board using show_widget**
After all four steps are complete, the Facilitator renders the Phase 1 outcome
as an interactive sprint board using show_widget β before presenting Decision
Point 1 for confirmation.
The board contains four sections:
**Section A β Sprint configuration stickies (yellow)**
Three sticky notes: project type, timebox, agent roster.
Each shows the selected value and a consequence line.
Expand trigger: "See what this configures" β reveals timebox bar +
phase-by-phase targets + what changes based on project type.
**Section B β Problem sensemaking affinity map**
Two clusters side by side:
- "Root cause" cluster: who, pain, context, effect β blue stickies
- "Scoped out" cluster: symptoms and assumptions that were parked β same blue
but at 50% opacity with a strikethrough label
Expand trigger: "See sharpened problem statement" β reveals the full confirmed
one-sentence problem statement and why the scoped-out items were deferred.
**Section C β Principles wall**
Product principles (P1βP4) in indigo left border.
Design principles (D1βD3) in pink left border.
Success metrics (M1βM3) in teal left border β same card treatment.
One card per principle/metric: code + name only in default view.
An empty "+ Add principle" card at the end β tapping sends a prompt to add one.
Expand trigger: "See what each principle means" β reveals full descriptions.
**Section D β DP1 confirm card**
Single summary bar showing:
- Sprint config (project type Β· timebox Β· agents) with a coloured icon
- Problem statement (one line) with a coloured icon
- Principles count confirmed with a coloured icon
- Success metrics confirmed with a coloured icon
- Secondary agents activated with a coloured icon
Confirm button (sendPrompt "Confirm β proceed to Phase 2 Discover")
Edit button (sendPrompt "I want to edit something in Phase 1")
**Collapsed detail rule:**
Every section has an expand trigger that reveals detail on tap.
Default view is visual and scannable. No walls of text in the default state.
The human should be able to read the entire board in under 30 seconds.
βΈοΈ **Decision Point 1:** Human confirms sharpened problem statement + principles
+ success metrics + approved secondary agents via the visual board confirm card.
**Brief auto-update:** Facilitator writes DP1 entry to sprint-brief.md:
```
[DP1 Β· human] Confirmed problem statement Β· principles P[N]+D[N] Β· metrics M[N] Β· agents [list]
```
---
### π Phase 2 β Discover (Agent-executed)
**Goal:** Research and understand the problem, including what's already shipped in
this space. Identify and prioritize opportunity areas.
π₯ Stakeholder input gate β open before agents run.
**Step 1 β Competitive Analysis (Design Agent + Business Agent, mandatory):**
Before synthesizing insights, find out what competitors have already shipped in this
problem space. Skipping this risks reinventing β or shipping a worse version of β
something that already exists.
1. Identify 3β6 relevant competitor or adjacent products for the problem space.
2. Pull real screens and flows for each using the **Mobbin MCP** tools β `search_apps`,
`get_app_screens`, `get_app_flows`, and `search_screens` / `search_flows` for
pattern-specific searches. See Sprint Toolkit above for the full tool list.
3. If the Mobbin MCP isn't connected in this session, fall back to `WebFetch` against
mobbin.com search results, and flag to the human that fallback results may be
lower-fidelity or incomplete.
4. **Screenshots are required, not descriptions.** Never write "Competitor X has a
clean onboarding flow" without an actual retrieved screenshot backing it. A text
summary with no visual evidence is not evidence β mark it β οΈ Assumed instead of
presenting it as a finding.
5. Render retrieved screenshots directly in the Phase 2 board (Section A below) so
the human sees the actual competitive landscape, not a paraphrase of it.
```
π COMPETITIVE ANALYSIS
ββββββββββββββββββββββββββββββββββββββββ
[Competitor/Product]: [what they do in this space]
β³ Screenshot: [retrieved image β via Mobbin MCP or WebFetch fallback]
β³ Pattern: [the specific interaction or design pattern shown]
β³ Implication: [gap, bar to clear, or pattern to avoid copying]
```
Gap-check: if no real evidence could be retrieved for a claimed competitor pattern,
flag it β do not carry an unverified assumption into insight gathering.
**Step 2 β Insight Gathering:**
All core agents (+ approved secondary agents) synthesize insights through their lens.
All insights must include source citations and source quality rating.
β See `references/agent-prompts.md` for each agent's reasoning pattern.
Gap-check: flag missing input sources before proceeding. Rate overall data quality.
**Step 3 β HMW Generation:**
Agents generate "How Might We" questions derived from their insights.
Each HMW traces back to the insight(s) that generated it.
Facilitator clusters into 3β5 themes.
π Content Agent reviews HMW phrasing for hedging and vagueness before clustering locks.
**Step 4 β Prioritization:**
Agents score themes against: impact, feasibility, principle alignment.
Facilitator presents scoring with rationale and source confidence level.
**Step 5 β Render insight board using show_widget**
After all agents complete their insight gathering, the Facilitator renders the
Phase 2 output as an interactive sprint board using show_widget. This replaces
walls of text with a visual research board the human can scan and interact with.
The board contains four sections:
**Section A β Competitive evidence strip**
The screenshots retrieved in Step 1 render first, as a horizontal strip of actual
screenshots with source app name and pattern label beneath each β before the
insight stickies below.
**Section B β Insight sticky notes (with tappable source links)**
Each insight is rendered as a colour-coded sticky note:
- Yellow = User Agent insights
- Blue = Business Agent insights
- Green = Data Agent insights
- Orange = System Agent insights
- Pink = Customer Success / secondary agent insights
Each sticky note must include:
- Agent label + insight category (e.g. "User Β· silent drift")
- Insight text (2 sentences max)
- Source quality dot: π’ strong / π‘ moderate / π΄ weak
- Tappable source link using `<a href="[real URL]" target="_blank">` β not a
placeholder. If the source has a real URL, link to it directly. If it's an
assumed or inferred source, omit the link and show "β οΈ Assumed" instead.
Sticky note footer format:
```html
<div class="sticky-footer">
<span class="sticky-source">
<span class="src-dot src-[green|yellow|red]"></span>[Quality]
</span>
<a class="sticky-link" href="[real URL]" target="_blank">[Short source name] β</a>
</div>
```
**Section C β HMW affinity map**
HMW questions grouped into clusters using dashed-border containers.
Each cluster has a theme label. Stickies inside match the agent colour of
the agent who generated the HMW.
**Section D β Priority themes**
Ranked themes shown as horizontal score bars with agent attribution.
Expand/collapse for scoring breakdown detail.
The board ends with a DP2 confirm card β summary of themes + Confirm/Edit buttons
that use sendPrompt() to proceed or go back.
**Collapsed detail rule:**
Every section has an expand trigger ("See full notes", "See scoring breakdown")
that reveals detail on tap. The default view is visual and scannable β not a
wall of text. Detail is available but not forced.
βΈοΈ **Decision Point 2:** Human reviews the visual board, confirms competitive
findings, themes, and priority order, or redirects before Phase 3 begins.
---
### π Phase 3 β Diverge (Two-round ideation with agent voting)
**Goal:** Generate, sketch, and vote on ideas in two rounds β forcing genuine divergence
before convergence. Every idea must be made visible through a low-fidelity wireframe.
Human sees and reacts at every checkpoint.
π₯ Stakeholder input gate β open before agents run.
---
**Round 1 β First burst of ideas**
**Step 1A β Generate 8 ideas:**
All core agents (+ active secondary agents) collaborate to generate exactly 8 ideas
across the top-rated opportunity areas from Phase 2.
Ideas are distributed across agents by lens β not one agent per idea, but each idea
tagged with which agent perspective it comes from.
Each idea format:
```
IDEA [1β8]: [Name]
Agent lens: [which agent perspective drives this]
What it is: [1-2 sentences β concrete, not abstract]
Why it matters: [which principle it serves]
Opportunity area: [which Phase 2 theme it addresses]
Source: [insight it came from, with source citation]
```
**Step 1B β Wireframe all 8 ideas (mandatory):**
Every idea must be made visible as a wireframe β this is not optional and is never
satisfied by a text description alone. Load the design skill stack (see Sprint
Toolkit β Anthropic Frontend Design, Impeccable, UI/UX Pro Max, Vercel Web Design
Guidelines, Vercel React Best Practices; invoke whichever are installed), then
generate a mid-fidelity wireframe for each of the 8 ideas via show_widget.
These are rough, fast, intentional β not polished UI. Goal: make thinking visible,
not make it pretty. If wireframes cannot be rendered for any reason, escalate to
the human before proceeding β never substitute a written description for the
missing wireframe.
**Wireframe rendering rules:**
Each wireframe follows this spec: a browser-framed HTML mockup β the same style as
the key screens in Phase 5, but intentionally rougher. Every wireframe must show:
- Browser chrome with a descriptive URL (e.g. "sprint.app / contention")
- One key screen or moment β the most critical interaction
- Real copy, not placeholder text ("Lorem ipsum" is forbidden)
- 2β3 floating annotation labels pointing to the key design decisions
- The agent lens that inspired the idea shown as a small badge
**Sketch vs polished distinction:**
- Sketch (Phase 3): rough layouts, minimal colour, boxes for images, key copy only
- Polished (Phase 5 Output 3): full design system, real visual language, all states
Present all 8 wireframes in a 2Γ4 grid using show_widget so the human can
compare ideas visually side by side. Below each wireframe: idea name, agent lens,
and vote count (empty at this stage β filled after voting).
```
βββββββββββββββββββββββββββββββββββββββββββββββββββ
β #1 [Name] #2 [Name] β
β [wireframe] [wireframe] β
β [Agent] Β· 0 votes [Agent] Β· 0 votes β
β β
β #3 [Name] #4 [Name] β
β [wireframe] [wireframe] β
β ... ... β
βββββββββββββββββββββββββββββββββββββββββββββββββββ
```
After voting, re-render the grid with vote counts filled in and top-voted
ideas highlighted (indigo border + star badge).
**Step 1C β Agent voting (Round 1):**
Each agent votes for their top 3 ideas from the 8 β cannot vote for their own idea.
Agents must give a one-line reason for each vote, tied to a principle.
Vote format:
```
π³οΈ ROUND 1 VOTING
ββββββββββββββββββββββββββββββββββββββββββββββββ
π§βπ» User Agent votes: #[N] β [reason] Β· #[N] β [reason] Β· #[N] β [reason]
πΌ Business Agent votes: #[N] β [reason] Β· #[N] β [reason] Β· #[N] β [reason]
π Data Agent votes: #[N] β [reason] Β· #[N] β [reason] Β· #[N] β [reason]
ποΈ System Agent votes: #[N] β [reason] Β· #[N] β [reason] Β· #[N] β [reason]
π¨ Design Agent votes: #[N] β [reason] Β· #[N] β [reason] Β· #[N] β [reason]
[Secondary agents if active]
TALLY:
#1 [Idea name]: β β β β β (4 votes)
#2 [Idea name]: β β β ββ (3 votes)
...
TOP 3 IDEAS: #[N], #[N], #[N]
SURPRISE: #[N] β voted for by unexpected agents
TENSION: #[N] vs #[N] β agents disagreed most here
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
βΈοΈ **Human checkpoint 1:** Review 8 ideas, wireframes, and vote results.
β React to what you see. Any ideas to kill? Any that surprised you?
β Proceed to Round 2, or adjust direction first.
---
**Round 2 β Inspired by each other**
**Step 2A β Generate 8 new ideas:**
Agents now generate a second set of 8 ideas β this time inspired by what they saw
in Round 1. Cross-pollination is the goal: an idea from the User Agent might inspire
the System Agent to think differently. Ideas can combine, evolve, or deliberately
challenge the Round 1 top-voted concepts.
Rules for Round 2:
- Cannot repeat a Round 1 idea verbatim
- Must reference at least one Round 1 idea as inspiration
- At least 2 ideas must combine perspectives from different agents
- At least 1 idea must challenge the Round 1 top vote β "what if the opposite were true?"
Each idea format: same as Round 1, plus:
```
Inspired by: [Round 1 idea # and what it sparked]
```
**Step 2B β Wireframe all 8 new ideas (mandatory):**
Wireframe every Round 2 idea β the same non-negotiable requirement as Round 1,
with the design skill stack loaded. Same format as Round 1 β rough browser-framed
mockup, real copy, 2β3 annotations. Cross-pollination should be visible in the
wireframe β if idea #9 combines #3 and #7 from Round 1, the wireframe should
visually reference both.
Present all 8 Round 2 wireframes in a 2Γ4 grid, same style as Round 1.
After Round 2 voting, render a combined view: all 16 ideas with vote counts,
clearly marking which ideas climbed, held, or fell between rounds.
**Carry-forward highlight:**
After DP3 confirmation, render the 4 selected ideas as a focused 2Γ2 grid
with a "Carrying into Phase 4" label β this becomes the handoff board for convergence.
**Step 2C β Agent voting (Round 2):**
Agents vote across the full pool β Round 1 + Round 2 ideas combined.
Each agent picks their top 3 from the full 16.
Vote format: same as Round 1, but labeled Round 2 Voting.
Tally shows movement: did Round 2 ideas displace Round 1 leaders?
```
π³οΈ ROUND 2 VOTING (full pool of 16)
ββββββββββββββββββββββββββββββββββββββββββββββββ
[same format as Round 1]
MOVEMENT:
β Ideas that climbed: [which Round 2 ideas overtook Round 1 leaders]
β Ideas that held: [which Round 1 ideas survived both rounds]
β Ideas that fell: [which Round 1 top-voted ideas lost support]
FINAL TOP IDEAS GOING INTO CONVERGENCE:
#[N], #[N], #[N] β carried forward to Phase 4
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
βΈοΈ **Decision Point 3:** Human reviews Round 2 ideas, wireframes, and final vote tally.
β Confirm which ideas carry forward into Phase 4.
β Can add ideas to carry forward, remove any, or flag a tension to resolve first.
---
### π Phase 4 β Converge (Agents as a team)
**Goal:** Take the top-voted ideas from Phase 3 and converge them into one group solution β
designed around a real user story, connected into flows, tested against real behavior.
π₯ Stakeholder input gate β open before convergence runs.
---
**Step 1 β Identify up-voted concepts and patterns:**
Facilitator reads across all Phase 3 voting to surface:
- Which ideas won across multiple rounds (durable signal)
- Which patterns appear across multiple ideas (recurring motif worth keeping)
- Which tensions between top ideas need resolving before combining
```
π§ PATTERN RECOGNITION
ββββββββββββββββββββββββββββββββββββββββββββββββ
DURABLE IDEAS (top-voted across both rounds):
β’ [Idea] β voted by [agents], survived [rounds]
β’ [Idea] β voted by [agents], survived [rounds]
RECURRING PATTERNS (appear in multiple ideas):
β’ [Pattern] β seen in ideas #[N], #[N], #[N]
β Worth making a core part of the group solution
TENSIONS TO RESOLVE:
β’ [Idea A] vs [Idea B] β incompatible because [reason]
β Group must pick one or find a synthesis
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
---
**Step 2 β Agents work as a team to build one group solution:**
This is the critical shift. Agents stop competing and start collaborating.
Every agent contributes their perspective to ONE unified solution.
The group solution must:
- Incorporate the top-voted patterns from both rounds
- Resolve the tensions identified above
- Pass all confirmed principles
- Be expressed as a named concept with a clear point of view
Group solution format:
```
π€ GROUP SOLUTION: [Name]
ββββββββββββββββββββββββββββββββββββββββββββββββ
Concept: [1 sentence β what this is]
Point of view: [what makes this distinctively ours]
Each agent's contribution:
π§βπ» User: [what this does for the user]
πΌ Business: [how this creates value]
π Data: [what signal this uses or generates]
ποΈ System: [how this is structured underneath]
π¨ Design: [what makes this visually and experientially distinctive]
[Secondary agents if active]
Principles passed:
P2 β /β οΈ P3 β /β οΈ P5 β /β οΈ P6 β /β οΈ
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
---
**Step 3 β Think like a real user. Build the story and flow:**
Before wireframing the group solution, agents run a story test:
write the experience as a real user living through it β not a feature description,
a human moment. If the story doesn't feel natural, the solution has a gap.
Story format:
```
π€ USER STORY
ββββββββββββββββββββββββββββββββββββββββββββββββ
Character: [Name], [brief description β real person, not a persona archetype]
Situation: [specific moment β not "a user wants to..." but "It's Thursday evening
and Maya just realized her best friend's birthday is in 4 days..."]
Step by step β what Maya actually does:
1. [She opens the app because...]
2. [She sees... and feels...]
3. [She does... and the product responds with...]
4. [The moment of truth: does the solution solve the problem?]
5. [How does she feel at the end?]
Does the solution solve Maya's problem? β / β οΈ / β
Where does the story break or feel unnatural? [specific moment]
What's missing from the solution to make this story work?
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
If the story breaks β agents identify the gap and patch the solution before flows.
If the story holds β proceed to storyboarding.
---
**Step 3b β Storyboard first, Mobbin first (mandatory before final screens):**
Before designing final hi-fidelity screens, storyboard the journey as a sequence
of rough panels β comic-strip style β mapped 1:1 to the user story steps from
Step 3. This locks the *sequence and moments* before locking pixels.
For each storyboard panel:
1. **Storyboard it** β one panel per story beat: who's on screen, what state
they're in, what they see and do. Simple boxes with a 1-line caption β not a
wireframe yet.
2. **Mobbin it** β before designing the panel, search Mobbin (`search_screens` /
`search_flows`, filtered to the relevant pattern β e.g. "onboarding interview,"
"comparison result") for real shipped examples of this exact moment. Pull 2β3
reference screens per panel.
3. **Design it** β only once the panel and its references are set, produce the
final hi-fidelity screen for that panel with the design skill stack loaded
(see Sprint Toolkit), using the Mobbin references as grounding β not to copy,
but to know the bar and avoid reinventing a solved pattern.
Storyboard panel format:
```
PANEL [N]: [story beat this maps to]
Who/state: [character, emotional state]
Sees/does: [what happens in this panel]
Mobbin references: [2-3 screens pulled, with source app name + link]
β Design once panel is confirmed
```
Skip storyboarding only if the flow has 2 or fewer screens β for anything longer,
storyboard-first is mandatory before final screens are built.
---
**Step 4 β Connect concepts into flows using show_widget:**
Agents produce a connected flow β not isolated screens, but the full journey
from trigger to resolution, using the group solution.
**Flow rendering rules:**
The Facilitator renders the connected flow as an interactive show_widget board:
**Section A β Flow map (top strip)**
A horizontal sequence of screen thumbnails connected by arrows showing:
- Each screen as a small labelled node
- The transition trigger between screens (tap, scroll, AI response, time)
- The key AI moment or decision point marked with a distinct badge
- Any branches or edge cases shown as secondary paths below the main flow
**Section B β Key screens (main section)**
For each major screen in the flow, produce a full browser-framed wireframe β
mandatory, not optional, same as Phase 3, with the design skill stack loaded β
using the storyboard panel and Mobbin references from Step 3b as grounding. Same
quality level as Phase 3 wireframes but showing the actual group solution, not a
rough concept. Render via show_widget. Each screen includes:
- Browser chrome with real URL
- Real copy reflecting the confirmed group solution name and tone
- 3β4 annotation labels per screen: what user sees, what they do,
what AI does, and which principle this moment expresses
- Principle badge (e.g. "D1 Β· visible reasoning") on the annotation
that is most clearly expressing a design principle
**Section C β Story test overlay (collapsible)**
An expand trigger reveals the user story mapped onto the flow:
- Each story step pinned to the relevant screen in the flow
- Emotional state shown at each moment (the "feels like" label from the story)
- Gap flag if a story step has no screen β shown as a red dotted outline
**Section D β DP4 + DP5 confirm card**
Facilitator synthesis summary + human decision card at the bottom.
Two buttons: "Approve and proceed to Phase 5" and "Flag a gap and rework."
Flow structure text format (for reference, rendered visually):
```
TRIGGER β [what causes the user to open the product]
β
SCREEN 1: [name] β [what user sees and does]
β
SCREEN 2: [name] β [what user sees and does]
β
[key AI moment / decision point]
β
SCREEN 3: [name] β [resolution or next step]
β
OUTCOME: [how the user feels β and what the product learned]
```
Facilitator then runs final checks:
```
π§ FACILITATOR SYNTHESIS
ββββββββββββββββββββββββββββββββββββββββββββββββ
[See full format in Responsibility 3 above]
STORY TEST RESULT: β holds / β οΈ gaps found / β needs rework
FLOW COHERENCE:
β’ Does each screen lead naturally to the next?
β’ Is the AI's role clear at every step?
β’ Where might a real user hesitate or drop off?
RECOMMENDED DIRECTION: [Group solution name]
ALTERNATIVE: [if story test revealed a fallback]
THE DECISION ONLY YOU CAN MAKE: [human judgment call]
SUGGESTED NEXT STEP: [one action, one owner, one deadline]
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
βΈοΈ **Decision Point 4:** Human reviews group solution, user story, and connected flow.
β Does the story feel true? Does the flow solve the problem?
β Approve for Phase 5, or identify which part of the story breaks and needs fixing.
βΈοΈ **Decision Point 5 (Final):** Human makes the call.
β See `references/convergence-guide.md` for full scoring and recommendation formats.
---
### π Phase 5 β Build (Agent-executed)
**Goal:** Translate sprint direction into four artifacts β for the team, leadership, design
and engineering, and stakeholders who weren't in the room.
Phase 3 wireframes are the foundation β Phase 5 refines and finalizes them, it does not
start from scratch.
π₯ Stakeholder input gate β open before build artifacts are created.
**Wireframe continuity rule:**
The preferred sketch selected at Decision Point 3 anchors Phase 5 design work.
Design Agent refines Phase 3 sketches into final-fidelity screens β mandatory,
same as Phase 3/4, with the design skill stack loaded β then renders via
show_widget. New screens are added only where Phase 3 sketches left gaps.
**Rendering approach for Output 3 and Output 4:**
All key screens are rendered as high-fidelity HTML mockups using show_widget β not
placeholder boxes. Every screen must show real UI: actual conversation flows, real copy,
real interaction states, real visual language from the confirmed design principles.
Each rendered screen includes:
- Browser chrome with URL
- Actual content (not "Lorem ipsum" or grey boxes)
- Annotations explaining the key design decision on that screen
**Content pass:** the Content Agent runs a full writing pass on all five outputs
before any of them is shown to the human β see Content Agent under Meta-Roles.
**Optional: export to Google Workspace.** Ask the human once, before producing
outputs: "Want these exported to Google Docs/Slides as we go β Build Brief and
Executive Brief to Docs, Sprint Recap Deck to Slides? I can help you connect the
Google Workspace MCP if you'd like." Only proceed if the human says yes β never
install or connect it automatically. If declined, or the MCP isn't available, all
outputs are still produced natively (show_widget / pptxgenjs / markdown) with no
loss of functionality.
**If the human says yes and the MCP isn't connected yet**, guide the install β
don't just say "connect it and come back":
1. **Check first** β Google Docs/Drive/Slides/Calendar tools may already be
connected in the session (look for Google Workspace tools in the available
tool list, including deferred tools). If present, skip install entirely.
2. **Connector directory (Claude app / Cowork / claude.ai):** if a connector
directory or MCP registry tool is available in the session, use it to surface
the Google Workspace / Google Drive / Google Slides connector so the human can
approve it in the app's own UI β the OAuth consent screen is theirs to click
through, never Claude's.
3. **Claude Code CLI:** tell the human to run `claude mcp add` for the Google
Workspace MCP server they prefer, or point them to Settings β Connectors.
Claude may draft the exact command, but the human runs or approves it.
4. **Auth stays with the human.** Never enter Google credentials, click through
OAuth consent, or approve scopes on the human's behalf β pause and hand off
whenever a sign-in screen appears.
5. **Resume gracefully.** Once connected, verify with a lightweight read-only
call (e.g. list recent files) before exporting. If the human abandons the
install midway, fall back to native outputs without re-asking.
Five outputs are produced in priority order: Output 4 β Output 3 β Output 1 β Output 2 β Output 5 (auto, 2 days later)
---
**Output 4: Sprint Recap Deck** β produce first
Owned by Sprint Facilitator.
Audience: stakeholders who weren't in the room.
Purpose: tell the story of how you got to the decision, not just what the decision was.
Every slide uses rendered visuals β no placeholder boxes. Produce using show_widget or
a PPTX generation script (pptxgenjs). Render key screens as actual HTML/CSS mockups and
embed them into the relevant slides.
Slide structure:
```
Slide 1 β Cover (product name, sprint type, tagline)
Slide 2 β The problem (Phase 1: problem statement + user types + key stat)
Slide 3 β Confirmed principles (product principles + design principles, rendered)
Slide 4 β What we learned (Phase 2: 6 themes grid with agent attribution, rendered)
Slide 5 β Two rounds of ideation (Phase 3: all 16 ideas, vote tallies, what carried forward)
Slide 6 β The group solution (Phase 4: concept name + all agent contributions, rendered)
Slide 7 β Story test (Phase 4: character + 6-step journey + verdict, rendered)
Slide 8 β Key screens 1+2 (actual high-fidelity UI, dark background)
Slide 9 β Key screens 3+4 (actual high-fidelity UI, dark background)
Slide 10 β End-to-end flow (4 screens in sequence + story steps mapped below)
Slide 11 β The decision (recommendation + the human call + next step)
Slide 12 β What we're building (12-week plan + metrics + out-of-scope)
Slide 13 β Next step (one action, one owner, one deadline β close slide)
```
**Deck rendering rules:**
- Every slide that describes a feature or screen must show actual rendered UI
- Dark slides (#2C2C2C background) for group solution and key screens β creates contrast
- Warm off-white (#FAF7F2) for all other slides
- Terracotta (#C4785A) as the accent β headers, highlights, CTAs
- Georgia serif for titles, Calibri/Arial for body
- No placeholder boxes, no wireframe-style grey rectangles in the final deck
---
**Output 3: End-to-end flow + key screens** β produce second
Owned by Design Agent + System Agent.
Audience: the team building the product (design + engineering).
Purpose: replace the design system doc β what the team actually needs is the flow and
screens, not a token list. They can extract tokens from the screens.
Render ALL key screens using show_widget as high-fidelity HTML mockups:
- Browser chrome with real URL
- Real copy, real interaction states, real visual language
- Fit/miss guide, conversation bubbles, relationship arc β all rendered properly
- Annotations per screen explaining the key design decision
Screen set (minimum for web app):
```
Screen 1: Homepage β the invitation (no account, one CTA)
Screen 2: Relationship interview β conversation, live tags
Screen 3: Warm reveal β one rec, fit/miss guide side by side
Screen 4: Account prompt β value shown before ask
Screen 5: Dashboard β relationship arc, circle view
Screen 6: Loop close β how did it land? one tap
```
Flow document structure:
```
FLOW: [trigger β screen 1 β screen 2 β ... β outcome]
Per screen:
- What the user sees
- What they do
- What the system does next
- Key design decision annotated
Edge cases:
- Returning user (known relationship β skip to recs)
- Second gift (profile already built)
- No profile yet (cold start)
```
---
**Output 1: Build Brief** β produce third
Owned by System Agent + Business Agent + Quality Agent (if active).
Audience: the team building the product.
```
ποΈ BUILD BRIEF
ββββββββββββββββββββββββββββββββββββββββββββββββ
What we're building: [chosen direction, 1-2 sentences]
Why: [rationale from sprint β connects to principles]
Owner: [team / person accountable]
Success metrics:
β’ [Metric 1 + target]
β’ [Metric 2 + target]
Scope (in):
β’ [what's included in v1]
Scope (out):
β’ [explicitly excluded to keep v1 focused]
Technical starting point: [System Agent input]
Acceptance criteria: [Quality Agent if activated]
Launch criteria: [Marketing Agent if activated]
Next action: [one specific thing, owner, deadline]
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
---
**Output 2: Executive Brief** β produce last
Owned by Executive Agent.
Audience: leadership, board, stakeholders.
Purpose: translate sprint output into the language leadership cares about β
vision alignment, strategic fit, business impact, resource ask, and risk.
This is NOT a decision-making document β it prepares the human to walk into
any exec conversation fully prepared.
```
π― EXECUTIVE BRIEF
ββββββββββββββββββββββββββββββββββββββββββββββββ
FOR: [Leadership / Board / Stakeholders]
RE: [What we're building β one line]
VISION & MISSION ALIGNMENT
β [How this connects to company vision]
β [How this advances the mission]
STRATEGIC FIT
β [How this maps to current company goals / OKRs]
β [What strategic bet this represents]
WHY NOW
β [Market signal, user signal, or competitive context]
WHAT WE'RE ASKING FOR
β Resources: [team, time, budget]
β Decision needed: [what leadership needs to approve]
EXPECTED IMPACT
β [Metric 1 + target]
β [Metric 2 + target]
RISK & MITIGATION
β [Key risk and how we're managing it]
WHAT HAPPENS IF WE DON'T DO THIS
β [Opportunity cost or competitive consequence]
DESIGN DIFFERENTIATION
β [What makes this visually and experientially
different from what's already in market]
ββββββββββββββββββββββββββββββββββββββββββββββββ
```
---
**Output 5: Decision Quality Score** β runs 2 days after sprint closes
Owned by Sprint Facilitator.
Audience: the designer who ran the sprint.
Purpose: close the learning loop. Turn the decision log into a learning system β not just
a record of what happened, but a signal about which phases and decision points need more
rigor next time.
**Trigger:** 2 days after sprint end. Facilitator sends a lightweight nudge (email or
in-tool message). Subject: "How did [project name] sprint decisions hold up?"
**Format:**
```
π§ DECISION QUALITY CHECK
ββββββββββββββββββββββββββββββββββββββββββββββββββββ
How did your sprint decisions hold up?
Rate each 1β3. Pattern appears immediately. Takes 60 seconds.
[DP1] [what was decided] [1 wrong] [2 ok] [3 right]
[DP2] [what was decided] [1 wrong] [2 ok] [3 right]
[DP3] [what was decided] [1 wrong] [2 ok] [3 right]
[DP4] [what was decided] [1 wrong] [2 ok] [3 right]
[DP5] [what was decided] [1 wrong] [2 ok] [3 right]
PATTERN:
[Which DP scored lowest and what it means for next sprint]
β [One specific suggestion for next sprint]
ββββββββββββββββββββββββββββββββββββββββββββββββββββ
```
**Pattern rules:**
- Show pattern immediately after rating β not stored in a dashboard nobody visits
- Pattern references the specific DP, not generic advice
- One suggestion only β concrete and actionable
- Store ratings in sprint-brief.md under Sprint history for cross-sprint comparison
**Cross-sprint learning** (after 3+ sprints):
```
## Sprint history (in sprint-brief.md)
Sprint 1: DP3 weakest (2/3) β carry-forward decisions
Sprint 2: DP2 weakest (2/3) β theme prioritization
Sprint 3: Pattern β Phase 3 β Phase 4 transition is consistently weakest
Suggestion: Add more time to Round 2 voting before carrying forward
```
**On web/async:** Re-engagement relies on email. Must be opt-in at account or session
creation. One-click link back to the quality check screen β no login wall on return.
βΈοΈ **Decision Point 6:** Human approves all five outputs before execution begins.
Decision Quality Score runs automatically 2 days later β no further approval needed.
---
## Sprint Modes
### Full Sprint
All 5 phases with all decision points and stakeholder input gates active.
Default for new products or complex, high-stakes decisions.
Recommended timebox: 1 hour (compressed) Β· 3 days (standard) Β· 1 week (deep/high-stakes)
**A 1-hour timebox still runs all 5 phases** β see the note under "Timebox
adjustment rules" below. Only agent output depth scales down, never the phase count.
### Express Sprint
Phase 1 (abbreviated) β Phase 2 (core agents only) β Phase 4 (convergence).
Decision Points 1 and 5 still required. Others skipped.
Recommended timebox: under 30 minutes, or by explicit human request to skip phases
regardless of time available.
**Do not default to Express Sprint just because the timebox is short.** A 1-hour
timebox runs the Full Sprint, not Express β see "Timebox adjustment rules" below.
### Single Agent Mode
Run one agent's analysis on demand. Output: Insights β HMWs β Implications.
Source citations still required. No decision points needed.
No timebox needed β run as long as useful.
### Timebox adjustment rules
When a timebox is active, the Facilitator scales agent output depth automatically.
**All rows below assume all 5 phases run β only depth changes, never the phase
count** β unless the human explicitly requests Express Sprint or Single Agent Mode.
| Time available | Agent output depth |
|---|---|
| < 30 min | Top 3 insights only Β· 2 HMWs Β· no source deep-dives |
| 30β60 min | 5 insights Β· 3β4 HMWs Β· key sources cited |
| 60β90 min | Full output Β· all sources Β· full HMW set |
| No limit | Full output + extended source research on request |
**1-hour timebox, concretely:** Phase 1 and 2 compress to top insights only, Phase 3
(Diverge) may reduce to a single round instead of two if time is tight β flag this
trade-off explicitly and get the human's confirmation before cutting it β Phase 4
stays fully intact (it's the convergence the sprint exists to reach), and Phase 5
produces Build Brief + key screens first, with Executive Brief, Deck, and the
Decision Quality Score following async if time runs out.
If a phase runs over the timebox target, Facilitator flags and asks:
β "We're at [X min] on Phase [N] β want to wrap and move on, or keep going?"
Never cuts off silently. Always human's call.
---
## Sprint Review Page Template
Create this live HTML page immediately after the welcome screen and share its link
before Phase 1, Question 1. Begin with the five-phase structure and an empty starting
state; fill it in as the sprint progresses. Update it after every meaningful human
input, agent round, decision, or artifact. Do not wait for a Decision Point.
The default view shows the current outcome and what needs attention. Supporting notes
may be collapsible. Keep artifacts in chronological phase order:
Define β Discover β Diverge β Converge β Build.
Required HTML structure:
```
[Project name] β Sprint Review
βββββββββββββββββββββββββββββββββββββββββ
UNIVERSAL PHASE NAVIGATION
1 Define Β· 2 Discover Β· 3 Diverge Β· 4 Converge Β· 5 Build
[active phase is visually distinct]
STATUS
Phase [N] of 5 Β· [in progress / DP[N] confirmed]
Timebox: [mode] Β· [target] Β· [elapsed]
CURRENT OUTCOME
[the latest confirmed framing, recommendation, or decision]
YOUR INPUT
[human statements, confirmations, redirects, and feedback needed]
PHASE 1 Β· DEFINE
Problem Β· examples Β· principles Β· success metrics
PHASE 2 Β· DISCOVER
Competitive evidence Β· agent insights Β· opportunity areas
PHASE 3 Β· DIVERGE
Round 1 concepts + votes
Round 2 concepts + votes
PHASE 4 Β· CONVERGE
Selected patterns Β· solution directions Β· recommendation Β· human decision
PHASE 5 Β· BUILD
Connected prototype Β· final artifacts Β· validation status
DECISION LOG
[DP1] [what was decided]
[DP2] [what was decided]
...
LINKS
Sprint Recap Deck Β· Build Brief Β· Executive Brief Β· Flow + Key Screens
```
HTML behavior requirements:
- The universal phase navigation is visible across the page and links to each phase.
- Every artifact appears inside the phase that produced it.
- Round 1 appears before Round 2; concepts and votes stay together.
- Human input is visually distinct from agent proposals and facilitator synthesis.
- The newest outcome and current feedback request are easy to find without scrolling
through the entire history.
- Embed competitive screenshots, concept sketches, and prototypes when available;
include their source links and labels.
- Use stable character avatars or role badges for agents. Use the human's supplied
avatar beside their input and feedback requests.
- Carry the confirmed visual language into new sections and prototypes. If the human
provides a reference, record what visual properties were applied.
- Verify the shared link and key navigation targets whenever the page structure changes.
Before each handoff, run one visual QA pass at phone and desktop widths. Check for
horizontal overflow, clipped or colliding text, broken images, weak contrast, inconsistent
spacing, misplaced artifacts, mismatched device-frame colors, and unclear active-phase
states. Fix observed defects, confirm once, and stop unless a specific defect remains.
Share the link as soon as the initial shell exists so the human has a working sprint
surface from the beginning. Reuse that link for the entire sprint.
---
## Reference Files
- `INSTALL.md` β One-time setup guide: installing the skill, the Google Workspace
MCP, the Mobbin MCP, and the recommended design skill stack
- `references/agent-prompts.md` β Reasoning patterns for all core, secondary, and
cross-cutting agents (5 core + 6 secondary + Content)
- `references/convergence-guide.md` β HMW scoring, idea evaluation, final recommendation format
- `references/writing-guide.md` β Content principles applied by the Content Agent:
voice, tone, and copy rules for every phase
Read these when running a full sprint or when you need deeper guidance on a specific phase.