Crew Expansion Map
Stage 3 is complete for the Crew product path. The approved non-canonical working packet has been archived, the canonical expansion artifacts have been written, and this alignment page is now the confirmed record for the completed cycle.
alignment_status: confirmed visual tier research/crew 2026-06-24
Table Of Contents
Confirmed Status
alignment_status: confirmed
Confirmation date: 2026-06-24
Finished but amendable: this page is current for the completed Crew expansion-map cycle. Later research may amend it only by archiving this confirmed page first and highlighting the new evidence, decisions, or scope changes.
research/crew/expansion-map.mdresearch/crew/expansion-map-interview.mdalignment/expansion-map-crew.htmldocs/history/archive/2026-06-24/134853/research/crew/_working/preliminary-expansion-map-research.mdWhat Changed During Confirmation
- Archived the Stage 2 review page before replacement.
- Archived and removed the active Stage 2 working packet.
- Wrote the approved canonical Crew expansion map and interview/evidence artifact.
- Converted this page from review to confirmed and removed review gates, feedback inputs, and response compilation controls.
Approval Record
The final compiled response YAML approved Stage 3 with no unresolved required questions, no explicit edit notes, and no negative section feedback.
| Gate | Gate Type | Approved Decision | Target |
|---|---|---|---|
| Evidence Matrix | research completeness | Evidence is sufficient for Stage 3 canonicalization after applying any explicit edits in this response. | Stage 3 evidence readiness |
| User Format Preferences | user format preferences | Approve the Stage 2 review format as written. | Review format |
| Final Artifact Approval | final artifact approval | Approve Stage 3: archive/remove the working packet, write canonical Crew expansion-map and interview artifacts, and confirm the alignment page. | research/crew/expansion-map.md |
Final Approval YAML
alignment_page: "alignment/expansion-map-crew.html"
response_status: complete
approval_status: ready-for-agent-review
required_gate_status: complete
unanswered_required_questions:
[]
gate_answers:
- section: "Evidence Matrix"
gate_type: "research completeness"
status: "answered"
answer: |-
Evidence is sufficient for Stage 3 canonicalization after applying any explicit edits in this response.
- section: "User Format Preferences"
gate_type: "user format preferences"
status: "answered"
answer: |-
Approve the Stage 2 review format as written.
- section: "Final Artifact Approval"
gate_type: "final artifact approval"
status: "answered"
answer: |-
Approve Stage 3: archive/remove the working packet, write canonical Crew expansion-map and interview artifacts, and confirm the alignment page.
target_path: "research/crew/expansion-map.md"
Canonical Expansion Map
Based on: research/crew/journey-map.md, research/crew/icp.md, research/crew/positioning.md, research/crew/competitive-analysis.md, Trace sibling artifacts for cross-sell mechanics, and current external validation captured in research/crew/expansion-map-interview.md.
Date: 2026-06-24
Summary
Crew's primary expansion model is shared-account scope expansion, not a Solo-to-Platform tier ladder. A retained Crew account should grow when the same CalcLLM account needs to answer a neighboring cost question: Crew users need production LLM cost attribution through Trace, or Trace users need engineering AI-tool spend governance through Crew.
Crew-only expansion begins when visibility becomes governance. A monthly dashboard habit is retention; expansion starts when Crew adds team leads, Finance/FP&A, budget thresholds, cost-center allocation, executive reporting, chargeback, enterprise controls, a sibling product module, or advocacy proof.
Role expansion matters as much as seat expansion. Crew should track which new stakeholder consumes a workflow artifact and which role gains budget authority: Platform Eng Lead champion, platform/admin owner, team leads, Finance/FP&A, VP Eng or CTO, Security/Compliance, and Procurement.
Native vendor controls validate the problem and raise the bar. GitHub Copilot, Claude Code, Cursor, Devin/Windsurf, and adjacent vendors expose budget, usage, admin, or enterprise controls. Crew should therefore lead with cross-tool normalization, team attribution, outcome linkage, and shared Crew-plus-Trace account intelligence rather than generic budget alerts.
Expansion Triggers
| Trigger | Signal | Expansion Motion | Evidence | Confidence |
|---|---|---|---|---|
| Shared-account production-cost question | Crew user asks why production LLM cost changed or which feature/customer drove production spend | Add Trace to the same CalcLLM account | Crew and Trace journey maps both name cross-sell as a gap | Medium-high |
| Trace account scales engineering team | Trace user asks what engineers spend across AI coding tools | Add Crew to the same CalcLLM account | Trace journey and positioning name the Trace-to-Crew trigger | Medium |
| Engineering headcount growth | 50 engineers becomes 100-150, or more teams need cost-center mapping | Add teams, team dashboards, cost centers, user-level views, Slack alerts | Crew journey map lists headcount growth and team-level rollup | High |
| New AI dev-tool adoption | Account adds Cursor, Claude Code, Copilot, Devin/Windsurf, or another coding agent | Add provider integration and cross-tool normalization | Crew positioning says aggregation across vendors is the wedge | High |
| Budget volatility under usage-based plans | Shared pools burn early, power users consume more than planned, or caps become necessary | Add forecasting, anomaly workflows, budget guardrails, executive reporting | Current GitHub and Claude docs validate usage volatility and spend limits | High |
| Showback becomes chargeback | Finance asks for allocation by team/cost center and teams dispute numbers | Add FP&A export, audit trail, allocation rules, role-specific dashboards | Crew journey and ICP name Finance's "who spent this?" question | Medium-high |
| Enterprise readiness | SSO, RBAC, audit logs, SCIM, invoice/PO, support, security review | Upgrade tier or sales-assisted motion | Crew ICP and current vendor pricing pages expose enterprise controls | Medium |
| ROI proof | Cost-per-PR, budget surprise prevented, or tool consolidation win | Referral, case study, platform-community talk, FinOps story | Crew journey-map advocacy paths | Medium |
Upgrade And Seat Growth Paths
| Path | Baseline State | Expansion State | Roles Added | Product Surface | Caveat |
|---|---|---|---|---|---|
| Product-line account | Crew only | Crew plus Trace | CTO, production platform owner, product/feature owners | Shared account navigation, shared billing, shared identity/team taxonomy | Cross-sell needs a specific production-cost question to avoid feeling forced |
| More engineers | One platform lead reviews account spend | Multiple teams and cost centers review their own spend | Team leads, engineering managers | Team dashboards, cost-center filters, Slack digest subscriptions | No monetization artifact defines whether this is seat-based or usage-based |
| More tools | One or two providers connected | Three or more AI coding tools normalized | Tool owners, platform admins | Provider connectors, vendor-normalized usage and cost views | Per-developer granularity depends on vendor API detail |
| Governance workflow | Monthly visibility | Showback, chargeback, budget caps, allocation reports | Finance/FP&A, VP Eng, team leads | FP&A export, allocation rules, budget alerts, audit trail | Chargeback may be too heavy pre-PMF; treat packaging as provisional |
| Higher assurance | Self-serve admin | Enterprise controls | Security, compliance, procurement | SSO/SAML, SCIM, RBAC, audit logs, custom data retention, invoice/PO | Defer full enterprise sales until post-PMF unless pulled by customers |
| Advocacy | Private ROI proof | Public or peer proof | Champion, comms/exec reviewer | Case-study export, referral prompt, community-talk support | Needs customer permission and evidence quality |
Team Or Account Rollout
| Step | Account State | Primary User Job | Expansion Test | Health Signal | Risk Signal |
|---|---|---|---|---|---|
| 1. Single champion | Platform lead connects first provider | Answer "who spent what?" | Can Crew reveal an unexpected per-tool or per-user fact? | First dashboard viewed in the first session | No cross-tool reveal within 5 minutes |
| 2. Team mapping | Teams and cost centers assigned | Make spend legible by team | Does the champion invite team leads or add cost centers? | 80%+ spend attributed | 20%+ spend remains shared or ambiguous |
| 3. Governance ritual | Monthly review becomes recurring | Explain budget movement to Finance | Does Finance ask for recurring exports? | Monthly report generated twice | Dashboard viewed only at renewal or crisis |
| 4. Budget controls | Forecasts and alerts guide behavior | Prevent surprise spend or justify productive spend | Does account configure thresholds? | Alert acted on within one business day | Hard caps block productive work or get disabled |
| 5. Product-line scope | Production-cost or dev-tool-cost question crosses domains | Add Trace or Crew inside the same account | Does an unanswered question belong to the sibling product? | Cross-sell preview clicked or requested | Cross-sell shown before current module proves value |
| 6. Executive proof | VP Eng/CTO uses ROI narrative | Defend AI tooling investment | Does buyer use a board/CFO-ready artifact? | Export shared with leadership | Leadership sees only raw spend, no outcome context |
| 7. Advocacy | Champion has a credible win | Refer, review, or speak publicly | Is there a specific public-safe proof point? | Referral or case-study permission | Ask arrives before proof or without data |
Referral And Advocacy Paths
| Advocacy Path | Trigger | Proof Artifact | Best Ask | Confidence |
|---|---|---|---|---|
| Peer referral | Platform lead can answer the Finance question quickly | Before/after attribution screenshot or sanitized table | Intro to another platform lead | Medium |
| Engineering blog post | Tool consolidation or budget surprise prevention | Cost-per-PR or vendor-normalized spending story | Co-authored technical post | Medium |
| PlatformCon/KubeCon talk | Governance workflow becomes repeatable | Showback-to-chargeback operating model | Conference talk support | Medium-low |
| FinOps case study | Finance uses Crew in monthly budget review | Allocation accuracy, forecast accuracy, saved or avoided spend | Case study or webinar | Medium |
| Product review | Startup CTO gets a fast surprise insight | Self-serve first-value story | Product Hunt, G2, or dev-tool community review | Medium-low |
Risks And Constraints
| Risk | Evidence | Impact | Mitigation |
|---|---|---|---|
| Native vendor controls catch up | GitHub docs show AI credits, pooled allowances, user/cost-center/enterprise budgets, and blocking behavior | Weakens "budget controls" as standalone wedge | Anchor Crew on cross-tool aggregation, team attribution, outcome linkage, and shared-account cross-sell |
| Vendor branding shifts | Windsurf pricing now redirects to Devin pricing and labels Windsurf as Devin Desktop | Crew docs may sound stale if they keep "Windsurf" without current naming | Use "Devin/Windsurf" until canonical docs are refreshed |
| Per-developer attribution gaps | Crew journey and positioning note pooled plans may expose only aggregate spend | Account health and chargeback may be unreliable | Add confidence/coverage indicators and progressive integration options |
| Cross-sell feels forced | User approved cross-sell emphasis, but no customer interviews exist | Premature upsell can damage trust | Trigger cross-sell only from a concrete missing answer in the current workflow |
| No monetization artifact | No scoped Crew monetization, lifecycle metrics, GTM, specs, or feedback docs found | Tier and packaging recommendations are provisional | Keep package names conceptual and recommend lifecycle metrics/monetization follow-up |
| Cohrint competes more directly than older repo framing suggests | Current Cohrint page covers Claude Code, Copilot, Cursor, Codex CLI, Gemini CLI, MCP, CLI, OTel, SDK, dashboards, budget enforcement | Crew's "near-empty" claim should be narrowed | Position Crew as read-only governance and cross-product account intelligence, not routing optimization |
Product Gaps
| Gap | Why It Matters | Follow-Up |
|---|---|---|
| Expansion packaging | This map identifies upgrade moments but cannot price them confidently | Run monetization or packaging research before final pricing decisions |
| Lifecycle metrics | Health thresholds need instrumentation definitions | Run $lifecycle-metrics research/crew after expansion map confirmation |
| Onboarding activation | Cross-sell timing depends on first-value and retained usage | Run onboarding map before implementing in-product prompts |
| Customer proof | Advocacy and referral paths are inferred from journey docs | Gather customer interviews or simulated discovery notes |
| Integration feasibility | Per-user/per-team attribution may vary by vendor API | Technical validation/spec interview before committing to chargeback claims |
Product Path Implications
Crew should stay scoped to AI dev-tool cost intelligence while treating Trace as the first sibling expansion path. The shared CalcLLM account should carry common identity, teams, cost centers, billing, Slack reporting, and account navigation so cross-sell can be presented as adding an answer inside the same operating model.
Do not create a Trace expansion artifact from this Crew map. Trace should receive its own expansion-map pass later if the product path remains active and the Trace-to-Crew trigger needs canonical treatment.
Next Steps
- Run
$lifecycle-metrics research/crewto define expansion triggers, account-health thresholds, cross-sell events, and advocacy readiness signals. - Run monetization or packaging research before finalizing tier names, prices, or plan limits.
- Validate integration feasibility for per-user and per-team attribution before promising chargeback workflows.
- Gather customer discovery around Crew-to-Trace and Trace-to-Crew trigger questions before implementing in-product cross-sell prompts.
Visual Account-Flow Diagrams
Shared CalcLLM Account Expansion Flow
Crew Expansion Trigger Map
| Trigger | Expansion Motion | Terminal Outcome |
|---|---|---|
| More teams | Add cost centers and team dashboards | Expanded account |
| New AI tool | Add provider integration and cross-tool normalization | Expanded account |
| Finance allocation | Add showback, chargeback, and FP&A export | Expanded account |
| Budget volatility | Add forecasting, budget guardrails, and anomaly workflow | Expanded account |
| Production cost | Add Trace to the same account | Expanded account |
| Proof artifact | Ask for referral, case study, or public story | Advocacy |
Interview And Evidence Artifact
Evidence Matrix
| Claim | Evidence | Inference | Confidence | Assumption Status | Decision Impact |
|---|---|---|---|---|---|
| Shared-account cross-sell is the primary expansion model | Crew and Trace journey, positioning, and ICP docs describe a shared product-line account and bidirectional cross-sell | Expansion should be a domain-scope addition, not a maturity graduation | High | Evidence-backed for strategy; product mechanics unbuilt | Put cross-sell first in the canonical expansion map |
| Crew-only expansion begins with governance workflows | Crew journey-map tasks include monthly budget review, anomaly investigation, team budgets, and ROI justification | A retained dashboard becomes expansion when roles and workflows multiply | High | Packaging still provisional | Separate retention from expansion |
| Role expansion matters | Crew ICP names champion, economic buyer, Finance influencer, and security blocker | Account rollout should map roles and artifacts, not only seats | High | Evidence-backed | Include role-aware rollout paths |
| GitHub validates budget-control urgency but is also a risk | Current GitHub docs describe AI credits, pooled allowances, budgets at multiple levels, and usage blocking | Vendor-native controls can solve single-vendor budget problems | High | Current as of 2026-06-24 | Crew must lead with cross-tool value |
| Claude validates per-developer cost variability | Claude Code docs describe token-based charges, active-day averages, monthly averages, /usage, spend limits, and reporting | Teams need baseline pilots and cost tracking | High | Current as of 2026-06-24 | Use Claude as trigger evidence, not as the sole wedge |
| Cursor and Devin/Windsurf support enterprise or tier expansion surfaces | Current pricing pages list team billing, usage analytics/admin dashboards, SSO, enterprise controls, usage quotas, and extra usage | Buyers expect admin and governance controls in AI coding tools | Medium | Public pricing pages lack deep API details | Keep enterprise controls as likely tier drivers |
| Cohrint narrows the empty-space claim | Current Cohrint page claims routing, dashboards, Claude Code/Copilot/Cursor support, budget enforcement, and savings pricing | Crew should avoid claiming no AI coding cost player exists | Medium-high | Vendor claims unverified by customers | Say Crew lacks a validated cross-tool read-only governance layer, not that there are no competitors |
| Advocacy follows proof | Crew journey-map lists platform talks, peer referrals, engineering blogs, and FinOps case studies | Ask for advocacy after a proof artifact exists | Medium | No customer feedback yet | Add advocacy thresholds rather than generic referral prompts |
Source Coverage
| Source | Coverage | Notes |
|---|---|---|
research/crew/journey-map.md | High | Expansion triggers, task journeys, cross-product gap, advocacy paths |
research/crew/icp.md | High | Buyer/user roles, pain map, willingness-to-pay signals, account model |
research/crew/positioning.md | High | Shared account, cross-sell, anti-positions, product/sales/pricing implications |
research/crew/competitive-analysis.md | Medium-high | Competitor framing, GTM lessons, pricing expectations, native-convergence risk |
research/trace/* | Medium | Sibling context only for cross-sell mechanics |
| Crew retention map, lifecycle metrics, monetization, GTM, customer feedback, specs | Missing | Keep affected claims provisional |
| Source | Relevant Observation | Use In This Artifact |
|---|---|---|
| GitHub Copilot licenses | Copilot licensing combines seats and AI credits; org/enterprise owners can use budgets for AI-credit consumption | Validates AI-credit budget governance as a current market reality |
| GitHub Copilot org/enterprise usage-based billing | Business and Enterprise include pooled monthly AI credits; promotional amounts run June 1 to September 1, 2026; budgets can halt user access and additional spend | Makes GitHub native controls Crew's main competitive risk |
| GitHub Copilot budget controls | GitHub recommends enterprise guardrails because a heavy user or automated agent session can consume a disproportionate shared pool early in the cycle | Validates Crew's power-user and pooled-budget trigger |
| GitHub Copilot models and pricing | Copilot converts token consumption into AI credits at 1 credit = $0.01; model and token count determine interaction cost | Supports the need for normalized model/usage explanations |
| Claude Code costs | Claude Code charges by API token consumption; enterprise deployment averages around $13 per developer active day and $150-250 per developer per month; admins can set limits and view reporting | Validates per-developer cost volatility and spend-limit workflows |
| Claude pricing | Claude Pro includes Claude Code; Team and Enterprise expose central billing/admin, SSO, usage analytics, spend controls, audit logs, and usage-at-API-rates enterprise pricing | Supports higher-tier control expectations |
| Cursor pricing | Cursor lists Individual at $20/month, Teams at $40/user/month, usage-based continuation, admin dashboard metrics, SSO, enterprise pooled usage, SCIM, audit logs, and controls | Supports tier/readiness triggers and shows vendor-native admin surfaces |
| Devin/Windsurf pricing | Windsurf pricing redirects to Devin; Devin Desktop pricing lists Pro at $20/month, Max at $200/month, Teams at $80/month plus $40/full dev seat, quotas, admin dashboard analytics, and enterprise controls | Flags naming drift and validates quota/admin-surface expansion |
| Cohrint | Cohrint claims savings routing for Claude Code, Copilot, and Cursor; supports MCP, CLI, SDK, OTel, dashboards, budget enforcement, and savings-based Growth pricing | Updates closest-competitor risk and narrows Crew's positioning |
| DX pricing | DX sells modular engineering intelligence and AI measurement with developer-license pricing and annual contracts | Confirms DX remains broad engineering intelligence, not Crew's direct cost-input wedge |
| Worklytics pricing | Worklytics Business starts at $2,500/month with up to 200 users and integrations including ChatGPT, Microsoft Copilot, GitHub, and more; AI adoption insights are included | Confirms Worklytics is adoption/work-pattern analytics, not spend attribution |
Assumptions And Confidence Register
| Assumption | Status | Confidence | What Would Change It |
|---|---|---|---|
| Crew and Trace will share one account, identity layer, and billing relationship | Repo-supported | High | Product architecture decision changes |
| Shared-account cross-sell should be first in the canonical map | User-approved and repo-supported | High | User later narrows scope to Crew-only |
| Chargeback is a genuine expansion motion | Provisional | Medium | Monetization/lifecycle work shows chargeback must be baseline activation |
| Cross-tool aggregation remains differentiated despite vendor controls | Evidence-backed but competitive | Medium-high | GitHub, Claude, Cursor, or Devin ships robust cross-vendor import/normalization, or teams consolidate to one vendor |
| "Windsurf" should be refreshed to "Devin/Windsurf" in future Crew docs | Current-source observation | Medium | Product docs clarify separate branding or market usage keeps Windsurf dominant |
| Per-developer attribution is technically feasible from read-only sources | Known gap | Medium-low | Vendor APIs expose insufficient user-level data |
| Advocacy can become a repeatable expansion loop | Inferred | Medium | Customer interviews fail to show willingness to refer or publish |
Interview Prompts To Validate
- When does a monthly Crew budget review become an account expansion opportunity?
- Expected answer to validate: when it adds team leads, Finance/FP&A, budget thresholds, cost-center allocation, or chargeback.
- Evidence needed: lifecycle metrics or customer discovery.
- What should trigger Crew-to-Trace cross-sell?
- Expected answer to validate: a production-cost attribution question that Crew cannot answer, such as "why is our production LLM bill 3x expected?" or "which feature/customer drove the production spend?"
- Evidence needed: onboarding flow, dashboard event, or discovery interview.
- What should trigger Trace-to-Crew cross-sell?
- Expected answer to validate: a Trace account's CTO scales engineering headcount and begins asking what engineers spend across AI coding tools.
- Evidence needed: Trace onboarding and expansion map.
- Which roles must be mapped before chargeback is credible?
- Expected answer to validate: platform admin, team lead, Finance/FP&A, economic buyer, and a dispute owner.
- Evidence needed: chargeback workflow spec.
- What makes an advocacy ask appropriate?
- Expected answer to validate: budget surprise prevented, ROI proof, cost-per-PR narrative, or tool-consolidation win.
- Evidence needed: customer feedback and public-story permission.
Current-Market Validation Interview
- Do GitHub AI Credits and budget controls reduce Crew demand or create urgency? Current synthesis: both. They validate volatility and governance needs, but weaken single-vendor budget-alert differentiation.
- Does Claude Code's own cost tooling make Crew redundant? Current synthesis: not by itself. It helps within Claude, but Crew's value depends on cross-tool and cross-team normalization.
- Does Cursor or Devin/Windsurf already solve team cost intelligence? Current synthesis: they expose admin/usage surfaces and enterprise controls, but public pages do not show cross-vendor, per-outcome cost intelligence.
- Does Cohrint directly compete? Current synthesis: yes, as cost optimization/routing and dashboards for AI coding spend. Crew should differentiate on read-only governance, attribution, account rollout, and Trace cross-sell.
Follow-Up Research Questions
- Which account-health event best predicts Crew-to-Trace interest: dashboard repeat use, unanswered production-cost question, executive export, or provider-integration milestone?
- Which expansion package should own budget thresholds, forecasting, and chargeback: baseline Crew, a governance add-on, or an enterprise tier?
- What data quality threshold is required before Team Lead and Finance views become trustworthy?
- Which advocacy artifact is easiest to produce without exposing sensitive spend data?
- Which vendor APIs expose enough user/team-level data for reliable cost-center allocation?
Archive Record
| Artifact | Archived Path | Active State |
|---|---|---|
| Stage 2 review page | docs/history/archive/2026-06-24/134853/alignment/expansion-map-crew.html | Replaced by this confirmed page |
| Stage 2 working packet | docs/history/archive/2026-06-24/134853/research/crew/_working/preliminary-expansion-map-research.md | Removed from research/crew/_working/ |
Downstream Implications
The next research dependency is lifecycle instrumentation for Crew expansion signals. The confirmed map recommends $lifecycle-metrics research/crew as the next concrete skill command, with monetization/packaging and integration feasibility work following before final pricing or chargeback commitments.