Every multi-city product launch sounds the same in the kickoff meeting. "We want to launch in 25 markets simultaneously. Big bang. National press. National impact." And every multi-city launch that follows that brief ends the same way: 60% of the budget got spent in markets that did not move the needle, 30% got spent in markets that moved the needle a lot, and nobody can tell you which is which until the postmortem.
The Test → Validate → Scale framework below is the one we use across 500+ campaigns. It is not glamorous. It will not give you a national launch in week one. But it gets you to 25 cities with 60-80% efficiency, instead of the 30-40% efficiency you get from a national-bang launch. Read the framework, then read the example campaign at the end.
Table of Contents
Why National-Bang Launches Burn Budget
The temptation to launch nationally in week one is real. The CMO wants press. The board wants velocity. The agency wants the bigger fee. But the data on simultaneous multi-city launches is brutal:
- Variance in CPI across markets is typically 3-5x. NYC might cost $4/install, Phoenix $1.20. If you book equal hours in both, you waste in NYC.
- Per-market creative tweaks improve performance by 15-40%, but you cannot make tweaks if every market is live simultaneously.
- Operational risk concentrates. One bad team lead in one market is a contained problem if you launch 3 markets. It is a brand-wide problem if you launched 25.
- Press lift compounds when smaller markets feed into the larger story. You can build a press narrative across 8 weeks. You cannot build one in 5 days.
The Test → Validate → Scale Framework
The framework has three sequential phases. Each phase has a budget allocation, a primary KPI, and a go/no-go gate before the next phase begins.
| Phase | Cities | Duration | Budget Share | Primary KPI |
|---|---|---|---|---|
| Test | 2-3 | 2-4 weeks | 10-15% | Per-market unit economics |
| Validate | 5-8 | 3-4 weeks | 25-35% | Repeatability across geographies |
| Scale | 15-25 | 4-8 weeks | 50-65% | National volume at validated CAC |
Phase 1: Test (2-3 Cities, 2-4 Weeks)
Goal
Prove the unit economics of your launch model in a controlled environment. Establish a baseline CAC, conversion rate, and per-shift output you can defend.
City Selection
Pick 2-3 markets that represent your target audience but are not your most expensive or your most competitive. Good Phase 1 markets: Denver, Austin, Charlotte, Nashville, Phoenix, Minneapolis. Avoid NYC, LA, SF for the test phase.
Tactics
- 2-4 brand ambassadors (BAs) per city, 1 team lead
- 4-6 activation locations per market
- Single creative version, single offer, single QR code structure
- Daily KPI reporting
Budget
10-15% of the total program budget. A 2-3 city test is deliberately lean, the goal is a defensible unit-economics baseline, not volume.
Go / No-Go Gate
You should see a defensible CAC, repeatable conversion rate, and clear creative/offer insights before approving Phase 2. If CAC is 2x+ above target, fix the unit economics before scaling.
Phase 2: Validate (5-8 Cities, 3-4 Weeks)
Goal
Prove the model works across diverse geographies. Stress-test for variance in cost, conversion, and operational complexity.
City Selection
5-8 cities that span geographic and demographic diversity. Mix coastal and middle-America, high-density and suburban. Good Phase 2 mix: Chicago, Atlanta, Dallas, Boston, Seattle, Miami, Detroit, Portland.
Tactics
- 4-8 BAs per city, 1-2 team leads, 1 regional manager
- 6-12 activation locations per market
- 2-3 creative variants A/B/C tested
- Real-time dashboard with city-by-city KPIs
- Weekly creative iteration based on top-performing markets
Budget
25-35% of the total program budget for 5-8 cities over 3-4 weeks, the step up from Phase 1 reflects the added markets, team leads, and creative variants.
Go / No-Go Gate
If 6 of 8 markets hit target CAC and conversion, you are validated. If only 3-4 hit target, you have a market-fit problem that scaling will not fix. Pause and diagnose before going to Phase 3.
Phase 3: Scale (15-25 Cities, 4-8 Weeks)
Goal
Hit national volume at the validated CAC. The strategic question is now executional: can you operate consistently at 3-5x the scale of Phase 2?
City Selection
Top 15-25 US markets by your target audience density. Tier them: Tier 1 (NYC, LA, Chicago, SF, DC, Boston, Miami), Tier 2 (Dallas, Atlanta, Houston, Phoenix, Philadelphia, Seattle, San Diego), Tier 3 (Denver, Austin, Charlotte, Minneapolis, Portland, Nashville, Tampa, Orlando, St. Louis, Pittsburgh).
Tactics
- 6-15 BAs per Tier 1 city, 4-8 per Tier 2/3
- National account director plus regional managers (1 per 5-7 cities)
- Validated creative + offer from Phase 2, locally adapted
- Integrated paid social amplification per market
- National PR push timed to Phase 3 kickoff
Budget
50-65% of the total program budget for 15-25 cities over 4-8 weeks, the largest phase since it carries the national volume push.
Operational Risks at Scale
- Recruiting bottlenecks in Tier 2/3 markets — start 4 weeks early
- Quality drift — rotate top team leads through new markets in the first week
- Reporting overload — consolidate dashboards, kill noisy metrics
Walkthrough: A 20-City CPG Launch
Here is how the framework plays out for an illustrative 20-city CPG SKU launch over a 14-week window.
Phase 1: Test (Weeks 1-3)
Denver and Austin. 3 BAs per market, 1 team lead, 4 high-foot-traffic activation locations per city. A test phase at this scale typically produces a cost-per-trial in the low single digits, giving you a defensible baseline before committing to Phase 2. Creative version B outperforming version A by 30-40% is a common Phase 1 finding worth locking in before scaling.
Phase 2: Validate (Weeks 4-7)
Chicago, Atlanta, Dallas, Boston, Seattle, Miami. Adopt the winning creative nationally and double the BA count per market. A validated cost-per-trial in line with Phase 1 across 5 of 6 markets is the signal to proceed to Phase 3; an underperforming market like Miami getting reallocated to weekend hours is a normal mid-phase adjustment, not a reason to pause.
Phase 3: Scale (Weeks 8-14)
Add the remaining 12 cities, launch national PR coverage, and integrate paid social amplification of the street team content. Cost-per-trial typically improves at this stage as scale efficiencies kick in, even as total sample and redemption volume grows several times over from Phase 2.
Final Numbers
Across all three phases, a launch at this scale typically distributes several hundred thousand samples and produces a blended cost-per-trial meaningfully lower than the Phase 1 baseline, driven by the efficiencies unlocked in Phase 3. Every launch is custom-quoted based on the specific city list, timeline, and team size.
The Decision Rules That Matter
- If Phase 1 CAC is 2x+ over target, do NOT proceed to Phase 2. Fix the model first.
- If Phase 2 shows fewer than 70% of markets hitting target, do NOT proceed to Phase 3. You have a strategy problem, not a scale problem.
- Re-allocate weekly during Phase 3. Hours from weak markets to strong markets is the highest-ROI lever you have.
- Reserve 10% of the budget for opportunistic adds. A market that overperforms in Phase 2 should get more budget in Phase 3 than the original plan called for.
- Lock the creative after Phase 2. If you are still iterating creative in Phase 3, you are not scaling, you are testing in production.
Frequently Asked Questions
What is the Test, Validate, Scale framework for a multi-city product launch?
It is three sequential phases, each with a go/no-go gate before the next one begins. Test (2-3 cities, 2-4 weeks, 10-15% of budget) proves per-market unit economics. Validate (5-8 cities, 3-4 weeks, 25-35% of budget) proves the model is repeatable across diverse geographies. Scale (15-25 cities, 4-8 weeks, 50-65% of budget) hits national volume at the CAC validated in the earlier phases.
How many cities should the Test phase of a multi-city launch include?
2-3 markets that represent your target audience but are not your most expensive or most competitive, avoid NYC, LA, and SF for the test phase. Good Phase 1 markets include Denver, Austin, Charlotte, Nashville, Phoenix, and Minneapolis. The goal is a defensible CAC and repeatable conversion rate, not volume.
When should a brand not move from Phase 1 to Phase 2, or Phase 2 to Phase 3?
If Phase 1 CAC is 2x or more over target, do not proceed to Phase 2, fix the unit economics first. If Phase 2 shows fewer than 70% of markets hitting target, do not proceed to Phase 3, that is a strategy problem, not a scale problem that more cities will fix.
Why does cost per acquisition vary so much between cities in a multi-city launch?
Variance in cost-per-acquired-customer across cities in a typical multi-city launch is typically 3-5x. That is the reason for the phased Test, Validate, Scale rollout: it identifies which higher-cost markets to avoid before committing 50-65% of the budget to the Scale phase.
For App Launches Specifically
If you are launching an app rather than a CPG product, the framework holds but the metrics shift. Replace cost-per-trial with cost-per-install, swap coupon redemptions for D7 retention, and add MMP attribution from day one. The full app-launch version of this playbook lives in our app launch street team guide.
Key Resources
Related Reading
- Street Team Marketing Cost in 2026
- App Launch Street Team Guide
- What is Field Marketing? The 2026 Guide
Build a Multi-City Launch Plan
We have run multi-city launches in 1,000+ cities. Share your target markets and product and we will sketch a Test → Validate → Scale plan within 24 hours.
Plan My LaunchOr email hello@streetteamsco.com