The short answer
Run tactical urbanism pilots as reversible experiments, not aesthetic gestures. Prove impact by collecting the same quantitative and qualitative data before, during, and after the intervention — counts, speeds, crossing behaviour, dwell time and user perception, ideally against a control street. A short interim design exists to de-risk permanent investment: it either earns the evidence to justify scaling, or it fails cheaply and comes out. Lock the hypothesis, measurement windows and decision date before anyone picks up a paint roller.
Key takeaways
- A pilot is an experiment with a testable question, distinct from a one-off demonstration — that difference decides which data you need.
- Credibility comes from comparability: repeat identical counts and observations at the same days and times before, during and after, recording weather and events.
- For safety, use behavioural proxies — speeds, yielding, crossing compliance, waiting space — not only crash statistics.
- Treat perception as data: intercept surveys plus broader channels, always reporting response volumes so a handful of comments is not called 'the community'.
- A control street and independent yardsticks such as iRAP star ratings protect conclusions from seasonal and external noise.
- Reversibility is the point: on a 'no' the pilot comes out cheaply; on a 'yes' you plan maintenance and the capital phase from day one.
Decide what the pilot must prove before you touch the street
Tactical urbanism is powerful when a temporary intervention is framed as an experiment with a testable question rather than an end in itself. A one-day pop-up can build enthusiasm and show a possible future, but a true reversible pilot runs for weeks or months so behaviour has time to settle and data can accumulate. What you collect depends entirely on the decision the pilot is meant to inform.
Before anyone holds a paint roller, name the decision you are de-risking. Is the open question about safety, pedestrian volumes, dwell time, business impact or public acceptance? Write a short hypothesis in plain language, such as 'widening this corner will cut the share of people crossing mid-block.' Every metric you choose should test one clause of that hypothesis, and the pilot should end on a pre-agreed date when someone formally decides to scale, adjust or dismantle.
- Demonstration and pilot are not the same: the first is for excitement, the second for evidence.
- Write one plain-language hypothesis per intervention and pin it where the team can see it.
- Fix the decision date and the go / adjust / stop criteria in advance.
- Keep it reversible: paint, planters, modular kerbs and removable bollards.
- Name one accountable owner for evaluation and the final report.
Measure the same way before, during and after
Credible evidence comes from comparing like with like. Define your baseline before construction, then repeat identical counts, speeds and observations at the same days and times during the pilot and after removal or conversion. The team behind a 2017 São Paulo pop-up, supported by the Global Designing Cities Initiative, documented existing conditions first and then collected the same metrics during the intervention — pedestrian flow, vehicle volume, crossing behaviour, signal timing and user perception — which is why its results could be stated in clean percentages.
Practitioners who build monitoring into the design stage, such as urban designer Fabrizio Prati (a co-author of the Global Street Design Guide), advise pairing quantitative counts with qualitative material — interviews, focus groups and field notes — and collecting an identical set before and after. Keep collection conditions consistent: weekday versus weekend, peak versus off-peak, dry versus wet weather. If you cannot replicate conditions, the comparison weakens and you should say so in the report.
- Lock the measurement calendar before opening day and stick to it.
- Reuse the same surveyor instructions, counting forms and camera positions.
- Record weather, events and nearby works that could distort the counts.
- Archive raw data and timestamped photos for the final report.
Pick metrics a council or a court of opinion can trust
Different decisions need different proof. If the question is safety, look at objective behaviour rather than only crashes: measured speeds, the share of vehicles yielding to pedestrians, the share crossing at the marked facility, waiting space and turning speeds. In the São Paulo experiment these behavioural proxies moved quickly — a 75% fall in crossing outside the designated area, a 40% rise in cars yielding to pedestrians and 23% lower bus turning speeds — enough to shift local engineers toward permanent geometry.
If the question is whether a space is actually used and valued, measure dwell and activity. Bologna's school-square programme, which grew from tactical pilots into permanent city assets, combined counting, mapping, surveys and interviews and recorded an 87% increase in time spent in pilot areas during school drop-off and a 270% increase on weekends. Such numbers tell a story crash statistics cannot, because they capture life rather than just the absence of harm.
- Safety: speeds, yielding, compliance, waiting space, near-miss proxies.
- Use: counts, dwell time, group size, time-of-day profiles.
- Economy: spending proxies and vacancy, collected cautiously over longer windows.
- Sustainability: modal shift and walking / cycling share.
- Claim only what your chosen metric actually measures.
Treat perception as data, not decoration
Evidence that a design 'works' is incomplete without evidence that people want it. In the same São Paulo pop-up, 91% of people interviewed on site liked the new layout and 88% wanted it permanent. Engagement held before opening — workshops where residents mapped safe and unsafe places and reacted to draft designs — turned the intervention into a shared hypothesis rather than a top-down surprise. That preparation is part of the evidence chain, because acceptance is often the deciding factor for permanence.
Beware of perception data gathered only from people who already show up. Pair intercept surveys with broader channels — online consultation, local business interviews, letters to affected residents — and report response volumes so nobody overstates a handful of comments as 'the community.' Communicate findings visually: simple before/after graphics and time-lapses carry more weight in a decision meeting than a wall of tables.
- Run workshops before opening and intercept surveys during the pilot.
- Collect views through several channels and always state the number of respondents.
- Prepare visual before/after materials for the people who will decide.
- Capture negative feedback too — it shows what to refine.
Add controls and independent yardsticks to stay honest
Short pilots are noisy. Traffic varies by season, school terms, weather and nearby events, so a change in your numbers is not automatically caused by your design. Wherever possible compare your site with a similar 'control' street that received no intervention, measuring both on the same schedule; evaluators of active-street programmes call this a before-after-control design. At minimum, acknowledge the confounders and avoid overclaiming causation.
For safety claims, independent, standards-based tools strengthen credibility. iRAP star ratings, adopted within UN road safety targets as a global benchmark for the safety built into a design, let you score a proposed or interim configuration for pedestrians, cyclists, motorcyclists and vehicle occupants and compare it with the baseline street. Automating collection with cameras and sensors — as piloting cities in the CIVITAS network have done for bike-lane use, illegal parking and school-street trials — widens coverage and reduces observer bias, though interpretation still needs human judgement.
- Select a control street with similar land use and traffic before you start.
- Measure the pilot and control sites on the same schedule.
- Use iRAP star ratings as an external safety benchmark for interim designs.
- Cameras and sensors reduce human bias but do not replace analysis.
Convert the evidence into a permanent decision — or remove it
Reversibility is the point: if the data says no, the pilot should come down as cheaply as it went up, and that is a success because you avoided an irreversible mistake. If the data says yes, schedule the capital phase, secure budget and communicate the evidence that justified it. Interim materials deteriorate faster than permanent construction, so either a maintenance plan or a firm removal date must exist from the first day, otherwise a 'temporary' intervention quietly turns into a liability.
Close the loop publicly. Publish a short evaluation that restates the hypothesis, the method, the numbers, what changed relative to the control and the explicit decision. Agencies that treat pilots as a learning loop — measure, trial, refine — convert experiments into policy more reliably than those that stage one-off events, because the evidence, not the enthusiasm, carries the argument forward.
- Schedule maintenance or a removal date at the same moment you open the pilot.
- On a 'yes', fix capital-phase timing and a named budget source.
- Publish the final evaluation with method and numbers to build trust.
- Treat a 'no' as a win: you saved money on a wrong irreversible decision.
Put it into practice
The Reversible Pilot Evidence Plan (field template)
A single working sheet that keeps the team honest, from hypothesis to decision. Fill it in before opening day and return to it at every gate.
- Site and control street named, with similar land use and traffic load.
- One plain-language hypothesis per intervention, agreed by the team.
- Decision date and go / adjust / stop criteria fixed before opening.
- Baseline collection: at least two weekday and two weekend windows before any paint.
- Metric sheet: counts, dwell time, speeds, yielding, crossing compliance, perception.
- Identical post-opening windows at the same days, times and weather conditions.
- Perception plan: on-site intercepts + online channel + business interviews, with response counts.
- Comparison method (before/after vs control) and a stated list of confounders.
- Agreed thresholds that will trigger scale, adjust or remove decisions.
- Archive of photos, time-lapses and raw data with timestamps.
- Maintenance or removal plan and a capital-phase handover date.
Questions people ask
How long should a reversible pilot run before I judge it?
There is no universal duration, but practical guidance points to at least several weeks with both weekday and weekend measurement windows, so behaviour settles and data accumulates. Windows of a few days amplify the effect of weather, holidays and events and weaken conclusions. Fix the decision date in advance, account for season and school calendars, and gather several repeated windows under consistent conditions rather than relying on a single day.
Do I really need a control street, and what if I cannot get one?
A control street without the intervention, similar in land use and traffic, helps separate the effect of your design from seasonal and external fluctuations. If no suitable street exists, you can still build an argument, but you must honestly list confounders — season, holidays, weather — and avoid overclaiming causation. The minimum credible fallback is comparing several before and after periods on the same site and stating the limitation in the report.
What are the cheapest reliable metrics for a small team?
Start with what is visible and can be recorded on a template: pedestrian and cycle counts in fixed windows, dwell time, the share crossing at the marked facility and, where feasible, turning speeds. Add a short intercept survey of perception. Repeat the same measures before and after under identical conditions, keep raw data and photos, and you will have most of the evidence you need without expensive sensors.
How do I handle residents who oppose the pilot from the start?
Opposition is often driven less by the design than by a sense that the decision was made without them. Run workshops before opening where residents, businesses and schools map safe and unsafe places and react to draft designs, rather than merely being informed. Early engagement turns the pilot into a shared hypothesis. Capture negative views systematically, show that they were heard, and use data rather than argument to show whether the intervention works.
What should I do if results are mixed — safety improves but support drops?
That is a normal experimental outcome, not a failure. Separate the decisions: safety is an objective fact, while support can be rebuilt by refining the design and communication. Ask what exactly people dislike and adjust the pilot — more greenery, quieter traffic or better access for businesses — then run a second measurement round. Report both dimensions honestly and present scenarios rather than burying the project over one metric or ignoring the discontent.
How do I budget maintenance affordably if the pilot becomes permanent?
Interim materials wear faster than permanent construction, so a maintenance plan should exist from day one. Budget for replacing paint, fixings, planters and modular elements, and decide who is responsible — the city, businesses or volunteers. Run the capital phase in parallel: while the temporary solution proves its effect, design the permanent version with a realistic budget. If funding for permanence does not materialise, set a firm removal date so 'temporary' does not quietly become a cost liability.
Sources and further reading
Sources were checked when this page was generated. Confirm changing dates, rules and prices with the original publisher.
- People, Participation, and Pop-ups: Lessons in Tactical Urbanism in São Paulo, BrazilGlobal Designing Cities Initiative (NACTO)
- Star-Rated Street Designs: iRAP Star Ratings and the Global Street Design GuideGlobal Designing Cities Initiative (NACTO)
- Tactical Urbanism for Urban Design Interventions: In Conversation with Fabrizio PratiTransform Transport
- Designing Safer Spaces: Best Practices from CIVITAS CitiesCIVITAS (ICLEI Europe / European Commission)