FINDING WHAT
USERS NEED BEFORE
A LINE IS
DRAWN

I run the full research lifecycle myself: brief, study design, recruiting and screening, moderation, and synthesis into deliverables a team can act on. Five threads here shaped real product direction, from foundational discovery to multi-group usability testing to the operations that let several studies run at once.

Role
Lead Product Designer, research DRI
Discipline
Discovery, usability testing, synthesis, research ops
Context
Enterprise B2B SaaS platform
Threads in this study
05
THREAD 01 / 05
Multi-group usability study
3
User groups tested, moderated and unmoderated
2.9/5
Average untrained confidence, the metric that reframed everything
5/7
Named error recovery top blocker, prioritized before re-test
The problem

A new bulk creation feature was in prototype, and engineering wanted evidence it worked before committing. One user group wouldn't tell the whole story: internal experts, paying customers, and external marketers each bring different context.

What I did

I ran usability testing across three groups, internal team members moderated, current customers and external growth marketers unmoderated, writing the screeners, recruiting participants, moderating the think-aloud sessions, and building the unmoderated tests myself. One cross-group view separated universal findings from group-specific ones, with severity scored by product familiarity. The findings reshaped the plan: we cut a complex in-platform editing approach that proved hard to use and rarely needed, and shipped the simpler, higher-value path first.

The concept landed in every group. Confusion clustered in a few moments, and severity tracked product familiarity. Cross-group synthesis
MODERATED UNMODERATED CUT THREE USER GROUPS SAME PROTOTYPE, MATCHED METHODS INTERNAL TEAM MODERATED THINK-ALOUD CURRENT CUSTOMERS UNMODERATED TASKS GROWTH MARKETERS UNMODERATED TASKS SYNTHESIZE ONE CROSS-GROUP VIEW SEVERITY SCORED BY PRODUCT FAMILIARITY UNIVERSAL FINDINGS TRUE FOR EVERY GROUP GROUP-SPECIFIC TRUE FOR SOME RESHAPE THE PLAN SIMPLER PATH SHIPPED FIRST IN-PLATFORM EDITING CUT: HARD TO USE, RARELY NEEDED
Fig. Three groups, one signal: matched methods, one synthesis, and the plan reshaped
THREAD 02 / 05
Discovery interviews
5+
Industries: retail, grocery, hospitality, media, sports
1
Dominant signal: workaround systems built outside the product
100%
Of themes shipped with implications and next steps
The problem

Before designing anything, we needed to know how enterprise customers worked at scale, where the tooling fell short, and what they'd built to compensate. Foundational discovery: no prototype, just open questions about real workflows.

What I did

I ran discovery sessions with enterprise customers across retail, grocery, hospitality, media, and sports, then synthesized them into cross-interview theme documents. The strongest signal recurred in every vertical: customers had independently built elaborate parallel spreadsheet systems because the in-platform experience didn't match how they worked. The recurring themes, naming and discoverability at scale, export and audit gaps, a hard code dependency for non-technical users, each became a structured deliverable with quotes, implications, and next steps the team could design against.

Customers had quietly built their own spreadsheet systems around the product. That workaround was the clearest signal in the study. Discovery synthesis
SESSIONS THEME STRONGEST SIGNAL DISCOVERY SESSIONS ENTERPRISE CUSTOMERS, FIVE VERTICALS RETAIL GROCERY HOSPITALITY MEDIA SPORTS CLUSTER RECURRING THEMES EACH SHIPPED WITH QUOTES, IMPLICATIONS, NEXT STEPS PARALLEL SPREADSHEET WORKAROUNDS SYSTEMS BUILT OUTSIDE THE PRODUCT TO GET WORK DONE THE STRONGEST SIGNAL NAMING + DISCOVERABILITY AT SCALE EXPORT + AUDIT GAPS CODE DEPENDENCY FOR NON-TECHNICAL USERS
Fig. Interviews across five verticals, clustered into themes the team could act on
THREAD 03 / 05
Research operations
3
Regions scaled: Europe, North America, Asia-Pacific
100%
Founding cohort retention, voluntary attrition aside
8
Roadmap influences traced to the founding cohort
The problem

One study at a time couldn't keep pace with the roadmap, and ad hoc recruitment was slow and inconsistent. We needed a repeatable way to source the right participants and run parallel studies across regions without findings bleeding together.

What I did

I founded a customer advisory community as a standing recruitment pool: a hand-selected founding cohort of enterprise customers across travel, food, fintech, retail, and tech; a year-long cadence of an in-person kickoff, monthly virtual sessions, and quarterly events and newsletters; and a dedicated channel where members answered each other's product questions within the first week. Its scoring model expanded the program across Europe, North America, and Asia-Pacific. With the pool in place, concurrent studies ran across product areas at the roadmap's pace, with screener quality and clear synthesis boundaries keeping every insight attributable.

A good screener decides who shows up; parallel studies stay useful only when each synthesis has clear boundaries. Research ops principle
PARTICIPANT SCREENER SYNTHESIS BOUNDARY STANDING POOL THE ADVISORY COMMUNITY, ALWAYS ON ENTERPRISE CUSTOMERS, THREE REGIONS SCREENER: DECIDES WHO SHOWS UP PARALLEL STUDIES SEPARATE SYNTHESIS KEEPS INSIGHTS ATTRIBUTABLE STUDY A RECRUIT SESSIONS SYNTHESIS FINDINGS STUDY B RECRUIT SESSIONS SYNTHESIS FINDINGS STUDY C RECRUIT SESSIONS SYNTHESIS FINDINGS RUN AT THE SAME TIME, NEVER BLENDED
Fig. One standing pool feeding parallel studies with clean synthesis boundaries
THREAD 04 / 05
Study design
2
Methods matched: moderated for reasoning, unmoderated for volume
1
Falsifiable hypothesis set per study, before sessions
0
Scope surprises: every study declared its limits upfront
The problem

Research pays off only when the study is built to answer the real question. The wrong method, or framing so loose that any result confirms the plan, wastes participants and yields findings nobody can act on.

What I did

I matched method to question: moderated think-aloud sessions to understand reasoning and probe hesitation, unmoderated prototype tests for independent task completion at volume. Study goals were written as falsifiable hypotheses with explicit success metrics, scope lines made clear what each study could and couldn't produce, and discussion guides kept sessions consistent, so no participant was wasted and no finding arrived unactionable.

Write the hypothesis as something the study could disprove, then draw the scope lines before any session runs. Study design principle
MATCH THE METHOD TO THE QUESTION THE QUESTION PICKS THE METHOD, NOT HABIT RESEARCH QUESTION ASKING WHY, OR HOW OFTEN? WHY HOW OFTEN MODERATED REASONING, PROBE HESITATION UNMODERATED TASK COMPLETION AT VOLUME THINK-ALOUD SESSIONS PROTOTYPE TESTS BUILT TO BE DISPROVED SCOPE SET BEFORE ANY SESSION RUNS FALSIFIABLE HYPOTHESIS SOMETHING THE STUDY COULD PROVE WRONG EXPLICIT SUCCESS METRIC DECIDED UP FRONT, NOT AFTER SCOPE LINES WHAT THE STUDY CAN AND CANNOT ANSWER RUN THE SESSIONS
Fig. The method matched to the question, the study built to be disproved
THREAD 05 / 05
Multi-source demand synthesis
340
Logged product requests coded into themes across an 18-month window
3
Independent channels triangulated: sales log, research repository, meeting notes
20+
Enterprise accounts represented across the channels
The problem

Demand arrived through three disconnected channels: a sales request log, a research repository, and live meeting notes. Each has its own bias. The log favors the loudest account, the repository favors whoever was studied, the meetings favor whoever was in the room. Prioritizing from one channel meant building for the loudest voice.

What I did

I synthesized each channel on its own terms: 340 logged requests from an 18-month window coded into weighted themes, a multi-year research repository mined, months of meetings across 20+ enterprise accounts distilled. Cross-reading the three produced one demand picture where confidence comes from agreement: themes confirmed by all three channels became high-confidence priorities, two-channel themes carried their caveats, and single-channel signals were held for validation. The roadmap now prioritizes from evidence no single channel could provide.

A theme that survives the sales log, the research repository, and the meeting notes isn't an anecdote anymore. It's demand. Triangulation principle
ALL THREE TWO CHANNELS ONE CHANNEL THREE CHANNELS EACH WITH ITS OWN BIAS SALES REQUEST LOG 340 REQUESTS, 18 MONTHS RESEARCH REPOSITORY MULTI-YEAR STUDY ARCHIVE MEETING NOTES 20+ ENTERPRISE ACCOUNTS SYNTHESIZED SEPARATELY FIRST, SO NO CHANNEL SPEAKS FOR ANOTHER CROSS-READ ONE DEMAND PICTURE CONFIDENCE SET BY AGREEMENT, NOT VOLUME VALIDATED BY ALL THREE CHANNELS HIGH-CONFIDENCE ROADMAP PRIORITIES TWO-CHANNEL THEMES CARRIED FORWARD WITH THEIR CAVEATS SINGLE-CHANNEL SIGNALS HELD FOR VALIDATION, NOT PROMOTED FEEDS ROADMAP PRIORITY THE STANDING REFERENCE FOR WHAT CUSTOMERS ACTUALLY DEMAND
Fig. Three biased channels cross-read into one demand picture, confidence set by agreement
A-4 / CASE STUDY
← ALL CASE STUDIES