12

Search across five brands,
two viewports

Expert benchmark · FindabilityEstée Lauder Companies
UX ResearchExperience StrategyProduct & Design LeadershipAI & Insight Systems
Methods
Expert evaluation Competitive benchmark Heuristic grading Hypothesis development
No participants and no sessions. This is expert evaluation against a published standard.
Inputs and tools
Baymard Institute guidelines 5 commerce sites Desktop and mobile, scored separately Internal search analytics
The analytics leg ran at the time. Those files were not retained, so no figures from it appear here.
The problem

Every brand in the portfolio had built its own site search independently. The interactions diverged, the predictive overlays diverged, and nobody could say which decisions were deliberate and which were inherited. Product wanted a single global standard for search, and wanted it defensible rather than preferred.

My role

Led the analysis end to end as Director of UX Strategy and Analysis: framed the questions, wrote the hypotheses before the audit, graded the experiences, and delivered the global recommendations that fed the design and A/B roadmap.

Methodology
  • Structured expert review, desktop and mobile evaluated separately for every brand
  • Competitive benchmark across five commerce sites, in and out of category
  • Graded against Baymard Institute search and autocomplete guidelines
  • Three hypotheses written before the audit, then tested against what the audit found
What this is, and is not

No participants and no sessions. This is expert evaluation against an external standard, run to generate testable hypotheses rather than to settle them. The study was paired at the time with internal search analytics; those files did not leave the company with me, so no figures from that leg appear here.

The research and analysis model showing where this work sat in the discover, define, develop, deliver sequence
/ Insight
12 / Search experience

The richest thing we were building on mobile was the thing nobody could see.

What the audit found
  • On a phone, the keyboard opens with the search field and covers the predictive product cards below it
  • Most people submit the query before scrolling, so the cards are rendered and never viewed
  • Every brand in the portfolio was investing design and engineering effort in that space anyway
  • The same overlay, on desktop, is fully visible and genuinely useful
Why it mattered
  • It reframed the mobile question from how many product cards to whether product cards
  • It made the case for text-first predictive results on small screens
  • It turned a styling debate into a spend question with a clear answer
Desktop and mobile were being designed as one surface. They are not one surface, and search is where that shows first.
/ Written before the audit
12 / Search experience

Four questions, three hypotheses,
committed to in advance

Stating the position first is what makes an expert review checkable. Two of the three held, one was overturned by the evidence, and the one that was overturned is the reason the recommendations changed the roadmap rather than confirming it.

H1 · Entry point is a brand choice

Exposed field or magnifying glass icon should be left to each brand, because hiding the field signals that category navigation is the intended path.

Held

Three of five competitors used an exposed field, two used an icon, and neither produced a worse experience.

H2 · Cap products at four

No more than four product results in the predictive overlay on desktop, and on mobile focus on text results only.

Held, and sharpened

Four became the portfolio standard. The mobile half went further than the hypothesis: not fewer cards, but a different model entirely.

H3 · Hide content behind a tab

Editorial and story content should sit behind a tab in the predictive overlay rather than alongside products.

Overturned in part

Exposing content made the overlay harder to read, so the tab survived, but the reasoning changed from hierarchy to legibility.

The business objective, four research questions, and three hypotheses as stated at the start of the engagement
The objective, the four questions, and the three hypotheses, as they were written at kickoff before any grading began.
The research package taxonomy showing which methods were selected for this engagement
The method taxonomy the team worked from. Quantitative analytics and competitive benchmarking were the two packages selected here, out of twelve available.
/ The benchmark
12 / Search experience

Five sites, six attributes,
both viewports each

Two brands from the portfolio, two beauty competitors, and one furniture retailer included deliberately: the goal was to see how search behaves as a commerce pattern, not just how beauty does it. Every site was walked twice, once on desktop and once on a phone, and scored separately.

Desktop and mobile search overlay compared side by side, with pros and cons listed for each viewport
Every brand was written up this way: desktop and mobile side by side, with the failures separated by viewport rather than merged. Oversized product cards, a long scroll, and a search input that does not stick were mobile problems only.
A second brand's desktop and mobile search overlay compared, showing inconsistency between viewports
The second portfolio brand surfaced a different class of problem: an add-to-favorites control that existed on mobile and not on desktop. Divergence between viewports inside one brand, not just across brands.
Competitor features matrix scoring five brands across six search attributes
The matrix that made the argument portable. Six attributes, five sites, one row each. This is what a stakeholder could take into a prioritization meeting without me in the room.
Five benchmark conclusions summarizing the competitive scan
The conclusions stated as counts rather than impressions: three of five exposed the field, two of five exposed content suggestions, one of five offered a clear control on the input.
/ Graded against a standard
12 / Search experience

Opinion is cheap in an audit.
The standard is what makes it hold.

Grading against published Baymard guidelines rather than personal judgment is what let this survive a room full of brand teams who each believed their own search was fine. When a finding is a guideline violation, the conversation moves from whether to when.

Baymard-derived recommendations on search entry point and product suggestions
The guideline reasoning behind two of the portfolio decisions: when category navigation is the dominant path, a quieter search field is a legitimate choice rather than a mistake.
Search audit table listing findings that applied to all brands and findings specific to individual brands
Findings split into what applied to every brand and what applied to one. Portfolio-wide items became the global standard; brand-specific items went to that brand's backlog.
Autocomplete queries first
Query suggestions are the primary output. Products and categories are additions to that, not replacements for it.
Persist the query
The search term should survive onto the results page. Several brands were dropping it.
Monitor zero-result logs
Recurring empty queries are a content and mapping problem, not a user error.
Support non-product search
People search for help, stores, and policies in the same field. Only some brands returned anything.
/ The standard that came out of it
12 / Search experience

Five decisions, applied
across every brand site.

Each one is a default with a stated reason, which matters in a portfolio: brand teams can deviate, but they have to say why, and the reason is on the record for the next person who asks.

01 Entry point

Exposed field or icon is a brand decision. Neither harms the experience, because category navigation is where most product discovery starts on a beauty site.

02 Suggestions

Ten suggested terms as the ceiling, scaled to how many categories and products the brand actually carries.

03 Product cards

Four maximum on desktop, to keep focus on the query. On mobile, text first, because the keyboard covers the cards and people submit before they scroll.

04 Content

A tab to switch between products and editorial, matched to the same count as products. Exposing both at once made the overlay unreadable.

05 Card contents

Drop add-to-bag from the overlay, since size and shade selection happens on the product page anyway. Add ratings, which users treat as a filter rather than a detail.

A portfolio does not need every brand to search the same way. It needs every brand to have a reason.
/ What I would do differently
12 / Search experience

Three things this study
could not settle

It never met a user

Expert review generates hypotheses well and confirms them badly. The mobile keyboard finding is the strongest thing here and it is still an inference from interaction design, not an observation of behavior. Five unmoderated sessions per viewport would have settled it in a week, and I would build that in from the start now.

The competitor set was small

Five sites is enough to see a pattern and not enough to weight it. Including a furniture retailer was the right instinct, and one out-of-category site out of five gives that instinct very little to stand on.

No measurement plan attached

The recommendations shipped as a standard without a defined success metric per decision. On the benchmark audit I ran later, every activation carried its own measurement plan, which is the direct lesson from this one.