Product discovery and search quality

Rate the search. Find the cause. Fix the ranking.

Buyers judge a marketplace by what comes back when they type. We rate the search experience the way a buyer would, trace poor results to their root cause and propose the fix to the ranking or the model. Agents read every query at volume. Named search analysts own the rating and the recommendation. Every change to a ranking is approved by the marketplace and kept on a record.

spine·vision v4 · search
queries 12,400 · conf 0.96
Search results audit · top queries · batch #33104 of 12,400
relevant · excellent
Thin laptop on a desk
Thin and light laptop, 14 inch
"laptop 14 inch"rank 1
irrelevant · flagged
Aviator sunglasses
Aviator sunglasses
"reading glasses"rank 1
attributes fixed
Trail running shoe
Trail running shoe
"trail shoes men"rank 3 → 1
relevant · good
Field watch with a black dial
Field watch, black dial
"black watch"rank 2
3 of 4 queries rated good or better · 1 root cause openRe-run analysis
result · irrelevant 0.93
query:"reading glasses" top:sunglasses → irrelevant · intent:vision aid · cause:attribute gap · fix:proposed → approval · query:"laptop 14 inch" rated:excellent · query:"trail shoes men" attrs:backfilled → rank 1 · audit:chained ✓ ·
Search results audit, batch #3310 · rendered by spine·vision
then a named human approves the ranking change →
SPINE · SEARCH QUALITY · batch #3310live
step 02Rate 12,400 head and torso queriesread only · 1,140 rated below averageAuto · read
step 05Apply 36 ranking and attribute changesrouted to the marketplace’s search ownerHeld · human
✓ approved · named reviewer · audit entry signed
Every query rated, every ranking change on the record Ratings five-point scale Root cause tested, not guessed Fixes proposed to the model Ranking changes owner-approved Audit tamper-evident
The five-step loop

Rate, locate, explain, fix, re-run.

Search quality is a loop, not a project. Each pass starts with a buyer’s rating and ends with proof that the change worked. It is one of the practices in our marketplace operations group and it runs on the same record as the rest.

01

Gauge the customer experience

Analysts rate the search experience for each query on a five-point scale: poor, bad, average, good or excellent. The rating is a buyer’s judgement, recorded with the query, the page and the date.

02

Identify the issues

For queries that rate below average, the team marks what is wrong on the results page: irrelevant products, missing attributes, duplicate listings, thin results or a wrong category.

03

Find the root cause

Tests are run to discover why the results drew fewer impressions or clicks than they should have. A broken synonym, a missing attribute, a stale ranking signal or a model that misread the intent.

04

Suggest the fix

Recommendations go to the search and machine learning teams as specific proposals: an attribute to backfill, a synonym to add, a ranking weight to revisit, a training example to correct.

05

Re-run the analysis

After the change ships, the same queries are rated again on the same scale. The before and the after sit side by side on the record, so the efficiency of the change is shown rather than claimed.

Head and torso queries

Automate the head. Optimise the attributes that decide the rank.

A small set of head queries carries most of the traffic and a long torso carries most of the variety. Agents cluster and classify both by intent, so analysts spend their time on the queries that matter rather than on sorting them. Query automation is available as a pilot.

  • Search query analysisQueries are grouped by intent and volume, and the top results for each are pulled with their impressions and clicks, so an analyst opens a case rather than a spreadsheet.
  • Head and torso query automationThe highest-volume queries and the long middle are monitored continuously. A drop in relevance on any of them opens a case for a named analyst. Available as a pilot.
  • Key attribute optimisationWhere the root cause is a listing rather than the algorithm, the missing or wrong attributes are drafted for the seller or the catalog team and applied only after approval.
  • Fixes for the top queries by intentA buyer typing "reading glasses" wants a vision aid, not sunglasses. The intent is read, the mismatch is documented and the fix is proposed with the evidence attached.
spine·vision // query audit · intent and ratingrows 5 of 12,400_
Search · head and torso queriesbatch #3310
QueryIntentTop resultRatingAction
"laptop 14 inch"specLaptopexcellentnone
"reading glasses"visionAviatorspoorheld
"trail shoes men"buySneakergoodfixed
"black watch"browseWatchgoodnone
"coffee mug set"bundleMugaveragesynonym
2 fixes proposed · 0 applied without approvalExport report
Query audit, batch #3310 · rendered by spine·vision
The buyer’s perception

Judge the results the way a buyer does.

Relevance is not a score in a log. It is what a person sees on page one, on page two and in the strip of things bought together. We rate all of it against agreed criteria and keep each rating with the evidence behind it.

01

Rate results against criteria

Each query’s results are scored against a written rubric: does the top result match the intent, are the attributes right, is the price band sensible, is anything missing that a buyer would expect.

02

Browse beyond page one

Analysts browse several pages of output for each query, because the tail of a results page tells you how the ranking degrades and where the duplicates and the mis-tagged products hide.

03

Rate the recommendation engine

Recommendations on product and cart pages are rated on the same scale as search. A relevant result followed by an odd recommendation costs the sale just as surely.

04

Analyse "bought together"

Bought-together and similar-item strips are checked for sense, for category fit and for the accidental pairings that a model learns from a few noisy orders.

05

Competitive information architecture

How competing marketplaces structure their categories and filters for the same products, documented as a comparison your category team can act on.

06

Competitive search results

The same queries run on competing platforms, rated on the same scale, so you know where your results are ahead, where they are behind and by how much in the buyer’s eyes.

Who changes the ranking

The team proposes. The marketplace approves.

A ranking change touches every buyer at once, so it is never made by an agent and never made quietly. Search quality runs on Spine, the same governed platform behind our regulated work since 1955.

01 · Proposal

Proposed by a named analyst, with the analysis attached

Every fix arrives as a proposal: the queries affected, the rating before, the root cause found, the tests run and the change recommended. The analyst who owns it is named on the proposal.

02 · Approval

Approved by the marketplace’s owner, on the record

The search or category owner on your side approves, amends or declines. The decision, the person and the time are written to a chained audit record, and the re-run that follows is filed against the same entry.

live
Computing
held
Eyewear
fixed
Footwear
live
Watches
live
Home
Talk to us

Let’s build smarter, faster and more scalable e-commerce experiences together.

Explore how AI + creativity can accelerate your next big move. Bring your top queries. Leave with a rating for each, the root causes behind the poor ones and a set of fixes ready for your search owner to approve.