Overview
Discovery API
Developers
Shopping Intelligence
Case studies
Canonical benchmark · Flights + Accommodation · 12 scenarios · September 2026

A smaller shopping space that still holds the right options.

Discovery removes 67.7% of the logical shopping space across flight and accommodation discovery while retaining 67.5% Recall@10 in the two-track mean. Reduction and relevance are reported together because neither number means anything alone.
Internal measurements across a 12-scenario canonical set.
Why relevance is the constraint

The reduction is only as good as what survives it.

Reduction alone proves nothing

Returning no candidates reduces the shopping space by 100%. Any reduction figure published without a relevance figure beside it is unfalsifiable.

Recall@10 shows the good options survived

At the ten-candidate cutoff Discovery retains a mean 67.5% of relevant options across the equally sized Flights and Accommodation tracks, including 69.5% on Flights.

NDCG@10 shows they are ranked, not just present

Mean NDCG@10 is 67.8% across the two reported tracks, including 76.3% on Flights. Relevant options appear near the top of the shortlist rather than at its tail.
Reduction on its own is not a result: returning nothing reduces the shopping space by 100%. These reductions are reported only alongside the recall and ranking quality measured on the same runs, which is what shows the discarded space was the part that did not matter.
Headline metric

Logical shopping-space reduction

1 - selected_destination_date_units / broad_destination_date_units
Numerator
Exact destination/date units returned by Discovery.
Denominator
Unique destination/date units actually exposed to the broad provider-backed lifecycle. Not planner expansion, not the candidate limit, and not provider HTTP calls.
Aggregation
Twelve scenarios across two reported tracks: Flights and Accommodation. Mean figures are the arithmetic mean of the track metrics because each track contains the same number of scenarios.
Per-track results

Full results across both shortlist cutoffs.

Results use the ten-candidate cutoff because that is the shortlist size Discovery is typically integrated at. Recall measures whether relevant options survived; NDCG measures whether they are ranked near the top.

Flights

Recall@5
66.0%
Recall@10
69.5%
NDCG@5
75.4%
NDCG@10
76.3%
Shopping-space reduction
67.9%

Accommodation

Recall@5
48.0%
Recall@10
65.5%
NDCG@5
55.9%
NDCG@10
59.2%
Shopping-space reduction
67.5%
What we do not publish

Provider-side reduction is measured but withheld.

Provider-side reduction — fresh HTTP requests, fresh provider request units, total provider units and configured estimated cost — uses the same formula but is not published. It is reported only when both arms have comparable cache state and complete accounting; unknown request counts remain unavailable and are never counted as zero.
Run it on your own traffic

The only benchmark that matters is yours.

A 4–6 week pilot measures recall, ranking quality and shopping-space reduction against your own scenarios rather than ours.

MyEscapePlan Travel Intelligence

Travel intent infrastructure: natural-language discovery, ranked shortlists and provider-backed verification, plus travel shopping data and market intelligence.
MYESCAPEPLAN LTD · Company no. 16395758 · Registered in United Kingdom
Discovery API is available now. Destination & Market Shopping Intelligence is coming soon.
Discovery benchmark figures, their formula and their limitations are published in full.
Read the benchmark methodology ↗
Shopping Intelligence case studies are a separate release, and will publish with the Competitive Shopping Audit benchmark.
Legal
Terms & conditions
Privacy
Cookies