Query
Capture the exact input like a customer used, including spelling and capitalization where relevant.
A representative test set makes ranking changes reproducible, protects strong queries and gives vendors the same practice test.
Capture the exact input like a customer used, including spelling and capitalization where relevant.
For example, classify brand, SKU, category, attribute, long-tail, typo, natural language or problem query.
Describe what the user is trying to find or achieve, regardless of the current results.
Write down which products, categories or properties are demonstrably relevant.
Combine volume, buying intent, strategic importance and regression risk.
Capture current results, first relevant position and known issues.
Highlight critical queries that need to be checked again with each relevant release.
Capture store, market, device, or other context when the same query needs to be judged differently.
Make it clear who can adjust judgments, priority and expected outcomes.
Include top queries, but also long-tail, typos, zero-result searches, technical codes and known regression risks. Otherwise, you only optimize for the easy part.