Using LLMs to Identify Problematic Search Queries
This article explores a practical use of LLMs in search relevance work: identifying poorly performing queries for human review. It describes the evaluation pipeline and compares results from a cross-encoder and a generative model using sample ecommerce searches.