The SemCacheSearchQueries benchmark is designed to evaluate semantic caching in open-domain search applications. Large-scale search engines, such as Google, increasingly rely on LLMs to generate direct answers to natural language queries. While this improves user experience, it introduces significant latency and cost, particularly at the scale of millions of daily queries. Many queries issued to search engines are paraphrased variations of earlier inputs, making semantic caching a natural fit… See the full description on the dataset page:
https://huggingface.co/datasets/vCache/SemBenchmarkSearchQueries.