A multilingual dataset for training and evaluating query parsing models that convert natural language queries into structured JSON for R&D project semantic search.
This dataset was created as part of the IMPULS project (AINA Challenge 2024), a collaboration between SIRIS Academic and Generalitat de Catalunya to build a multilingual semantic search system for R&D ecosystems.
The dataset contains natural language queries in… See the full description on the dataset page:
https://huggingface.co/datasets/SIRIS-Lab/impuls-query-parsing.