A dataset of 164 evaluation tasks for training and benchmarking RL agents on credit card optimization. Each task presents a user scenario with spending patterns, constraints, and preferences, and asks the agent to recommend optimal credit cards with expected value (EV) calculations.
This dataset is the task suite for the LexEnvs Harbor RL Environment, a stateless evaluation server that scores agent responses on… See the full description on the dataset page:
https://huggingface.co/datasets/Endishai-org/lexenvs-tasks.