This repository contains the data released alongside SLIP (Sensor Language-Informed Pretraining), a framework for learning language-aligned sensor representations that generalize across diverse sensor setups. It includes two components: (1) the pretraining corpus of 600K+ paired sensor time-series and hierarchical text captions, and (2) 11 downstream evaluation datasets spanning four sensor domains.