Paper | Code
This repository contains the LibriBrain data organised by book: MEG recordings (.h5), event annotations (.tsv), and the audiobook stimulus audio (.wav).
LibriBrain was first open-sourced as part of the 2025 PNPL Competition.
In addition, LibriBrain is used as a fine-tuning dataset in the paper "MEG-XL: Data-Efficient Brain-to-Text via Long-Context Pre-Training" to evaluate word decoding from brain data.
Sample Usage… See the full description on the dataset page: https://huggingface.co/datasets/pnpl/LibriBrain.