The Emergent Language Corpus Collection is collection of corpora and metadata
from a variety of emergent communication simulations.
You can clone this repository with git LFS and use the data directly or load
the data via the mlcroissant library. To install the mlcroissant library and
necessary dependencies, see the conda environment at util/environment.yml.
Below we show an example of loading ELCC's data via mlcroissant.
import mlcroissant as mlc
cr_url… See the full description on the dataset page:
https://huggingface.co/datasets/bboldt/elcc.