Paper:
https://arxiv.org/abs/2504.16438
This dataset contains speeches extracted from government institutions of the US, UK and Canada.
Each speech is a single partition and contains metadata such as speaker name, date, country, and the source URL.
This dataset is intended to evaluate federated learning algorithms. If using one speech as a client, then there are… See the full description on the dataset page:
https://huggingface.co/datasets/hazylavender/CongressionalDataset.