A 100-sample subset of the LDJnr/Capybara dataset.
This is a small, reproducible subset of the Capybara conversational dataset containing 100 randomly selected examples. It was created by shuffling the original dataset with seed 42 and selecting the first 100 examples.
Source: LDJnr/Capybara
Total original examples: 16,006
Subset size: 100
Shuffle seed: 42
Columns: source, conversation
from datasets… See the full description on the dataset page:
https://huggingface.co/datasets/lewtun/Capybara-100.