A question-answering dataset for supervised fine-tuning of language models on Formula 1 knowledge, 1950–2025.
This is the training corpus behind machina-sports/ayrton-1.
Race results, championship standings, driver/constructor history (1950–2025, backed by Jolpica-F1).
Session-level telemetry and strategy — lap times, pit stops, stints, compound usage, top speeds (2018–2025, backed by FastF1).… See the full description on the dataset page:
https://huggingface.co/datasets/machina-sports/ayrton-1-qa-v2.