The Basis-Latin-French dataset is an unannotated Latin and old French corpus, compiled from different resources from the web. This resources include the Corpus de la Bourgogne du Moyen Âge, The e-NDP project, HIMANIS Guérin and the HOME-Alcar project and the Corpus Cisterciens et Ressources.