This dataset has been done by M2 student in computational linguistics at the Paris Cité university.
It is part of a project for the course "NLP in Industry". The annotation are handmade and the gold should represent the publications date of the extracted text document.