This model is a fine-tuned version of
google-t5/t5-base on an MRQA sample.
It achieves the following results on the evaluation set:
T5 base but trained at FP16 in the MRQA sample dataset.
This model is the checkpoint at 3000 steps (3rd epoch), because there were instabilities during the late epochs.
Note that this model is the checkpoint at 3000 steps (3rd epoch), because there were instabilities during the late epochs.