heb-synthtiger-16k (Synthetic Hebrew Printed Text Dataset)
Dataset Summary
heb-synthtiger-16k is a dataset containing 16,000 synthetic Hebrew text images generated using SynthTIGER with 11 different printed fonts.
The text in these images consists primarily of single words, sampled from the 10,000 most frequent Hebrew words, along with words from additional sources such as: