I just tried the new recognition model on Hebrew data from Sofer Mahir, it doesn't look to OCR anything just output an almost the same phrase with some variants. Something wrong with the tokenizer or the decoder?
Original:
New model output:
Data:
data.zip
I just tried the new recognition model on Hebrew data from
Sofer Mahir, it doesn't look to OCR anything just output an almost the same phrase with some variants. Something wrong with the tokenizer or the decoder?Original:
New model output:
Data:
data.zip