You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
@theraysmith commented 2 days ago
Update: after going back to the www to get fresh data, I believe that my corpus text is now good for:
chr
dzo
iku
snd
syr
tgk
tir
I have put a lot of time into cleaners/filters for languages that use 'virama' characters.
I am not convinced that they are perfect, but I will add the code to the github repo in due course, so experts/native speakers can offer suggestions/fixes to make them better.
Hello, wanted to request if it would it be possible to add a new font called joyig for Dzongkha which is more commonly used. The ttf file can be downloaded from here https://www.dzongkha.gov.bt/en/tools
dzo - dz - Dzongkha - http://crubadan.org/languages/dz
Ref: tesseract-ocr/tesseract#654 (comment)
The text was updated successfully, but these errors were encountered: