Nakdan: Professional Hebrew Diacritizer

Avi Shmidman,Shaltiel Shmidman,Moshe Koppel,Yoav Goldberg
DOI: https://doi.org/10.48550/arXiv.2005.03312
2020-05-07
Abstract:We present a system for automatic diacritization of Hebrew text. The system combines modern neural models with carefully curated declarative linguistic knowledge and comprehensive manually constructed tables and dictionaries. Besides providing state of the art diacritization accuracy, the system also supports an interface for manual editing and correction of the automatic output, and has several features which make it particularly useful for preparation of scientific editions of Hebrew texts. The system supports Modern Hebrew, Rabbinic Hebrew and Poetic Hebrew. The system is freely accessible for all use at <a class="link-external link-http" href="http://nakdanpro.dicta.org.il" rel="external noopener nofollow">this http URL</a>.
Computation and Language
What problem does this paper attempt to address?