Nakdan: Professional Hebrew Diacritizer

Avi Shmidman, Shaltiel Shmidman, Moshe Koppel, Yoav Goldberg


Abstract
We present a system for automatic diacritization of Hebrew Text. The system combines modern neural models with carefully curated declarative linguistic knowledge and comprehensive manually constructed tables and dictionaries. Besides providing state of the art diacritization accuracy, the system also supports an interface for manual editing and correction of the automatic output, and has several features which make it particularly useful for preparation of scientific editions of historical Hebrew texts. The system supports Modern Hebrew, Rabbinic Hebrew and Poetic Hebrew. The system is freely accessible for all use at http://nakdanpro.dicta.org.il
Anthology ID:
2020.acl-demos.23
Volume:
Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics: System Demonstrations
Month:
July
Year:
2020
Address:
Online
Editors:
Asli Celikyilmaz, Tsung-Hsien Wen
Venue:
ACL
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
197–203
Language:
URL:
https://aclanthology.org/2020.acl-demos.23
DOI:
10.18653/v1/2020.acl-demos.23
Bibkey:
Cite (ACL):
Avi Shmidman, Shaltiel Shmidman, Moshe Koppel, and Yoav Goldberg. 2020. Nakdan: Professional Hebrew Diacritizer. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics: System Demonstrations, pages 197–203, Online. Association for Computational Linguistics.
Cite (Informal):
Nakdan: Professional Hebrew Diacritizer (Shmidman et al., ACL 2020)
Copy Citation:
PDF:
https://aclanthology.org/2020.acl-demos.23.pdf
Software:
 2020.acl-demos.23.Software.zip
Dataset:
 2020.acl-demos.23.Dataset.pdf
Video:
 http://slideslive.com/38928593