George Sanchez


pdf bib
Sentence Boundary Detection in Legal Text
George Sanchez
Proceedings of the Natural Legal Language Processing Workshop 2019

In this paper, we examined several algorithms to detect sentence boundaries in legal text. Legal text presents challenges for sentence tokenizers because of the variety of punctuations and syntax of legal text. Out-of-the-box algorithms perform poorly on legal text affecting further analysis of the text. A novel and domain-specific approach is needed to detect sentence boundaries to further analyze legal text. We present the results of our investigation in this paper.