HAADS: A Hebrew Aramaic abbreviation disambiguation system

Yaakov HaCohen-Kerner, Ariel Kass, Ariel Peretz

Research output: Contribution to journalArticlepeer-review

24 Scopus citations

Abstract

In many languages abbreviations are very common and are widely used in both written and spoken language. However, they are not always explicitly defined and in many cases they are ambiguous. This research presents a process that attempts to solve the problem of abbreviation ambiguity using modern machine learning (ML) techniques. Various baseline features are explored, including context-related methods and statistical methods. The application domain is Jewish Law documents written in Hebrew and Aramaic, which are known to be rich in ambiguous abbreviations. Two research approaches were implemented and tested: general and individual. Our system applied four common ML methods to find a successful integration of the various baseline features. The best result was achieved by the SVM ML method in the individual research, with 98-07% accuracy.

Original languageEnglish
Pages (from-to)1923-1932
Number of pages10
JournalJournal of the American Society for Information Science and Technology
Volume61
Issue number9
DOIs
StatePublished - Sep 2010
Externally publishedYes

Fingerprint

Dive into the research topics of 'HAADS: A Hebrew Aramaic abbreviation disambiguation system'. Together they form a unique fingerprint.

Cite this