Brierley, C, Sawalha, M, Heselwood, B et al. (1 more author) (2016) A Verified Arabic-IPA Mapping for Arabic Transcription Technology, Informed by Quranic Recitation, Traditional Arabic Linguistics, and Modern Phonetics. Journal of Semitic Studies, 61 (1). pp. 157-186. ISSN 1477-8556
Abstract
In this paper, we present a detailed mapping from the graphemes of Modern Standard Arabic (MSA) to symbols from the International Phonetic Alphabet (IPA) for automated transcription of Arabic text. This mapping is distinctive in several ways. First, the corpus used in rule development is the full text of the Qur’ān rendered in fully pointed MSA. Second, we validate our scheme via automaticallygenerated frequency distributions of Arabic letters and diacritics over the whole corpus to anticipate and disambiguate non-trivial, compound grapheme-to-phoneme events, thus reducing the number of letter-to-sound rules. Such difficult cases include: the definite article; the letters alif, wāw, and yāʼ; the variant forms of hamza; the tanwīn case mark; and words with special pronunciations. Finally, our mapping scheme is informed by theory and practice from medieval Arabic linguistics and traditional Quranic recitation or tajwīd; we make a novel contribution with new translations for ancient terms which incorporate concepts familiar to modern phoneticians. Our principal objective in automating Arabic-IPA transcription is to generate phonemic citation forms of Arabic words to enhance Arabic dictionaries, to facilitate Arabic language learning, and for natural language engineering applications.
Metadata
Item Type: | Article |
---|---|
Authors/Creators: |
|
Keywords: | Quran |
Dates: |
|
Institution: | The University of Leeds |
Academic Units: | The University of Leeds > Faculty of Engineering & Physical Sciences (Leeds) > School of Computing (Leeds) |
Funding Information: | Funder Grant number EPSRC EP/K015206/1 |
Depositing User: | Symplectic Publications |
Date Deposited: | 12 Mar 2015 15:24 |
Last Modified: | 01 Mar 2017 13:26 |
Published Version: | http://dx.doi.org/10.1093/jss/fgv035 |
Status: | Published |
Publisher: | Oxford University Press (OUP): Policy K - Oxford Open Option D |
Identification Number: | 10.1093/jss.fgv035 |
Related URLs: | |
Open Archives Initiative ID (OAI ID): | oai:eprints.whiterose.ac.uk:84153 |