DocumentCode
3540086
Title
Towards Sinhala Tamil machine translation
Author
Pushpananda, Randil ; Weerasinghe, Ruvan ; Niranjan, Mahesan
Author_Institution
Sch. of Comput., Language Technol. Res. Lab., Univ. of Colombo, Colombo, Sri Lanka
fYear
2013
fDate
11-15 Dec. 2013
Firstpage
288
Lastpage
288
Abstract
Statistical Machine Translation is a well established data-driven approach to translate source language text to target language text using statistical methods using bilingually aligned corpora. However there has been little research in this area for less well-resourced languages such as Sinhala and Tamil. The focus of this research is to investigate how translation performance varies with the amount of parallel training data in order to find out the minimum needed to develop a baseline machine translation system for the Sinhala-Tamil language pair.
Keywords
language translation; statistical analysis; Sinhala-Tamil language pair; bilingually aligned corpora; data-driven approach; machine translation system; parallel training data; source language text; statistical machine translation; statistical methods; target language text; translation performance; Buildings; Educational institutions; Laboratories; Measurement; Training; Tuning; Language Modeling; Sinhala Language Processing; Statistical Machine Translation; Tamil Language Processing;
fLanguage
English
Publisher
ieee
Conference_Titel
Advances in ICT for Emerging Regions (ICTer), 2013 International Conference on
Conference_Location
Colombo
Print_ISBN
978-1-4799-1275-9
Type
conf
DOI
10.1109/ICTer.2013.6761202
Filename
6761202
Link To Document