Published January 1, 2013 | Version v1
Conference paper Open

Rewriting Turkish texts written in English alphabet using Turkish alphabet

  • 1. TUBITAK BILGEM, Bilisim & Bilgi Guvenligi Ileri Teknol Arastirma, TR-41470 Kocaeli, Turkey
  • 2. Cumhuriyet Univ, Dept Comp Engn, Sivas, Turkey
  • 3. Dept Comp Engn, Comp Vis Lab, Kocaeli, Turkey

Description

Turkish texts written by English characters are easily comprehended by people, although performing this process by machines is still one of the unsolved Word Sense Disambiguation problems. Rewriting texts in English characters using Turkish characters is a natural language processing problem special to Turkish. Choosing the right Turkish word among different alternatives requires consideration of the text semantically. In this study, the effect of examination of the text either sentence or whole text based, on the right word determination is investigated. Performance of machine learning methods and statistical methods in right word determination is examined. The study is tested on randomly selected news texts. It is shown that examination of the text as a whole provides more information compared to sentence based methods and machine learning methods provides better results compared to statistical studies.

Files

bib-da6e2181-ed25-420f-89d5-a5f6f55cf97b.txt

Files (192 Bytes)

Name Size Download all
md5:893f234cd019aa854cc5754fd433b36e
192 Bytes Preview Download