Statistical Machine Translation for Indian Languages: Mission Hindi 2

Prakash B. Pimpale; Raj Nath Patel

arxiv: 1610.08000 · v1 · pith:V75BUVIDnew · submitted 2016-10-25 · 💻 cs.CL

Statistical Machine Translation for Indian Languages: Mission Hindi 2

Raj Nath Patel , Prakash B. Pimpale This is my paper

classification 💻 cs.CL

keywords indianlanguagesmachinestatisticaltranslationcontestadvancedbengali-hindi

0 comments

read the original abstract

This paper presents Centre for Development of Advanced Computing Mumbai's (CDACM) submission to NLP Tools Contest on Statistical Machine Translation in Indian Languages (ILSMT) 2015 (collocated with ICON 2015). The aim of the contest was to collectively explore the effectiveness of Statistical Machine Translation (SMT) while translating within Indian languages and between English and Indian languages. In this paper, we report our work on all five language pairs, namely Bengali-Hindi (\bnhi), Marathi-Hindi (\mrhi), Tamil-Hindi (\tahi), Telugu-Hindi (\tehi), and English-Hindi (\enhi) for Health, Tourism, and General domains. We have used suffix separation, compound splitting and preordering prior to SMT training and testing.

This paper has not been read by Pith yet.

Statistical Machine Translation for Indian Languages: Mission Hindi 2

discussion (0)