pith. machine review for the scientific record. sign in

arxiv: 1810.10254 · v2 · submitted 2018-10-24 · 💻 cs.CL

Recognition: unknown

Learn to Code-Switch: Data Augmentation using Copy Mechanism on Language Modeling

Authors on Pith no claims yet
classification 💻 cs.CL
keywords code-switchinglanguagemodelsentencesdataparalleltrainingalign
0
0 comments X
read the original abstract

Building large-scale datasets for training code-switching language models is challenging and very expensive. To alleviate this problem using parallel corpus has been a major workaround. However, existing solutions use linguistic constraints which may not capture the real data distribution. In this work, we propose a novel method for learning how to generate code-switching sentences from parallel corpora. Our model uses a Seq2Seq model in combination with pointer networks to align and choose words from the monolingual sentences and form a grammatical code-switching sentence. In our experiment, we show that by training a language model using the augmented sentences we improve the perplexity score by 10% compared to the LSTM baseline.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.