Although there are increasing and significant ties between China and Portuguese-speaking countries, there is not much parallel corpora in the Chinese–Portuguese language pair. Both languages are very populous, with 1.2 billion native Chinese speakers and 279 million native Portuguese speakers, the language pair, however, could be considered as low-resource in terms of available parallel corpora. In this paper, we describe our methods to curate Chinese–Portuguese parallel corpora and evaluate their quality. We extracted bilingual data from Macao government websites and proposed a hierarchical strategy to build a large parallel corpus. Experiments are conducted on existing and our corpora using both Phrased-Based Machine Translation (PBMT) an...
The translation quality of Neural Machine Translation (NMT) systems depends strongly on the training...
Data sparsity is one of the challenges for low-resource language pairs in Neural Machine Translation...
Although, Chinese and Spanish are two of the most spoken languages in the world, not much research h...
Although there are increasing and significant ties between China and Portuguese-speaking countries, ...
Treball de fi de màster en Lingüística Teòrica i Aplicada. Directora: Dra. Maite Melero i NoguésNeur...
Parallel corpus is a valuable resource for cross-language information retrieval and data-driven natu...
Treball de fi de màster en Lingüística Teòrica i Aplicada. Directora: Dra. Maite MeleroThe lack of p...
In an increasingly globalized world, being able to understand texts in different languages (even mor...
Machine translation (MT), as a high level application of natural language pro-cessing (NLP), is a po...
Two of the most popular Machine Translation (MT) paradigms are rule based (RBMT) and corpus based, w...
Neural machine translation (NMT) has been a mainstream method for the machine translation (MT) task....
Thesis (Master's)--University of Washington, 2018In an emergency, machine translation systems can be...
This paper discusses the role played by parallel corpora in the design and implementation of fully a...
Abstract. This article presents and describes an experimental proto-type system for performing Chine...
This article presents some experimental results on Chinese to Spanish machine translation. The imple...
The translation quality of Neural Machine Translation (NMT) systems depends strongly on the training...
Data sparsity is one of the challenges for low-resource language pairs in Neural Machine Translation...
Although, Chinese and Spanish are two of the most spoken languages in the world, not much research h...
Although there are increasing and significant ties between China and Portuguese-speaking countries, ...
Treball de fi de màster en Lingüística Teòrica i Aplicada. Directora: Dra. Maite Melero i NoguésNeur...
Parallel corpus is a valuable resource for cross-language information retrieval and data-driven natu...
Treball de fi de màster en Lingüística Teòrica i Aplicada. Directora: Dra. Maite MeleroThe lack of p...
In an increasingly globalized world, being able to understand texts in different languages (even mor...
Machine translation (MT), as a high level application of natural language pro-cessing (NLP), is a po...
Two of the most popular Machine Translation (MT) paradigms are rule based (RBMT) and corpus based, w...
Neural machine translation (NMT) has been a mainstream method for the machine translation (MT) task....
Thesis (Master's)--University of Washington, 2018In an emergency, machine translation systems can be...
This paper discusses the role played by parallel corpora in the design and implementation of fully a...
Abstract. This article presents and describes an experimental proto-type system for performing Chine...
This article presents some experimental results on Chinese to Spanish machine translation. The imple...
The translation quality of Neural Machine Translation (NMT) systems depends strongly on the training...
Data sparsity is one of the challenges for low-resource language pairs in Neural Machine Translation...
Although, Chinese and Spanish are two of the most spoken languages in the world, not much research h...