 Corpus Name: CAPES
     Package: CAPES.en-pt in Moses format
     Website: http://opus.nlpl.eu/CAPES-v1.php
     Release: v1
Release date: Mon Jun 17 05:42:59 EEST 2019
     License: CC–BY 4.0

This corpus is part of OPUS - the open collection of parallel corpora
OPUS Website: http://opus.nlpl.eu

Please cite the following article if you use any part of the corpus in your own work: J. Tiedemann, 2012, Parallel Data, Tools and Interfaces in OPUS. In Proceedings of the 8th International Conference on Language Resources and Evaluation (LREC 2012)

Parallel corpus of theses and dissertation abstracts in Portuguese and English from CAPESSource: https://sites.google.com/view/felipe-soares/datasets
A parallel corpus of theses and dissertations abstracts in English and Portuguese were collected from the CAPES website (Coordenação de Aperfeiçoamento de Pessoal de Nível Superior) - Brazil. The corpus is sentence aligned for all language pairs. Approximately 240,000 documents were collected and aligned using the Hunalign algorithm.  DOI: https://doi.org/10.6084/m9.figshare.5995519.v2If you use this corpus, please cite the following work: @inproceedings{soares2018parallel,  title={A Parallel Corpus of Theses and Dissertations Abstracts},  author={Soares, Felipe and Yamashita, Gabrielli Harumi and Anzanello, Michel Jose},  booktitle={International Conference on Computational Processing of the Portuguese Language},  pages={345--352},  year={2018},  organization={Springer}}
