Utilize este identificador para referenciar este registo: http://hdl.handle.net/10071/16793
Registo completo
Campo DCValorIdioma
dc.contributor.authorBarreiro, A.-
dc.contributor.authorBatista, F.-
dc.date.accessioned2018-11-29T15:49:46Z-
dc.date.available2018-11-29T15:49:46Z-
dc.date.issued2018-
dc.identifier.isbn978-1-948087-54-4-
dc.identifier.urihttps://ciencia.iscte-iul.pt/id/ci-pub-50918-
dc.identifier.urihttp://hdl.handle.net/10071/16793-
dc.description.abstractThis paper performs a detailed analysis on the alignment of Portuguese contractions, based on a previously aligned bilingual corpus. The alignment task was performed manually in a subset of the English-Portuguese CLUE4Translation Alignment Collection. The initial parallel corpus was pre-processed and a decision was made as to whether the contraction should be maintained or decomposed in the alignment. Decomposition was required in the cases in which the two words that have been concatenated, i.e., the preposition and the determiner or pronoun, go in two separate translation alignment pairs (PT - [no seio de] [a União Europeia] EN - [within] [the European Union]). Most contractions required decomposition in contexts where they are positioned at the end of a multiword unit. On the other hand, contractions tend to be maintained when they occur at the beginning or in the middle of the multiword unit, i.e., in the frozen part of the multiword (PT - [no que diz respeito a] EN - [with regard to] or PT - [além disso] EN - [in addition]. A correct alignment of multiwords and phrasal units containing contractions is instrumental for machine translation, paraphrasing, and variety adaptationeng
dc.language.isoeng-
dc.publisherThe Association for Computational Linguistics-
dc.relationinfo:eu-repo/grantAgreement/FCT/3599-PPCDT/135285/PT-
dc.relationinfo:eu-repo/grantAgreement/FCT/SFRH/SFRH%2FBPD%2F91446%2F2012/PT-
dc.relationinfo:eu-repo/grantAgreement/FCT/5876/147282/PT-
dc.rightsopenAccess-
dc.titleContractions: to align or not to align, that is the questioneng
dc.typeconferenceObject-
dc.event.typeWorkshoppt
dc.event.locationSanta Feeng
dc.event.date2018-
dc.pagination122 - 130-
dc.peerreviewedyes-
dc.journalFirst Workshop on Linguistic Resources for Natural Language Processing, Coling 2018-
degois.publication.firstPage122-
degois.publication.lastPage130-
degois.publication.locationSanta Feeng
degois.publication.titleContractions: to align or not to align, that is the questioneng
dc.description.versioninfo:eu-repo/semantics/acceptedVersion-
dc.subject.fosDomínio/Área Científica::Ciências Naturais::Ciências da Computação e da Informaçãopor
dc.subject.fosDomínio/Área Científica::Engenharia e Tecnologia::Engenharia Eletrotécnica, Eletrónica e Informáticapor
dc.subject.fosDomínio/Área Científica::Humanidades::Línguas e Literaturaspor
Aparece nas coleções:CTI-CRI - Comunicações a conferências internacionais

Ficheiros deste registo:
Ficheiro Descrição TamanhoFormato 
LR4NLP_COLING2018_Barreiro_Batista_contractions-v2.pdfPós-print309,72 kBAdobe PDFVer/Abrir


FacebookTwitterDeliciousLinkedInDiggGoogle BookmarksMySpaceOrkut
Formato BibTex mendeley Endnote Logotipo do DeGóis Logotipo do Orcid 

Todos os registos no repositório estão protegidos por leis de copyright, com todos os direitos reservados.