Deep dialog act recognition using multiple token, segment, and context information representations

dc.contributor.authorRibeiro, E.
dc.contributor.authorRibeiro, R.
dc.contributor.authorMatos, D. M. de.
dc.date.accessioned2020-03-20T15:41:52Z
dc.date.available2020-03-20T15:41:52Z
dc.date.issued2019
dc.date.updated2024-11-07T09:45:24Z
dc.description.abstractAutomatic dialog act recognition is a task that has been widely explored over the years. In recent works, most approaches to the task explored different deep neural network architectures to combine the representations of the words in a segment and generate a segment representation that provides cues for intention. In this study, we explore means to generate more informative segment representations, not only by exploring different network architectures, but also by considering different token representations, not only at the word level, but also at the character and functional levels. At the word level, in addition to the commonly used uncontextualized embeddings, we explore the use of contextualized representations, which are able to provide information concerning word sense and segment structure. Character-level tokenization is important to capture intention-related morphological aspects that cannot be captured at the word level. Finally, the functional level provides an abstraction from words, which shifts the focus to the structure of the segment. Additionally, we explore approaches to enrich the segment representation with context information from the history of the dialog, both in terms of the classifications of the surrounding segments and the turn-taking history. This kind of information has already been proved important for the disambiguation of dialog acts in previous studies. Nevertheless, we are able to capture additional information by considering a summary of the dialog history and a wider turn-taking context. By combining the best approaches at each step, we achieve performance results that surpass the previous state-of-the-art on generic dialog act recognition on both the Switchboard Dialog Act Corpus (SwDA) and the ICSI Meeting Recorder Dialog Act Corpus (MRDA), which are two of the most widely explored corpora for the task. Furthermore, by considering both past and future context, similarly to what happens in an annotation scenario, our approach achieves a performance similar to that of a human annotator on SwDA and surpasses it on MRDA.eng
dc.description.versioninfo:eu-repo/semantics/publishedVersion
dc.identifier.citationRibeiro, E., Ribeiro, R., & Matos, D. M. de. (2019). Deep dialog act recognition using multiple token, segment, and context information representations. Journal of Artificial Intelligence Research, 66, 861-899. https://doi.org/10.1613/jair.1.11594
dc.identifier.doi10.1613/jair.1.11594
dc.identifier.issn1076-9757
dc.identifier.urihttp://hdl.handle.net/10071/20151
dc.journalJournal of Artificial Intelligence Research
dc.language.isoeng
dc.pagination861 - 899
dc.peerreviewedyes
dc.publisherAI Access Foundation
dc.relationUID/CEC/50021/2019
dc.rightsopen access
dc.subjectDialog processingeng
dc.subjectNatural languageeng
dc.subjectNeural networkseng
dc.subject.fosDomínio/Área Científica::Ciências Naturais::Ciências da Computação e da Informaçãopor
dc.subject.fosDomínio/Área Científica::Engenharia e Tecnologia::Engenharia Eletrotécnica, Eletrónica e Informáticapor
dc.subject.fosDomínio/Área Científica::Humanidades::Línguas e Literaturaspor
dc.titleDeep dialog act recognition using multiple token, segment, and context information representationseng
dc.typearticle
dc.volume66
degois.publication.firstPage861
degois.publication.lastPage899
degois.publication.titleDeep dialog act recognition using multiple token, segment, and context information representationseng
dspace.entity.typePublicationen
iscte.alternateIdentifiers.scopus2-s2.0-85077112723
iscte.alternateIdentifiers.wosWOS:WOS:000507355600013
iscte.identifier.cienciahttps://ciencia.iscte-iul.pt/id/ci-pub-64052
iscte.journalJournal of Artificial Intelligence Research

Ficheiros

Pacote original

A mostrar 1 - 1 de 1
A carregar...
Nome:
article_64052.pdf
Tamanho:
547.83 KB
Formato:
Adobe Portable Document Format
Descrição:
Versão Editora