ISO-Standard Domain-Independent Dialogue Act Tagging for Conversational Agents

Mezza, Stefano; Cervone, Alessandra; Tortoreto, Giuliano; Stepanov, Evgeny A.; Riccardi, Giuseppe

Dialogue Act (DA) tagging is crucial for spoken language understanding systems, as it provides a general representation of speakers’ intents, not bound to a particular dialogue system. Unfortunately, publicly available data sets with DA annotation are all based on different annotation schemes and thus incompatible with each other. Moreover, their schemes often do not cover all aspects necessary for open-domain human-machine interaction. In this paper, we propose a methodology to map several publicly available corpora to a subset of the ISO standard, in order to create a large task-independent training corpus for DA classification. We show the feasibility of using this corpus to train a domain-independent DA tagger testing it on out-of-domain conversational data, and argue the importance of training on multiple corpora to achieve robustness across different DA categories.

ISO-Standard Domain-Independent Dialogue Act Tagging for Conversational Agents / Mezza, S., Cervone, A., Tortoreto, G., Stepanov, E.A., Riccardi, G.. - ELETTRONICO. - (2018), pp. 3539-3551. (27th International Conference on Computational Linguistics, COLING 2018 Santa Fe 20th-26th August 2018).

ISO-Standard Domain-Independent Dialogue Act Tagging for Conversational Agents

Stefano Mezza;Alessandra Cervone;Giuliano Tortoreto;Evgeny A. Stepanov;Giuseppe Riccardi

2018-01-01

Abstract

Dialogue Act (DA) tagging is crucial for spoken language understanding systems, as it provides a general representation of speakers’ intents, not bound to a particular dialogue system. Unfortunately, publicly available data sets with DA annotation are all based on different annotation schemes and thus incompatible with each other. Moreover, their schemes often do not cover all aspects necessary for open-domain human-machine interaction. In this paper, we propose a methodology to map several publicly available corpora to a subset of the ISO standard, in order to create a large task-independent training corpus for DA classification. We show the feasibility of using this corpus to train a domain-independent DA tagger testing it on out-of-domain conversational data, and argue the importance of training on multiple corpora to achieve robustness across different DA categories.

Scheda breve

Scheda completa

Scheda completa (DC)

	Anno di pubblicazione (Date of publication)
	
				2018
			
	Titolo del volume (Proceedings title)
	
				Proceedings of the 27th International Conference on Computational Linguistics
			
	Luogo di edizione (Place of publication)
	
				Santa Fe
			
	Casa editrice (Publisher)
	
				Association for Computational Linguistics (ACL)
			
	ISBN
	
				978-1-948087-50-6
			
	Codice Scopus (Scopus Identifier)
	
				2-s2.0-85067137029
			
	Tutti gli autori
	
						Mezza, Stefano; Cervone, Alessandra; Tortoreto, Giuliano; Stepanov, Evgeny A.; Riccardi, Giuseppe
					
	Citazione
	
				ISO-Standard Domain-Independent Dialogue Act Tagging for Conversational Agents / Mezza, S., Cervone, A., Tortoreto, G., Stepanov, E.A., Riccardi, G.. - ELETTRONICO. - (2018), pp. 3539-3551. (27th International Conference on Computational Linguistics, COLING 2018 Santa Fe 20th-26th August 2018).
			
	Appare nelle tipologie:
	
				04.1 Saggio in atti di convegno (Paper in Proceedings)

File in questo prodotto:

File	Dimensione	Formato
C18-1300.pdf accesso aperto Tipologia: Versione editoriale (Publisher’s layout) Licenza: Creative commons Dimensione 210.41 kB Formato Adobe PDF Visualizza/Apri	210.41 kB	Adobe PDF	Visualizza/Apri