In plain words: A new collection of Italian speech examples, partly built automatically, marks what the speaker wants and the key details in each request. It is the first such Italian set, and it was used to compare open-source and commercial voice systems where none existed before.
Abstract · Almawave-SLU: A new dataset for SLU in Italian
The widespread use of conversational and question answering systems made it necessary to improve the performances of speaker intent detection and understanding of related semantic slots, i.e., Spoken Language Understanding (SLU). Often, these tasks are approached with supervised learning methods, which needs considerable labeled datasets. This paper presents the first Italian dataset for SLU. It is derived through a semi-automatic procedure and is used as a benchmark of various open source and commercial systems.
Valentina Bellomaria, Giuseppe Castellucci, Andrea Favalli, Raniero Romagnoli
arXiv:1907.07526 · cs.CL · submitted Jul 17, 2019
abstract · pdf · html