{"id":803,"date":"2022-10-28T12:10:05","date_gmt":"2022-10-28T12:10:05","guid":{"rendered":"http:\/\/iberspeech2022.ugr.es\/?page_id=803"},"modified":"2022-11-09T08:58:22","modified_gmt":"2022-11-09T08:58:22","slug":"technical-program-day-3","status":"publish","type":"page","link":"https:\/\/iberspeech2022.ugr.es\/?page_id=803","title":{"rendered":"Technical Program, Day 3"},"content":{"rendered":"<h3 style=\"background-color: #f3f3f3;\"><span style=\"color: #2e75b7;\"><strong>Wednesday, November 16<\/strong><\/span><\/h3>\n<p><strong>Oral 5: Natural Language Processing<\/strong><br \/>\n<span style=\"color: #d62013;\"><strong>Wednesday, 16\u00a0 November 2022 (9:00-10:40)<\/strong><\/span><br \/>\n<strong>Chair:\u00a0 Zoraida Callejas Carri\u00f3n<\/strong><\/p>\n<table width=\"0\">\n<tbody>\n<tr>\n<td rowspan=\"2\" width=\"140\"><strong>O5.1<\/strong><br \/>\n09:00\u00a0 &#8211; 09:20<\/td>\n<td>An Attentional Extractive Summarization Framework (<span class=\"collapseomatic \" id=\"id6a6eaab52ff88\"  tabindex=\"0\" title=\"abs\"    >abs<\/span><div id=\"target-id6a6eaab52ff88\" class=\"collapseomatic_content \">abstractive approaches, extractive methods can be specially adequate for some applications, and they can help with other tasks such as Question Answering or Information Extraction. In this paper, we propose a general framework for extractive summarization, the Attentional Extractive Summarization framework. The proposed approach is based on the interpretation of the attention mechanisms of hierarchical neural networks, that compute document-level representations of documents and summaries from sentence-level representations, which, in turn, are computed from word-level representations. The models proposed under this framework are able to automatically learn relationships among document and summary sentences, without requiring oracle systems to compute reference labels for each sentence before the training phase. We evaluate two different systems, formalized under the proposed framework, on the CNN\/DailyMail and the NewsRoom corpora, which are some of the reference corpora in the most relevant works in text summarization. The results obtained during the evaluation support<br \/>\nthe adequacy of our proposal and they suggest that there is still room for the improvement of our attentional framework.<\/div>)<\/td>\n<\/tr>\n<tr>\n<td><em>Jos\u00e9 \u00c1ngel Gonz\u00e1lez, Encarna Segarra, Fernando Garc\u00eda-Granada, Emilio Sanchis and Lluis-F Hurtado<\/em><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" width=\"140\"><strong>O5.2<\/strong><br \/>\n09:20\u00a0 &#8211; 09:40<\/td>\n<td>SUMBot: Summarizing Context in Open-Domain Dialogue Systems (<span class=\"collapseomatic \" id=\"id6a6eaab52fffb\"  tabindex=\"0\" title=\"abs\"    >abs<\/span><div id=\"target-id6a6eaab52fffb\" class=\"collapseomatic_content \"><span dir=\"ltr\" role=\"presentation\">In this paper, we investigate the problem of including relevant information as <\/span><span dir=\"ltr\" role=\"presentation\">context in open-domain dialogue systems.<\/span> <span dir=\"ltr\" role=\"presentation\">Most models struggle to identify and<\/span> <span dir=\"ltr\" role=\"presentation\">incorporate important knowledge from dialogues and simply use the entire turns as<\/span> <span dir=\"ltr\" role=\"presentation\">context, which increases the size of the input fed to the model with unnecessary<\/span> <span dir=\"ltr\" role=\"presentation\">information. Additionally, due to the input size limitation of a few hundred tokens <\/span><span dir=\"ltr\" role=\"presentation\">of large pre-trained models, regions of the history are not included and informative <\/span><span dir=\"ltr\" role=\"presentation\">parts from the dialogue may be omitted.<\/span> <span dir=\"ltr\" role=\"presentation\">In order to surpass this problem, we <\/span><span dir=\"ltr\" role=\"presentation\">introduce a simple method that substitutes part of the context with a summary <\/span><span dir=\"ltr\" role=\"presentation\">instead of the whole history, which increases the ability of models to keep track of <\/span><span dir=\"ltr\" role=\"presentation\">all the previous relevant information. We show that the inclusion of a summary may <\/span><span dir=\"ltr\" role=\"presentation\">improve the answer generation task and discuss some examples to further understand <\/span><span dir=\"ltr\" role=\"presentation\">the system\u2019s weaknesses.<\/span> <\/div>)<\/td>\n<\/tr>\n<tr>\n<td><em>Rui Ribeiro and Lu\u00edsa Coheur<\/em><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" width=\"140\"><strong>O5.3<\/strong><br \/>\n09:40\u00a0 &#8211; 10:00<\/td>\n<td>Automatic Detection of Inconsistencies in Open-Domain Chatbots (<span class=\"collapseomatic \" id=\"id6a6eaab53004c\"  tabindex=\"0\" title=\"abs\"    >abs<\/span><div id=\"target-id6a6eaab53004c\" class=\"collapseomatic_content \">Current pre-trained Large Language Models applied to chatbots are capable of producing good quality sentences, handling different conversation topics, and larger interaction times. Unfortunately, the generated responses highly depend on the data on which the chatbot has been trained on, the specific dialogue history and current turn used for guiding the response, the internal decoding mechanisms, ranking strategies, among others. Therefore, it may happen that for the same question asked by the user, the chatbot may provide a different answer, which in a long-term interaction may produce confusion. In this paper, we propose a new methodology based on three phases: a) automatic detection of dialogue topics using zeroshot learning approaches, b) automatic clustering of distinctive questions, and c) detecting inconsistent answers using K-Means clustering and the Silhouette coefficient. To test our proposal, we used the DailyDialog dataset to detect up to 13 different topics. To detect inconsistencies, we manually generated multiple paraphrased questions. Then, we used multiple pre-trained<br \/>\nchatbots to answer those questions. Our results in topic detection show a weighted F-1 value of 0.658, and a 3.4 MSE to predict the number of different responses.<\/div>)<\/td>\n<\/tr>\n<tr>\n<td><em>Jorge Mira Prats, Marcos Estecha-Garitagoitia, Mario Rodr\u00edguez-Cantelar and Luis Fernando D&#8217;Haro<\/em><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" width=\"140\"><strong>O5.4<\/strong><br \/>\n10:00\u00a0 &#8211; 10:20<\/td>\n<td>Ethics Guidelines for the Development of Virtual Assistants for e-Health (<span class=\"collapseomatic \" id=\"id6a6eaab530099\"  tabindex=\"0\" title=\"abs\"    >abs<\/span><div id=\"target-id6a6eaab530099\" class=\"collapseomatic_content \">The use of intelligent virtual assistants for human-machine communication is spreading across multiple applications. The latest breakthroughs in fields such as Natural Language Processing (NLP) or Natural Language Generation (NLG) make it possible to communicate with machines in a more natural and fluent way and in broader contexts, normalizing voice-based interactions with machines. These advances also lead to the appearance of new issues never seen before, especially when this technology extends to public services such as administration, education or health. The transfer of personal data, the opacity of decisions, the presence of bias or the exclusion of groups are critical aspects that cannot be controlled exclusively by economic interests. The design of conversational assistant solutions must be within an ethical, legal, socio-economic and cultural (ELSEC) framework, and it must be ensured that it preserves the dignity, freedom and autonomy of the users. In this paper, we analyse the Artificial Intelligence (AI) European regulatory framework, the issues that appear when designing and developing AI-based conversational solutions for e-health, and we present recommendations based on our experience and<br \/>\non the reflection from an ethical point of view.<\/div>)<\/td>\n<\/tr>\n<tr>\n<td><em>Andr\u00e9s Pi\u00f1eiro Mart\u00edn, Carmen Garc\u00eda Mateo, Laura Doc\u00edo Fern\u00e1ndez and Mar\u00eda del Carmen L\u00f3pez P\u00e9rez<\/em><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" width=\"140\"><strong>O5.5<\/strong><br \/>\n10:20\u00a0 &#8211; 10:40<\/td>\n<td>esCorpius: A Massive Spanish Crawling Corpus (<span class=\"collapseomatic \" id=\"id6a6eaab5300e8\"  tabindex=\"0\" title=\"abs\"    >abs<\/span><div id=\"target-id6a6eaab5300e8\" class=\"collapseomatic_content \"><span dir=\"ltr\" role=\"presentation\">In the recent years, transformer-based models have lead to significant advances <\/span><span dir=\"ltr\" role=\"presentation\">in language modelling for natural language processing. However, they require a vast <\/span><span dir=\"ltr\" role=\"presentation\">amount of data to be (pre-)trained and there is a lack of corpora in languages other <\/span><span dir=\"ltr\" role=\"presentation\">than English. Recently, several initiatives have presented multilingual datasets ob<\/span><span dir=\"ltr\" role=\"presentation\">tained from automatic web crawling. However, the results in Spanish present impor<\/span><span dir=\"ltr\" role=\"presentation\">tant shortcomings, as they are either too small in comparison with other languages,<\/span> <span dir=\"ltr\" role=\"presentation\">or present a low quality derived from sub-optimal cleaning and deduplication. In this<\/span> <span dir=\"ltr\" role=\"presentation\">paper, we introduce ESCORPIUS, a Spanish crawling corpus obtained from near 1 <\/span><span dir=\"ltr\" role=\"presentation\">PB of Common Crawl data.<\/span> <span dir=\"ltr\" role=\"presentation\">It is the most extensive corpus in Spanish with this<\/span> <span dir=\"ltr\" role=\"presentation\">level of quality in the extraction, purification and deduplication of web textual con<\/span><span dir=\"ltr\" role=\"presentation\">tent. Our data curation process involves a novel highly parallel cleaning pipeline and<\/span> <span dir=\"ltr\" role=\"presentation\">encompasses a series of deduplication mechanisms that together ensure the integrity <\/span><span dir=\"ltr\" role=\"presentation\">of both document and paragraph boundaries. Additionally, we maintain both the<\/span> <span dir=\"ltr\" role=\"presentation\">source web page URL and the WARC shard origin URL in order to complain with<\/span> <span dir=\"ltr\" role=\"presentation\">EU regulations. ESCORPIUS has been released under CC BY-NC-ND 4.0 license <\/span><span dir=\"ltr\" role=\"presentation\">and it is available on HuggingFace.<\/span><\/div>)<\/td>\n<\/tr>\n<tr>\n<td><em>Asier Guti\u00e9rrez-Fandi\u00f1o, David P\u00e9rez-Fern\u00e1ndez, Jordi Armengol-Estap\u00e9, David Griol and Zoraida Callejas<\/em><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p><strong>Keynote 3 <\/strong><br \/>\n<span style=\"color: #d62013;\"><strong>Wednesday, 16\u00a0 November 2022 (11:00-12:00)<\/strong><\/span><\/p>\n<table width=\"0\">\n<tbody>\n<tr>\n<td rowspan=\"2\" width=\"140\"><strong>KN3<\/strong><br \/>\n11:00\u00a0 &#8211; 12:00<\/td>\n<td>Utilizing context information in speech recognition for voice assistants (<span class=\"collapseomatic \" id=\"id6a6eaab530134\"  tabindex=\"0\" title=\"abs\"    >abs<\/span><div id=\"target-id6a6eaab530134\" class=\"collapseomatic_content \">Utilizing contextual information plays a key role in achieving accurate speech recognition in the voice assistant domain. Context can be available in a number of ways, such as the acoustic environment, conversational context, or personalized information about the user. Other sources of context information are trending content at the time the user is speaking and the speaker\u2019s location. While there are established methods for utilizing some of this information in traditional statistical speech recognition systems, contextualizing all-neural speech recognition systems is an active area of research. In my talk, I will present ongoing research at Amazon Alexa on this problem and discuss some of the challenges. <\/div>)<\/td>\n<\/tr>\n<tr>\n<td><em>Simon Wiesler<\/em><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p><strong>Entrepreneurship<\/strong><br \/>\n<span style=\"color: #d62013;\"><strong>Wednesday, 16\u00a0 November 2022 (12:00-13:00)<\/strong><\/span><br \/>\n<strong>Chair: Dayana Rivas\u00a0<\/strong><\/p>\n<table width=\"0\">\n<tbody>\n<tr>\n<td rowspan=\"2\" width=\"140\"><strong>EN<\/strong><br \/>\n12:00\u00a0 &#8211; 13:00<\/td>\n<td>Entrepreneurship Round Table (<span class=\"collapseomatic \" id=\"id6a6eaab53017e\"  tabindex=\"0\" title=\"abs\"    >abs<\/span><div id=\"target-id6a6eaab53017e\" class=\"collapseomatic_content \"> <\/div>)<\/td>\n<\/tr>\n<tr>\n<td><em>\u00a0<\/em><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<hr \/>\n","protected":false},"excerpt":{"rendered":"<p>Wednesday, November 16 Oral 5: Natural Language Processing Wednesday, 16\u00a0 November 2022 (9:00-10:40) Chair:\u00a0 Zoraida Callejas Carri\u00f3n O5.1 09:00\u00a0 &#8211; 09:20 An Attentional Extractive Summarization Framework () Jos\u00e9 \u00c1ngel Gonz\u00e1lez, Encarna Segarra, Fernando Garc\u00eda-Granada, Emilio Sanchis and Lluis-F Hurtado O5.2 09:20\u00a0 &#8211; 09:40 SUMBot: Summarizing Context in Open-Domain Dialogue Systems () Rui Ribeiro and Lu\u00edsa&hellip;&nbsp;<a href=\"https:\/\/iberspeech2022.ugr.es\/?page_id=803\" rel=\"bookmark\">Leer m\u00e1s &raquo;<span class=\"screen-reader-text\">Technical Program, Day 3<\/span><\/a><\/p>\n","protected":false},"author":1,"featured_media":0,"parent":0,"menu_order":0,"comment_status":"closed","ping_status":"closed","template":"","meta":{"neve_meta_sidebar":"","neve_meta_container":"","neve_meta_enable_content_width":"off","neve_meta_content_width":100,"neve_meta_title_alignment":"","neve_meta_author_avatar":"","neve_post_elements_order":"","neve_meta_disable_header":"","neve_meta_disable_footer":"","neve_meta_disable_title":"","footnotes":""},"class_list":["post-803","page","type-page","status-publish","hentry"],"_links":{"self":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages\/803","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages"}],"about":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/types\/page"}],"author":[{"embeddable":true,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=803"}],"version-history":[{"count":8,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages\/803\/revisions"}],"predecessor-version":[{"id":943,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages\/803\/revisions\/943"}],"wp:attachment":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=803"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}