{"id":77,"date":"2022-05-05T14:28:08","date_gmt":"2022-05-05T14:28:08","guid":{"rendered":"http:\/\/iberspeech2022.ugr.es\/wordpress\/?page_id=77"},"modified":"2022-10-17T12:01:40","modified_gmt":"2022-10-17T12:01:40","slug":"keynote-speakers","status":"publish","type":"page","link":"https:\/\/iberspeech2022.ugr.es\/?page_id=77","title":{"rendered":"Keynote Speakers"},"content":{"rendered":"<h3>Keynote Speakers at IberSPEECH 2022 are the following:<\/h3>\n<h3><u>Day: <strong>November 14th<\/strong><\/u><\/h3>\n<h3>Massimiliano Todisco, EURECOM Digital Security\u00a0Department, France.<\/h3>\n<p>Title of the talk:\u00a0 <strong>Secure and explainable voice biometrics<br \/>\n<\/strong><\/p>\n<p><em>Abstract<\/em>: Anti-spoofing for voice biometrics is now an established area of research, thanks to the four competitive ASVspoof challenges (the fifth is currently underway) that have taken place over the past decade. Growing research effort has invested, firstly, in the development of front-end representations that capture more reliably the tell-tale artefacts that are indicative of utterances generated with text-to-speech and voice conversion algorithms and, secondly, in the development of deep and end-to-end solutions. Despite enormous efforts and positive achievements, little is still known about the artefacts these recognisers use to identify spoofing utterances or distinguish between bona fide and spoofed. Although many unanswered questions remain, this talk aims to provide insights and inspirations, through examples, into the behaviour of voice anti-spoofing systems. Particular attention will be given to data augmentation and boosting methods that have been shown instrumental to reliability. The ultimate goal is to better understand these artefacts from a physical and perceptual point of view and how they are actually seen by automatic processes, which puts us in a better position to design more reliable countermeasures.<\/p>\n<p><strong><img loading=\"lazy\" decoding=\"async\" class=\"alignleft size-medium wp-image-288\" src=\"http:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/max-300x300.png\" alt=\"\" width=\"238\" height=\"238\" srcset=\"https:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/max-300x300.png 300w, https:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/max-150x150.png 150w, https:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/max.png 640w\" sizes=\"auto, (max-width: 238px) 100vw, 238px\" \/>Bio<\/strong>: Massimiliano Todisco is a professor of audio and speech technologies at the EURECOM Digital Security\u00a0Department in France. He received his PhD in Sensorial and Learning Systems Engineering from the University\u00a0of Rome Tor Vergata in 2012. From 2012 to 2015 he was a postdoctoral researcher at Fondazione Ugo Bordoni\u00a0and Tor Vergata University in Rome. From 2015 to 2020, he was a senior research fellow at EURECOM.\u00a0Massimiliano\u00a0is best known for contributions to fake audio detection. He is the inventor of the <em>constant Q\u00a0cepstral coefficients<\/em> (CQCCs), the most used features and source of inspiration for many researchers in speech\u00a0processing, speaker recognition and anti-spoofing field. For this work, he was honoured with the ISCA 2020\u00a0award for the best article published in the journal \u00abComputer Speech and Language\u00bb during the quinquennium\u00a02015-2019. He co-organises the ASVspoof challenge series, which is community-led challenges that promote\u00a0the development of countermeasures to protect automatic speaker verification (ASV) from the threat of\u00a0spoofing.\u00a0He is currently principal investigator and coordinator of TReSPAsS-ETN, an EU Marie Sk\u0142odowska-Curie Innovative\u00a0Training Network (ITN) project, and RESPECT,\u00a0a project funded by the national research agencies of France and Germany.\u00a0His current interests are in developing explainable DNN architectures for speech processing and speaker\u00a0recognition, fake audio detection and anti-spoofing, and the development of privacy preservation algorithms for\u00a0speech signals based on encryption solutions that support computation upon signals, templates and models in\u00a0the encrypted domain.<\/p>\n<h3><u>Day: <strong>November 15th<\/strong><\/u><\/h3>\n<h3>Isabel Trancoso, INESC-ID \/ IST \/ University of Lisbon, Portugal<\/h3>\n<p>Title of the talk: <strong>Disease biomarkers in speech<br \/>\n<\/strong><\/p>\n<p><em>Abstract<\/em>: Speech encodes information about a plethora of diseases, which go beyond the so-called speech and language disorders, and include neurodegenerative diseases, such as Parkinson\u2019s, Alzheimer\u2019s, and Huntington\u2019s disease, mood and anxiety-related diseases, such as Depression and Bipolar Disease, and diseases that concern respiratory organs such as the common Cold, or Obstructive Sleep Apnea. This talk addresses the potential of speech as a health biomarker which allows a non-invasive route to early diagnosis and monitoring of a range of conditions related to human physiology and cognition. The talk will also address the many challenges that lie ahead, namely in the context of an ageing population with frequent multimorbidity, and the need to build robust models that provide explanations compatible with clinical reasoning. That would be a major step towards a future where collecting speech samples for health screening may become as common as a blood test nowadays. Speech can indeed encode health information au par with many other characteristics that make it viewed as Personal Identifiable Information. The last part of this talk will briefly discuss the privacy issues that this enormous potential may entail.<\/p>\n<p style=\"text-align: left;\"><img loading=\"lazy\" decoding=\"async\" class=\"alignleft size-full wp-image-287\" src=\"http:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/ITrancoso.png\" alt=\"\" width=\"238\" height=\"238\" srcset=\"https:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/ITrancoso.png 238w, https:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/ITrancoso-150x150.png 150w\" sizes=\"auto, (max-width: 238px) 100vw, 238px\" \/><\/p>\n<p style=\"text-align: left;\"><strong>Bio<\/strong>: Isabel Trancoso is a full professor at Instituto Superior T\u00e9cnico (IST, Univ. Lisbon), the University where she got her PhD degree in 1987. She was\u00a0 the founder of the Human Language Technology Lab and the former President of the Scientific Council of INESC ID Lisbon. She chaired the ECE Department of IST, was Editor-in-Chief of the IEEE Transactions on Speech and Audio Processing and had many leadership roles in SPS (Signal Processing Society of IEEE) and ISCA (International Speech Communication Association), namely having been President of ISCA and Chair of the Fellow Evaluation Committees of both SPS and ISCA. She was elevated to IEEE Fellow in 2011, and to ISCA Fellow in 2014.<\/p>\n<p>&nbsp;<\/p>\n<h3><u>Day: <strong>November 16th<\/strong><\/u><\/h3>\n<h3>Simon Wiesler, Applied Science Manager at Amazon, Germany<\/h3>\n<p>Title of the talk: <strong>Utilizing context information in speech recognition for voice assistants<\/strong><\/p>\n<p><em>Abstract<\/em>: Utilizing contextual information plays a key role in achieving accurate speech recognition in the voice assistant domain. Context can be available in a number of ways, such as the acoustic environment, conversational context, or personalized information about the user. Other sources of context information are trending content at the time the user is speaking and the speaker\u2019s location. While there are established methods for utilizing some of this information in traditional statistical speech recognition systems, contextualizing all-neural speech recognition systems is an active area of research. In my talk, I will present ongoing research at Amazon Alexa on this problem and discuss some of the challenges.<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignleft size-medium wp-image-291\" src=\"http:\/\/iberspeech2022.ugr.es\/wp-content\/uploads\/2022\/07\/simon_wiesler2-1-259x300.jpg\" alt=\"\" width=\"238\" height=\"300\" \/><\/p>\n<p><strong>Bio<\/strong>:\u00a0 Simon Wiesler is a Science Manager in Alexa ASR at Amazon. He received his PhD from RWTH Aachen University, Germany, in 2016. Prior to his PhD, he completed a Diploma degree in Mathematics from the University of Marburg, Germany. At Amazon, he manages a team of scientists, which develops current and future technology for the Alexa cloud speech recognition system. His current research interests include machine learning for far-field speech recognition and use of context information in speech recognition systems.<\/p>\n<p>&nbsp;<\/p>\n<p>&nbsp;<\/p>\n<h3><\/h3>\n<h3><\/h3>\n<h3><\/h3>\n<p>&nbsp;<\/p>\n<hr \/>\n","protected":false},"excerpt":{"rendered":"<p>Keynote Speakers at IberSPEECH 2022 are the following: Day: November 14th Massimiliano Todisco, EURECOM Digital Security\u00a0Department, France. Title of the talk:\u00a0 Secure and explainable voice biometrics Abstract: Anti-spoofing for voice biometrics is now an established area of research, thanks to the four competitive ASVspoof challenges (the fifth is currently underway) that have taken place over&hellip;&nbsp;<a href=\"https:\/\/iberspeech2022.ugr.es\/?page_id=77\" rel=\"bookmark\">Leer m\u00e1s &raquo;<span class=\"screen-reader-text\">Keynote Speakers<\/span><\/a><\/p>\n","protected":false},"author":1,"featured_media":0,"parent":0,"menu_order":0,"comment_status":"closed","ping_status":"closed","template":"","meta":{"neve_meta_sidebar":"","neve_meta_container":"","neve_meta_enable_content_width":"off","neve_meta_content_width":100,"neve_meta_title_alignment":"","neve_meta_author_avatar":"","neve_post_elements_order":"","neve_meta_disable_header":"","neve_meta_disable_footer":"","neve_meta_disable_title":"","footnotes":""},"class_list":["post-77","page","type-page","status-publish","hentry"],"_links":{"self":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages\/77","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages"}],"about":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/types\/page"}],"author":[{"embeddable":true,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=77"}],"version-history":[{"count":28,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages\/77\/revisions"}],"predecessor-version":[{"id":487,"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=\/wp\/v2\/pages\/77\/revisions\/487"}],"wp:attachment":[{"href":"https:\/\/iberspeech2022.ugr.es\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=77"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}