Titouan is a Senior Research Scientist at the Samsung AI Research Center in Cambridge (UK) and an Affiliated Lecturer at the Cambridge Machine Learning Systems Lab from the University of Cambridge (UK). Previously, he was an Associate Professor in computer science at the Laboratoire Informatique d’Avignon (LIA), from Avignon University (FR). He also was a senior research associate at the University of Oxford (UK) within the Oxford Machine Learning Systems group. He received his PhD in computer science from the University of Avignon (FR) and in partnership with Orkis focusing on quaternion neural networks, automatic speech recognition, and representation learning. His current work involves efficient and large scale speech technologies. He is also collaborating with the University of Montréal (Mila, QC, Canada) as the co-leader of the SpeechBrain project.
Shalini is a Director of Research, New Experiences and a Distinguished Research Scientist at NVIDIA, where she leads the AI-Mediated Reality and Interaction Research Group. Previously, she was a researcher in the Learning and Perception Research Group at NVIDIA, from 2013-2023. Her research interests are in AI, computer vision and computer graphics, focusing on AI-mediated interactions between humans and machines. She has co-authored scores of peer-reviewed publications and patents. Her inventions have contributed to several NVIDIA AI products, including DriveIX, Maxine and TAO Toolkit. She has prevously also organized 15 tutorials and workshops at ECCV, ICCV and CVPR and regularly serves as an area chair at all major AI conferences.
Rogieris a Principal Research Scientist working for Samsung at the AI Center-Cambridge, UK. His main topic of research is speech processing, and he has side interest in federated learning and differential privacy. Previously, he worked at Apple, first on Siri and later on private federated learning. Before that, he was at the University of Cambridge, Department of Engineering, working on speech recognition and on assessing the English of non-native speakers.
Naomi is Professor in Speech Technology at the School of Engineering in Trinity College Dublin, Ireland. She is Co-PI and a founding member of the Research Ireland ADAPT Centre. She is also a lead academic of the Sigmedia Research Group in the School of Engineering. She is known for her innovative research in audio-visual speech recognition, multimodal speech-based interaction, and turn-taking in conversations. She currently serves on the Editorial Board of Computer Speech and Language, and was General Chair of INTERSPEECH 2023 in Dublin.
Slim Essid is an Applied Research Manager at NVIDIA, which he joined in June 2025. Previously, he was Full Professor of Télécom Paris and the coordinator of the Audio Data Analysis and Signal Processing (ADASP) group. His research interests are broadly in multimodal AI, especially Audio/Vision and Omni Language Models.
Over the past 20 years, he has been involved in various French and European collaborative research projects. He has collaborated with 18 post-docs and research engineers and has graduated 22 PhD students; he is currently co-advising 5 others. He has previously organized a number of special sessions and workshops and was general co-chair of ISMIR 2018 in Paris.