Keynote talks will provide great insights on the current trends in Real-Time Multimodal Conversational AI from renowned and impactful scientists within the community. The workshop will feature four keynote talks.
Title to be announced
Abstract: TBA
Biography: Alexandre is a co-founder of Kyutai, a non profit lab for research in artificial intelligence based in Paris committed to open science. His work covers generative speech and multimodal AI (Moshi, Hibiki, DSM) with a strong focus on handling multiple streams jointly across modalities in a streaming and low latency fashion. He is also a co-founder and Chief Science Officer at Gradium, a startup launched in 2025 whose mission is to commercialize the best possible voice AI experience. Before that, Alexandre was a scientist for 3 years at Facebook AI Research in Paris, where he led the development of models for audio compression and modeling (AudioCraft, MusicGen, EnCodec).
Title to be announced
Abstract: TBA
Biography: Catherine Pelachaud (CNRS-ISIR) is a CNRS Director of Research in the laboratory ISIR, Sorbonne University. Her research interests include socially interactive agents, nonverbal communication (face, gaze, gesture, and touch), and adaptive mechanisms in interaction. She has been developing an interactive virtual agent platform, Greta, with her research team, that can display emotional and communicative behaviors. She has participated in organizing international conferences such as ICMI, IVA, ACII, FG and AAMAS. She is and was associate editor of several journals, including ACM Transactions on Interactive Intelligent Systems, International Journal of Human-Computer Studies, and IEEE Transactions on Affective Computing. She is co-editor of the ACM handbook on socially interactive agents (2021-22).
Title to be announced
Abstract: TBA
Biography: TBA