Enhancing Voice Communication Services with Audio Fingerprinting Techniques
Kemal Altwlkany, University of Sarajevo
Toward a Holistic Modeling Paradigm for Accurate, Versatile, and Controllable Latent Audio Diffusion
Yuxuan Jiang, Tsinghua University
Community-Oriented Development of Speech Emotion Recognition Framework for te reo Māori
Himashi Rathnayake, University of Auckland
Detection and Perception of Smiled Speech
Rong Li, University of Twente
Computational Measures of Speech and Language for Schizophrenia Spectrum Disorders
Shrankhla Pandey, University of Cambridge
Physics-Based Simulation of Vowel Utterances using a Biomechanical-Acoustic Model
Debasish Mohapatra, University of British Columbia
Characterization and measurement of listenability for vulnerable individuals
Baptiste Ramonda, Institut de Recherche en Informatique de Toulouse (IRIT)
Natural Language as the Interface to Speaking Style: Describing and Controlling How Speech Is Delivered
Shreeram Suresh Chandra, The University of Texas at Dallas
Towards Agile and Generalisable Audio-Visual Speech Recognition
Zhaofeng Lin, Trinity College Dublin
Privacy Considerations for Audio-Visual Speech Enhancement
Poppy Welch, University of Southampton
Evaluating Phonological Knowledge in Foundational Speech Models
Sneha Ray Barman, Indian Institute of Technology Guwahati
From Representation to Retention: Locating and Protecting Knowledge for Continual Speech Adaptation
Yang Xiao, The University of Melbourne
Towards Trustworthy Speech-Based Multimodal Depression Detection
Dushanthi Manamalage, The University of Auckland