machine learning, machine vision, feature selection, texture classification, image analysis, texture analysis, local binary pattern, deep learning
A. Visual information Inference.
[1] Amanda Cardoso Duarte, Francisco Roldan, Miquel Tubau, Janna Escur, Santiago Pascual, Amaia Salvador, Eva Mohedano, Kevin McGuinness, Jordi Torres, and Xavier Giró-i-Nieto, “Wav2pix : Speech-conditioned face generation using generative adversarial networks,” in IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2019, Brighton, United Kingdom, May 12-17, 2019. 2019, pp. 8633–8637, IEEE.
[2] Zheng Fang, Zhen Liu, Tingting Liu, Chih-Chieh Hung, Jiangjian Xiao, and Guangjin Feng, “Facial expression GAN for voice-driven face generation,” Vis. Comput., vol. 38, no. 3, pp. 1151–1164, 2022.
[3] Tae-Hyun Oh, Tali Dekel, Changil Kim, Inbar Mosseri, William T. Freeman, Michael Rubinstein, and Wojciech Matusik, “Speech2face : Learning the face behind a voice,” in IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2019, Long Beach, CA, USA, June 16-20, 2019. 2019, pp. 7539–7548, Computer Vision Foundation / IEEE.
[4] Yandong Wen, Bhiksha Raj, and Rita Singh, “Face reconstruction from voice using generative adversarial networks,” in Advances in Neural Information Processing Systems 32 : Annual Conference on Neural Information Processing Systems 2019, NeurIPS 2019, December 8-14, 2019, Vancouver, BC, Canada, Hanna M. Wallach, Hugo Larochelle, Alina Beygelzimer, Florence d’Alché-Buc, Emily B. Fox, and Roman Garnett, Eds., 2019, pp. 5266–5275.