(2003) 4th International Conference on Audio- and Video-Based Biometric Person Authentication — Location: UNIV SURREY, GUILDFORD (England) (9.June.2003)
In this work, we present a multimodal identity verification system based on the fusion of the face image and the text independent speech data of a person. The system conciliates the monomodal face and speaker verification algorithms by fusing their respective scores. In order to assess the authentication system at different scales, the performance is evaluated at various sizes of the face and speech user template. The user template size is a key parameter when the storage space is limited like in a smart card. Our experimental results show that the multimodal fusion allows to reduce significantly the user template size while keeping a satisfactory level of performance. Experiments are performed on the newly recorded multimodal database BANCA.
Czyz, J., Bengio, S., Marcel, C., & Vandendorpe, L. (2003). Scalability analysis of audio-visual person identity verification. 4th Int. Conf. Audio and Video Based Biometric Person Authentication, 752-760. https://doi.org/10.1007/3-540-44887-X_87