MAVS (Multilingual Audio-Visual Smartphone dataset)
Introduced by Mandalapu et al. inMultilingual Audio-Visual Smartphone Dataset And Evaluation
MAVSis an audio-visual smartphone dataset captured in five different recent smartphones. This new dataset contains 103 subjects captured in three different sessions considering the different real-world scenarios. Three different languages are acquired in this dataset to include the problem of language dependency of the speaker recognition systems.
