We introduce the "LipBengal" dataset, marking a significant advancement in the field of Bengali lip-reading and visual speech recognition research. This dataset addresses a critical gap in the researc
The Lip Reading Vowel-Bangla (LRV-B) Dataset is a curated collection of video recordings focused on the articulation of Bangla vowels. It is designed to support research in lip reading and visual spee
The MODALITY corpus is one of the multimodal database of word recordings in English. It consists of over 30 hours of multimodal recordings. The database contains high-resolution, high-framerate stereo