CONFIG-body_parts-nose_eyes_face_3form-all_01_02_TS.wav
收藏官方服务:
资源简介:
:unav
应用场景:
创建时间:
2023-06-28
相关数据集
MAV-Celeb
MAV-Celeb是由林茨大学等机构联合构建的多模态说话人识别数据集,包含YouTube访谈、脱口秀等场景下的双语(英语-乌尔都语)音频-视觉样本。数据集包含4039个英语训练样本和9304个乌尔都语训练样本,每个样本均包含人脸图像(.jpg)和语音片段(.wav)的成对数据。数据通过预训练模型提取特征,并采用分层结构按模态、身份和语言组织。该数据集旨在解决多模态说话人识别中的模态缺失和跨语言泛化
arXiv2026-03-26 更新660
MAVS数据集
MAVS数据集是由挪威科技大学和印度理工学院合作开发的多语言视听智能手机数据集,旨在支持开发和评估先进的单模态或视听说话人识别系统。该数据集包含103名参与者(70名男性和33名女性),在三个不同会话中录制,考虑了不同的现实世界场景,并使用了五种不同的智能手机设备。数据集中的每个参与者用三种语言(英语、印地语和孟加拉语)说出六句话,每种语言三句。此外,数据集还包括两种类型的演示攻击,即物理访问攻击
arXiv2021-11-15 更新200
OMuSense-23++: A Multimodal dataset for contactless breathing pattern recognition and biometric analysis
OMuSense-23 is a multimodal dataset for non-contact biometric and breathing analysis. This database comprises RGBD and mmWave radar data collected from 50 participants.
Zenodo2025-06-06 更新70
LUTBIO multimodal biometric database
The LUTBIO database provides a comprehensive resource for research in multimodal biometric authentication, featuring the following key aspects: - Extensive Biometric Modalities: The database contains
doi.org2024-11-26 更新490
Performance of parallel multimodal system.
In the field of data security, biometric security is a significant emerging concern. The multimodal biometrics system with enhanced accuracy and detection rate for smart environments is still a signif
NIAID Data Ecosystem70



