遇见数据集

HAVIC MED Novel 1 Test -- Videos, Metadata and Annotation

收藏
Mendeley Data2024-01-31 更新2024-06-28 收录
官方服务:

资源简介:

Introduction HAVIC MED Novel 1 Test -- Videos, Metadata and Annotation was developed by the Linguistic Data Consortium (LDC) and is comprised of approximately 3,800 hours of user-generated videos with annotation and metadata. To advance multimodal event detection and related technologies, LDC developed, in collaboration with NIST (the National Institute of Standards and Technology), a large, heterogeneous, annotated multimodal corpus for HAVIC (the Heterogeneous Audio Visual Internet Collection) that was used in the NIST-sponsored MED (Multimedia Event Detection) task for several years. HAVIC MED Novel 1 Test is a subset of that corpus, specifically, a collection of videos, metadata and annotation for the HAVIC project originally released to support the 2015 Multimedia Event Detection tasks. Data The data consists of videos of various events (event videos) and videos completely unrelated to events (background videos) harvested by a large team of human annotators. Each event video was manually annotated with a set of judgments describing its event properties and other salient features. Background videos were labeled with topic and genre categories. All video files are in .mp4 format (h.264), with varying bit-rates and levels of audio fidelity and video resolution. Metadata and annotation for the videos are stored in a .tsv file. Samples Please view this video sample. Updates None at this time. Additional Licensing Instructions This 'members-only' corpus is available to current members. Contact ldc@ldc.upenn.edu for information about becoming a member. Portions © 2011-2016 YouTube, LLC, © 2011-2016, 2022 Trustees of the University of Pennsylvania

引言:HAVIC MED 新颖测试集1——视频、元数据与标注集由语言数据联盟(Linguistic Data Consortium, LDC)开发,包含约3800小时带标注及元数据的用户生成视频。为推动多模态事件检测及相关技术发展,LDC与美国国家标准与技术研究院(National Institute of Standards and Technology, NIST)合作开发了面向HAVIC(异构视听互联网资源集,Heterogeneous Audio Visual Internet Collection, HAVIC)的大型异构带标注多模态语料库,该语料库已在NIST资助的MED(多媒体事件检测,Multimedia Event Detection, MED)任务中使用多年。HAVIC MED 新颖测试集1是该语料库的子集,特指为支持2015年多媒体事件检测任务而发布的HAVIC项目相关视频、元数据及标注集。 数据说明:本数据集包含由大规模人工标注团队采集的各类事件视频(事件视频)与完全无关事件的背景视频(背景视频)。每段事件视频均经人工标注,附带描述其事件属性及其他显著特征的一系列判定结果;背景视频则被标注了主题与体裁类别。所有视频文件均采用.mp4格式(h.264编码),比特率、音频保真度与视频分辨率各不相同。视频的元数据与标注信息存储于.tsv文件中。 示例:请查看本视频示例。 更新情况:暂无更新。 额外授权说明:本"仅限会员"语料库仅对现有会员开放。如需了解会员注册相关信息,请联系ldc@ldc.upenn.edu。 本数据集部分内容 © 2011-2016 YouTube有限责任公司,© 2011-2016、2022 宾夕法尼亚大学校董会。

创建时间:
2024-01-31
搜集汇总
数据集介绍
HAVIC MED Novel 1 Test -- Videos, Metadata and Annotation 数据集图片
背景与挑战
背景概述
该数据集是HAVIC MED项目的一个子集,包含约3,800小时的用户生成视频,用于多媒体事件检测任务。视频分为事件视频和背景视频,事件视频有手动标注的事件属性,背景视频则标注了主题和类型。数据以.mp4视频文件和.tsv元数据文件提供,支持英语事件检测研究。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务