遇见数据集

HAVIC MED Novel 1 Test -- Videos, Metadata and Annotation

收藏
DataCite Commons2022-03-17 更新2024-07-13 收录
官方服务:

资源简介:

<h3>Introduction</h3><br> <p>HAVIC MED Novel 1 Test -- Videos, Metadata and Annotation was developed by the Linguistic Data Consortium (LDC) and is comprised of approximately 3,800 hours of user-generated videos with annotation and metadata.</p><br> <p>To advance multimodal event detection and related technologies, LDC developed, in collaboration with <a href="https://www.nist.gov/">NIST</a> (the National Institute of Standards and Technology), a large, heterogeneous, annotated multimodal corpus for <a href="https://www.ldc.upenn.edu/collaborations/past-projects/havic">HAVIC</a> (the Heterogeneous Audio Visual Internet Collection) that was used in the NIST-sponsored <a href="https://www.nist.gov/itl/iad/mig/trecvid-multimedia-event-detection-evaluation-track">MED</a> (Multimedia Event Detection) task for several years. HAVIC MED Novel 1 Test is a subset of that corpus, specifically, a collection of videos, metadata and annotation for the HAVIC project originally released to support the <a href="https://www.nist.gov/itl/iad/mig/med-2015-evaluation">2015</a> Multimedia Event Detection tasks.</p><br> <h3>Data</h3><br> <p>The data consists of videos of various events (event videos) and videos completely unrelated to events (background videos) harvested by a large team of human annotators. Each event video was manually annotated with a set of judgments describing its event properties and other salient features. Background videos were labeled with topic and genre categories.</p><br> <p>All video files are in .mp4 format (h.264), with varying bit-rates and levels of audio fidelity and video resolution. Metadata and annotation for the videos are stored in a .tsv file.</p><br> <h3>Samples</h3><br> <p>Please view this <a href="desc/addenda/LDC2022V01.mp4">video sample</a>.</p><br> <h3>Updates</h3><br> <p>None at this time.</p><br> <h3>Additional Licensing Instructions</h3><br> <p>This 'members-only' corpus is available to current members. Contact <a href="mailto:ldc@ldc.upenn.edu">ldc@ldc.upenn.edu</a>&nbsp;for information about becoming a member.</p></br> Portions © 2011-2016 YouTube, LLC, © 2011-2016, 2022 Trustees of the University of Pennsylvania

<h3>引言</h3><br><p>HAVIC MED Novel 1测试集——视频、元数据与标注(HAVIC MED Novel 1 Test -- Videos, Metadata and Annotation)由语言数据联盟(Linguistic Data Consortium, LDC)开发,包含约3800小时的用户生成视频,且均附带标注与元数据。</p><br><p>为推动多模态事件检测及相关技术发展,LDC与<a href="https://www.nist.gov/">美国国家标准与技术研究院(National Institute of Standards and Technology, NIST)</a>合作,为<a href="https://www.ldc.upenn.edu/collaborations/past-projects/havic">异构音视频互联网集合(Heterogeneous Audio Visual Internet Collection, HAVIC)</a>构建了大型异构带标注多模态语料库。该语料库曾多年应用于NIST主办的<a href="https://www.nist.gov/itl/iad/mig/trecvid-multimedia-event-detection-evaluation-track">多媒体事件检测(Multimedia Event Detection, MED)</a>评测任务。HAVIC MED Novel 1测试集正是该语料库的子集,特指为支持<a href="https://www.nist.gov/itl/iad/mig/med-2015-evaluation">2015年</a>多媒体事件检测任务而首次发布的HAVIC项目相关视频、元数据与标注集合。</p><br><h3>数据概况</h3><br><p>本数据集包含由大量人工标注人员采集的各类事件视频(事件视频)及与事件完全无关的背景视频(背景视频)。每段事件视频均经人工标注,包含描述其事件属性与其他显著特征的多组判定结果;背景视频则标注了主题与体裁类别。</p><br><p>所有视频文件均采用.mp4格式(h.264编码),码率、音频保真度与视频分辨率各不相同。视频对应的元数据与标注存储于.tsv格式文件中。</p><br><h3>样例</h3><br><p>请查看此<a href="desc/addenda/LDC2022V01.mp4">视频样例</a>。</p><br><h3>更新记录</h3><br><p>暂无更新。</p><br><h3>额外授权说明</h3><br><p>本专属会员语料库仅对现有会员开放。如需了解会员注册相关信息,请联系<a href="mailto:ldc@ldc.upenn.edu">ldc@ldc.upenn.edu</a>&nbsp;。</p></br><p>部分内容©2011-2016 YouTube有限责任公司,©2011-2016、2022 宾夕法尼亚大学理事会</p>

创建时间:
2022-03-10
搜集汇总
数据集介绍
HAVIC MED Novel 1 Test -- Videos, Metadata and Annotation 数据集图片
背景与挑战
背景概述
该数据集是HAVIC MED Novel 1 Test,包含约3,800小时用户生成的视频,附带注释和元数据,用于支持多媒体事件检测任务。数据由事件视频和背景视频组成,视频格式为.mp4,元数据和注释以.tsv文件存储,主要应用于英语环境的事件检测技术开发。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务