遇见数据集

TRECVID 2005 Keyframes & Transcripts

收藏
DataCite Commons2021-07-01 更新2025-04-16 收录
官方服务:

资源简介:

<h3>Introduction</h3> <p>This file contains documentation for TRECVID 2005 Keyframes &amp; Transcripts, Linguistic Data Consortium (LDC) catalog number LDC2007V01 and isbn 1-58563-437-9. </p><p> TREC Video Retrieval Evaluation (TRECVID) is sponsored by the National Institute of Standards and Technology (NIST) to promote progress in content-based retrieval from digital video via open, metrics-based evaluation. The keyframes in this release were extracted for use in the NIST TRECVID 2005 Evaluation. </p><p>TRECVID is a laboratory-style evaluation that attempts to model real world situations or significant component tasks involved in such situations. In 2005 there were four main tasks with associated tests: </p><ul> <li> shot boundary determination </li> <li> low-level feature extraction </li> <li> high-level feature extraction </li> <li> search (interactive, manual, and automatic) </li> </ul><p>For a detailed description of the TRECVID Evaluation Tasks, please refer to <a href="http://www-nlpir.nist.gov/projects/tv2005/" rel="nofollow"> the NIST TRECVID 2005 Evaluation Description.</a> </p><h3>Data</h3> <p>The source data is Arabic, Chinese and English language broadcast programming collected in November 2004 from the following sources: Lebanese Broadcasting Corp. (Arabic); China Central TV and New Tang Dynasty TV (Chinese); and CNN and MSNBC/NBC (English). </p><p>Shots are fundamental units of video, useful for higher-level processing. To create the master list of shots, the video was segmented. The results of this pass are called subshots. Because the master shot reference is designed for use in manual assessment, a second pass over the segmentation was made to create the master shots of at least 2 seconds in length. These master shots are the ones used in submitting results for the feature and search tasks in the evaluation. In the second pass, starting at the beginning of each file, the subshots were aggregated, if necessary, until the currrent shot was at least 2 seconds in duration, at which point the aggregation began anew with the next subshot. </p><p>The keyframes were selected by going to the middle frame of the shot boundary, then parsing left and right of that frame to locate the nearest I-Frame. This then became the keyframe and was extracted. Keyframes have been provided at both the subshot (NRKF) and master shot (RKF) levels. </p><p>In a small number of cases (all of them subshots) there was no I-Frame within the subshot boundaries. When this occured, the middle frame was selected. There is one anomaly: at the end of the first video in the test collection, a subshot occurs outside a master shot.) </p><p>The emphasis in the common shot boundary reference is on the shots, not the transitions. The shots are contiguous. There are no gaps between them. They do not overlap. The media time format is based on the Gregorian day time (ISO 8601) norm. Fractions are defined by counting pre-specified fractions of a second. </p><h3>Samples</h3> <p>The Keyframe below is a sample of the data contained in this corpus.</p> <p>For information about this frame, please examine this <a href="./desc/addenda/LDC2007V01.txt" rel="nofollow">annotation file</a>. </p> </br> Portions © 2004 Cable News Network, LP, LLLP, © 2004 China Central TV, © 2004 National Broadcasting Company, Inc., © 2004 New Tang Dynasty TV, © 2004 PAC, Ltd., © 2004, 2005, 2007 Trustees of the University of Pennsylvania

创建时间:
2020-11-30
搜集汇总
数据集介绍
TRECVID 2005 Keyframes & Transcripts 数据集图片
背景与挑战
背景概述
TRECVID 2005 Keyframes & Transcripts 是用于NIST TRECVID 2005评估的多语言(阿拉伯语、中文、英语)广播新闻视频数据集,包含关键帧和转录文本,支持基于内容的检索、事件检测和信息提取等任务。其特点是通过子镜头和主镜头分割提供关键帧,并采用2秒以上镜头策略,便于视频处理研究。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务