Interspeech 2016 - Experiment results for paper "Error correction in lightly supervised alignment of broadcast subtitles"
收藏资源简介:
The files in the dataset correspond to results that have been generated for the Interspeech 2016 article: "Error correction in lightly supervised alignment of broadcast subtitles" DOI: 10.21437/Interspeech.2016-56. The files in the zip file are of two types: - .ctm, which correspond to the output of the lightly supervised alignment system. - .sys, which correspond to scoring of the lightly supervised alignment system and includes the overall F measure as well as the precision and recall of the overall system and of each individual show. The following is a description about the naming convention of the files: TableX-LineY: This is the alignment and scoring output corresponding to Line Y of Table X in the article. TableX-LineY-[dev|eval]: This is the alignment and scoring output corresponding to Line Y of Table X for development (dev) or evaluation (eval) in the article. All two file types are standard outputs that are recognised by the speech technology community and can be opened using any text editor.
本数据集包含的文件均对应发表于Interspeech 2016的论文《广播字幕弱监督对齐中的误差校正》(Error correction in lightly supervised alignment of broadcast subtitles)的研究成果,该论文的DOI编号为10.21437/Interspeech.2016-56。 该压缩包内的文件分为两类: 1. .ctm格式文件:对应弱监督对齐(lightly supervised alignment)系统的输出结果; 2. .sys格式文件:对应弱监督对齐系统的性能评分结果,其中包含整体系统以及各单档节目的F值(F measure)、精确率(precision)与召回率(recall)指标。 以下为该数据集的文件命名规范说明: - TableX-LineY:对应论文中表X第Y行的对齐与评分输出结果; - TableX-LineY-[dev|eval]:对应论文中表X第Y行、针对开发集(dev)或评测集(eval)的对齐与评分输出结果。 上述两类文件均为语音技术社区公认的标准输出格式,可通过任意文本编辑器打开。



