遇见数据集

FPT Open Speech Dataset (FOSD) - Vietnamese

收藏
Mendeley Data2020-07-31 更新2026-04-09 收录
官方服务:

资源简介:

This dataset consists of 25,921 recorded Vietnamese speeches (with their transcripts and the labelled start and end times of each speech) manually compiled from 3 sub-datasets (approximately 30 hours in total) released publicly in 2018 by FPT Corporation. The speeches are in *.mp3 format while the transcript file is in *.txt format with utf-8 encoding scheme. The dataset is useful for several speech-related research topics, including but not limited to text-to-speech, speech-to-text applications, gender detection, mood detection, intent detection, onset detection, signal-to-noise improvement, signal processing, speech processing, etc. Copyright 2018 FPT Corporation Permission is hereby granted, free of charge, non-exclusive, worldwide, irrevocable, to any person obtaining a copy of this data or software and associated documentation files (the “Data or Software”), to deal in the Data or Software without restriction, including without limitation the rights to use, copy, modify, remix, transform, merge, build upon, publish, distribute and redistribute, sublicense, and/or sell copies of the Data or Software, for any purpose, even commercially, and to permit persons to whom the Data or Software is furnished to do so, subject to the following conditions: The above copyright notice, and this permission notice, and indication of any modification to the Data or Software, shall be included in all copies or substantial portions of the Data or Software. THE DATA OR SOFTWARE IS PROVIDED “AS IS”, WITHOUT WARRANTY OF ANY KIND, EXPRESS OR IMPLIED, INCLUDING BUT NOT LIMITED TO THE WARRANTIES OF MERCHANTABILITY, FITNESS FOR A PARTICULAR PURPOSE AND NONINFRINGEMENT. IN NO EVENT SHALL THE AUTHORS OR COPYRIGHT HOLDERS BE LIABLE FOR ANY CLAIM, DAMAGES OR OTHER LIABILITY, WHETHER IN AN ACTION OF CONTRACT, TORT OR OTHERWISE, ARISING FROM, OUT OF OR IN CONNECTION WITH THE DATA OR SOFTWARE OR THE USE OR OTHER DEALINGS IN THE DATA OR SOFTWARE. Patent and trademark rights are not licensed under this FPT Public License.

本数据集共包含25,921条录制的越南语语音数据,附带对应转录文本及每条语音的标注起始与结束时间,由FPT公司(FPT Corporation)于2018年公开发布的3个子数据集(总时长约30小时)手动汇编而成。语音文件采用*.mp3格式,转录文本文件则采用UTF-8编码的*.txt格式。本数据集可应用于多项语音相关研究主题,包括但不限于文本转语音(text-to-speech)、语音转文字(speech-to-text)应用、性别识别、情绪识别、意图识别、语音起始点检测、信噪比提升、信号处理、语音处理等研究方向。本数据集版权归2018年FPT公司所有。特此免费授予任何获得本数据或软件及其相关文档(以下简称"Data or Software")的个人全球范围内的非排他性、不可撤销的使用许可,允许其不受限制地处理数据或软件,包括但不限于为任何目的(包括商业用途)使用、复制、修改、改编、转换、合并、衍生创作、发布、分发、再分发、转授权以及/或者销售数据或软件的副本,并允许接收数据或软件的个人同样享有上述权利,但需遵守以下条件:上述版权声明、本许可声明以及对数据或软件所做修改的说明,必须包含在所有数据或软件的副本或实质部分中。本数据或软件按"现状"提供,不附带任何明示或暗示的担保,包括但不限于适销性、特定用途适用性以及不侵权的担保。在任何情况下,作者或版权持有人均不对因本数据或软件、或本数据或软件的使用或其他交易行为所产生的任何索赔、损害或其他责任承担责任,无论该责任产生于合同诉讼、侵权行为或其他情形。本FPT公开许可不授予专利与商标权。

创建时间:
2020-07-31
二维码
社区交流群
二维码
科研交流群
商业服务