United Nations Proceedings Speech
收藏DataCite Commons2021-07-01 更新2025-04-16 收录
下载链接:
https://catalog.ldc.upenn.edu/LDC2014S08
下载链接
链接失效反馈官方服务:
资源简介:
<h3>Introduction</h3><br>
<p>United Nations Proceedings Speech was developed by the <a href="http://www.un.org/">United Nations</a> (UN) and contains approximately 8,500 hours of recorded proceedings in the six official UN languages, Arabic, Chinese, English, French, Russian and Spanish. The data was recorded in 2009-2012 from sessions 64-66 of the <a href="http://www.un.org/en/ga/">General Assembly</a> (GA) and <a href="http://www.un.org/en/ga/first/">First Committee</a> (FC) (Disarmament and International Security), and meetings 6434-6763 of the <a href="http://www.un.org/en/sc/">Security Council</a>.</p><br>
<p>Recordings were made using a customized system following a daily internal circulated instruction from the <a href="http://www.un.org/depts/DGACM/mms.shtml">Meetings Management Section</a>. Most of the subjects and information related to a particular meeting or session are published in a UN Journal which can be found in the following link: <a href="http://www.un.org/en/documents/journal.asp">http://www.un.org/en/documents/journal.asp</a></p><br>
<h3>Data</h3><br>
<p>Data is presented either as mp3 or flac compressed wav and are 16-bit single channel files in either 22,050 or 8,000 Hz organized by committee and session number, then language. The folder labeled "Floor" indicates the microphone used by the particular speaker. Those files may include other languages, for instance, if the speaker's language was not among the six official UN languages.</p><br>
<p>File naming conventions for GA and FC data are in the form of LYY_ZZ_format.format and Security Council data is in the form of LYYYY_ZZ_format.format where L is a one letter language designation, YY is the meeting number, ZZ indicates the audio segment number and format.format is the wav or mp3 designation. Note that not all files are present for every language.</p><br>
<h3>Samples</h3><br>
<p>Please listen to the following samples</p><br>
<ul><br>
<li><a href="desc/addenda/LDC2014S08.flr.wav">Floor</a></li><br>
<li><a href="desc/addenda/LDC2014S08.A.wav">Arabic</a></li><br>
<li><a href="desc/addenda/LDC2014S08.C.wav">Chinese</a></li><br>
<li><a href="desc/addenda/LDC2014S08.E.wav">English</a></li><br>
<li><a href="desc/addenda/LDC2014S08.F.wav">French</a></li><br>
<li><a href="desc/addenda/LDC2014S08.R.wav">Russian</a></li><br>
<li><a href="desc/addenda/LDC2014S08.S.wav">Spanish</a></li><br>
</ul><br>
<h3>Updates</h3><br>
<p>None at this time.</p><br>
<p> </p></br>
Portions © 2009-2012, 2014 United Nations, © 2014 Trustees of the University of Pennsylvania
提供机构:
Linguistic Data Consortium
创建时间:
2020-11-30



