LJ Speech - Aligned IPA transcriptions
收藏资源简介:
Files: <code>grids.zip</code> contains TextGrids for all audio files containing three tiers <code>words</code>, <code>phonemes</code> and <code>transcription</code> <code>words</code> contains the aligned normalized English words <code>phonemes</code> contains IPA pronunciations transcribed using CMU dictionary which then were aligned with Montreal Forced Aligner. The pronunciations were then mapped from ARPAbet to IPA and duration marks were applied (without punctuation) <code>transcription</code> contains unaligned phonemes including punctuation and word boundary labels (SIL0) <code>preview.png</code> preview of the first TextGrid opened in Praat <code>words-vocabulary.txt</code> contains all words from tier <code>words</code> <code>phonemes-vocabulary.txt</code> contains all phonemes from tier <code>phonemes</code> <code>transcription-vocabulary.txt</code> contains all phonemes/punctuation from tier <code>transcription</code> <code>phonemes-durations.pdf</code> contains the plotted phoneme duration distribution of tier <code>phonemes</code> <code>phonemes-durations-simple.pdf</code> contains the plotted phoneme duration distribution of tier <code>phonemes</code> if all duration markers are ignored <code>pronunciations.dict</code> contains the pronunciations for each word including punctuation and weights (occurrence) <code>script.sh</code> contains the script to reproduce all results Phoneme duration marker: <code>˘</code> -> [0, 20) percentile <code>ˑ</code> -> [80, 90) percentile <code>ː</code> -> [90, inf) percentile Silence marker: <code>SIL0</code> -> no silence <code>SIL1</code> -> [0, 33.33) percentile <code>SIL2</code> -> [33.33, 66.66) percentile <code>SIL3</code> -> [66.66, inf) percentile



