<p>The <strong>ESC dataset</strong> is a collection of short environmental recordings available in a unified format (5-second-long clips, 44.1 kHz, single channel, Ogg Vorbis compres
Raw data used for the statistical analyses presented in the article "Anatomical unconstrained regions of the vocal tract determine acoustical correlates of individual identity in African penguins"
It seems trivial to identify sound sequences as music or speech, particularly when the sequences come from different sound sources, such as an orchestra and a human voice. Can we also easily distingui
we manually listened to all the sounds and classified them into four categories: [Type-0] Sounds that are completely irrelevant to the corresponding labels; [Type-1] Earcon (The identification of Earc