From Voice To Shell: Speech To Docker Dataset
收藏资源简介:
Dataset with 3,192 audio files, 3.92 hours, comprising voice recordings of 12 subjects (4 women and 8 men) while enunciating the prompts of 'text-to-docker' dataset test samples in English (EN). Followed protocol has been approved by the Social Research Ethics Committee (SREC) of the University of Castilla-La Mancha (UCLM) under reference number CEIS-2025-119450. Human participation is justified and the benefits and risks have been adequately assessed to participants, developing a participation mechanism which warrants equal opportunities for research collaboration. Every participant has been provided an informed consent document including the necessary information of research purposes and requirements. It also complies with the regulations in force regarding the personal data protection. Audio files are encrypted, which could be extracted using the decrypt.py script: user@host:#~/data/$ lsen decrypt.py metadatauser@host:#~/data/$ pip install cryptography[... install dependencies ...]user@host:#~/data/$ python3 decrypt.py -p <password> .[... decrypt files ...]



