audio-v2-opus
收藏资源简介:
This dataset contains >20k hours of Hebrew audio, all licensed under the [ivrit.ai v1 license](https://www.ivrit.ai/en/the-license/). It was released on April 20th, 2025. You can find the full list of sources in this dataset under the dataset's [sources.txt](https://huggingface.co/datasets/ivrit-ai/audio-v2/raw/main/sources.txt). Paper: https://arxiv.org/abs/2307.08720 If you use our datasets, the following quote is preferable: ``` @misc{marmor2023ivritai, title={ivrit.ai: A Comprehensive Dataset of Hebrew Speech for AI Research and Development}, author={Yanir Marmor and Kinneret Misgav and Yair Lifshitz}, year={2023}, eprint={2307.08720}, archivePrefix={arXiv}, primaryClass={eess.AS} } ``` # License The dataset is released under the ivrit.ai License, which enables broad research and commercial use. - Full license: https://www.ivrit.ai/en/the-license/ - FAQs: https://www.ivrit.ai/en/license-faqs/



