遇见数据集

audio-v2-opus

收藏
魔搭社区2026-07-02 更新2026-09-06 收录
官方服务:

资源简介:

This dataset contains >20k hours of Hebrew audio, all licensed under the [ivrit.ai v1 license](https://www.ivrit.ai/en/the-license/). It was released on April 20th, 2025. You can find the full list of sources in this dataset under the dataset's [sources.txt](https://huggingface.co/datasets/ivrit-ai/audio-v2/raw/main/sources.txt). Paper: https://arxiv.org/abs/2307.08720 If you use our datasets, the following quote is preferable: ``` @misc{marmor2023ivritai, title={ivrit.ai: A Comprehensive Dataset of Hebrew Speech for AI Research and Development}, author={Yanir Marmor and Kinneret Misgav and Yair Lifshitz}, year={2023}, eprint={2307.08720}, archivePrefix={arXiv}, primaryClass={eess.AS} } ``` # License The dataset is released under the ivrit.ai License, which enables broad research and commercial use. - Full license: https://www.ivrit.ai/en/the-license/ - FAQs: https://www.ivrit.ai/en/license-faqs/

提供机构:
maas
创建时间:
2026-02-06
二维码
社区交流群
二维码
科研交流群
商业服务