AI-Culture Commons Cultural Datasets: DOLMA + JSON + CSV
收藏官方服务:
资源简介:
Abstract (CC-BY-4.0, 12 languages, DOLMA + JSON + CSV) Overview This dump contains the complete content from the primary websites of the AI-Culture-Commons project - a non-profit digital humanities organization specializing in cultural AI research and development. Content This release provides a fully open datasets of our project: a 1-million-word source corpus (hitdarderut-haaretz.org) and its translations (degeneration-of-nation.org) into 11 major languages (English, Spanish, French, German, Portuguese, Italian, Japanese, Russian, Korean, Mandarin Chinese and Hindi). The dump contains DOLMA, CSV and JSON exports. Datasets All datasets
提供机构:
Zenodo创建时间:
2025-08-08



