遇见数据集

AI-Culture Commons Cultural Datasets: DOLMA + JSON + CSV

收藏
Zenodo2025-08-08 更新2026-05-29 收录
官方服务:

资源简介:

Abstract (CC-BY-4.0, 12 languages, DOLMA + JSON + CSV) Overview This dump contains the complete content from the primary websites of the AI-Culture-Commons project - a non-profit digital humanities organization specializing in cultural AI research and development. Content This release provides a fully open datasets of our project: a 1-million-word source corpus (hitdarderut-haaretz.org) and its translations (degeneration-of-nation.org) into 11 major languages (English, Spanish, French, German, Portuguese, Italian, Japanese, Russian, Korean, Mandarin Chinese and Hindi). The dump contains DOLMA, CSV and JSON exports. Datasets All datasets

提供机构:
Zenodo
创建时间:
2025-08-08
二维码
社区交流群
二维码
科研交流群
商业服务