Darshana Graph
收藏资源简介:
Darshana Graph是由独立研究员构建的一个大规模平行评论文本语料库,专注于印度古典哲学的比较研究。该数据集包含约12.5万条文本记录,核心内容涵盖印度教、佛教和耆那教哲学传统,其中巴利三藏占主导地位(114,591条),其独特价值在于约8,500条印度教和耆那教记录实现了跨学派对齐,同一原始经文与十八位不同评论家的注释被平行关联。数据主要来源于公共领域和开放许可的经典译本,如《薄伽梵歌》《梵经》等,通过系统化的对齐架构和模式设计,确保了多注释传统的直接可比性。该数据集旨在支持数字人文研究,特别是通过风格计量学和知识图谱提取等方法,深入探究不同哲学流派的论证风格、概念关系及解释传统差异。
Darshana Graph is a large-scale parallel commentary corpus constructed by independent researchers, focused on comparative studies of classical Indian philosophy. This dataset contains roughly 125,000 text records, with core content covering the philosophical traditions of Hinduism, Buddhism, and Jainism, where the Pāli Canon dominates with 114,591 records. Its unique value stems from approximately 8,500 Hindu and Jain records that enable cross-school alignment: the same original scriptures are parallelly aligned with commentaries from 18 distinct commentators. The dataset is primarily sourced from public domain and open-licensed classical translations, including works such as the Bhagavad Gita and Brahma Sutras. Through a systematic alignment framework and schema design, it ensures direct comparability across multiple commentary traditions. This corpus is intended to support digital humanities research, specifically enabling in-depth inquiries into argumentative styles, conceptual relationships, and disparities in interpretive traditions across different philosophical schools, via methodologies including stylometrics and knowledge graph extraction.




