JuDGE
收藏资源简介:
JuDGE数据集是由清华大学计算机科学与技术系构建的,包含真实案件的事实描述及其对应的完整判决书。该数据集还包括两个专门的法律语料库,一个是包含法律法规的语料库,另一个是包含大量过去判决书的语料库,作为任务中法律知识的外部来源。数据集的构建过程包括公开收集案件文档、法律条文,经过预处理和专家标注,以确保数据质量。该数据集旨在用于评估判决书生成性能,并推动相关领域的研究进展。
The JuDGE Dataset is constructed by the Department of Computer Science and Technology, Tsinghua University. It contains factual descriptions of real cases and their corresponding complete judicial judgments. The dataset also includes two specialized legal corpora: one is a corpus containing laws and regulations, and the other is a corpus comprising a large number of past judicial judgments, which serve as external sources of legal knowledge for relevant tasks. The dataset's construction process involves publicly collecting case documents and legal provisions, followed by preprocessing and expert annotation to ensure data quality. This dataset is designed to evaluate the performance of judicial judgment generation and advance research progress in related fields.




