遇见数据集

OpenCitations Index CSV dataset of the provenance information of all the citation data

收藏
Mendeley Data2024-06-29 更新2024-06-30 收录
官方服务:

资源简介:

This dataset contains the provenance information (in CSV format) of all the citation data included in the OpenCitations Index, released on 29 November 2023. In particular, each line of the CSV file defines a citation, and includes the following information:[field "oci"] the Open Citation Identifier (OCI) for the citation;[field "snapshot"] the identifier of the snapshot;[field "agent"] the name of the agent that have created the citation data;[field "source"] the URL of the source dataset from where the citation data have been extracted;[field "created"] the creation time of the citation data.[field "invalidated"] the start of the destruction, cessation, or expiry of an existing entity by an activity;[field "description"] a textual description of the activity made;[field "update"] the UPDATE SPARQL query that keeps track of which metadata have been modified.The size of the zipped archive is 14 GB, while the size of the unzipped CSV files is 312 GB.

本数据集包含2023年11月29日发布的开放引用索引(OpenCitations Index)所收录的全部引用数据的溯源信息,数据格式为CSV。具体而言,CSV文件的每一行均对应一条引用记录,包含以下字段: 1. 字段oci:该引用的开放引用标识符(Open Citation Identifier,OCI); 2. 字段snapshot:快照标识符; 3. 字段agent:生成该引用数据的主体名称; 4. 字段source:提取该引用数据的源数据集的URL; 5. 字段created:该引用数据的创建时间; 6. 字段invalidated:某一活动致使现有实体失效、终止或过期的起始时点; 7. 字段description:所执行活动的文本描述; 8. 字段update:用于追踪元数据修改情况的UPDATE SPARQL查询语句。 该压缩归档文件的大小为14 GB,解压后的CSV文件总容量为312 GB。

创建时间:
2023-10-27
二维码
社区交流群
二维码
科研交流群
商业服务