OpenCitations Corpus dataset of all the resource embodiments, archived on 2017-07-25
收藏资源简介:
This archive contains the dump of the OpenCitation Corpus (OCC, http://opencitations.net) dataset about resource embodiments, created regularly every month.After unzipping the archive, Disk ARchive (DAR, http://dar.linux.free.fr/, a multi-platform archive tool for managing huge amount of data) is needed for recreating the whole structure. For extracting the DAR archive, please run the commanddar -x [archive-name]Where "[archive-name"] is the name of the DAR file without final package number and extension. E.g.:dar -x 2016-09-23-corpus_reFor further questions, comments, and suggestions please don't hesitate to contact Silvio Peroni at essepuntato@opencitations.net.
本归档文件包含每月定期生成的开放引用语料库(OpenCitation Corpus, OCC,http://opencitations.net)中关于资源实体的数据集转储文件。解压该归档后,需使用Disk ARchive(DAR,http://dar.linux.free.fr/,一款用于管理海量数据的多平台归档工具)来还原完整的文件结构。若要提取DAR归档文件,请执行命令:dar -x [archive-name],其中"[archive-name]"为DAR文件的名称,无需携带末尾的包编号与扩展名。例如:dar -x 2016-09-23-corpus_re。如有进一步疑问、评论或建议,请随时联系Silvio Peroni,邮箱地址为essepuntato@opencitations.net。



