QAW: A Quality Assurance Workflow for Ontologies based on Detecting Semantic Regularities - Dataset on SNOMED
收藏资源简介:
This page contains supplementary material for the ISWC 2013 submission with title: <em>"QAW: A Quality Assurance Workflow for Ontologies based on Detecting Semantic Regularities" </em>Eleni Mikroyannidi, Manuel Quesada-Mart ́ınez, Dmitry Tsarkov, Jesualdo Tomas Fernandez Breis, Robert Stevens, Ignazio Palmisano. The fileset contains the data for the qualitative and quantititative analysis that were presented in the paper. In the <strong>qualitative analysis</strong>, six lexical patterns (keywords) that were processed. These are: "chronic","acute", "absent", "present", "right", "left". For these ones the reader can browse and download the following data: 1. XML Files with the generic name "<strong>keyword"_syntactic_usage.xml</strong> which contains the detected syntactic regularities for the referencing asserted axioms of the entities that contain the corresponding keyword in their label. There should be 6 files in total (for each keyword). 2. XML Files with the generic name "<strong>keyword"_semantic_usage.xml</strong> which contains the detected syntactic regularities for the referencing asserted axioms of the entities that contain the keyword in their label. In the <strong>quantitative analysis</strong>, 308 lexical patterns were processed, and corresponding syntactic and semantic regularities were detected. The dataset that is available for the reader contains the following: 1. <strong>LexAnal_Snomed_2013_NoSensitiveAnalysis_Cov_0.1_100.0.xml</strong>, which contains all lexical patterns that could be detected in the SNOMED-CT version January 2013. 2. <strong>Snomed_2013_LexAnal_Full_0.1-0.4Perc_.xml</strong>, which contains all lexical patterns with 0.1%-0.4% lexical pattern threshold. 3. <strong>syntactic_regularities_dataset.zip</strong> which contains 308 xml files with the syntactic regularities that were generated by RIO. 4. <strong>semantic_regularities_dataset.zip</strong> which contains 308 xml files with the semantic regularities that were generated by RIO. 5. <strong>quantitative_syntactic_regularity_analysis.csv</strong>, which contains the syntactic regularity stat analysis for the 308 processed cases. 6.<strong> quantitative_semantic_regularity_analysis.csv, </strong>which contains the semantic regularity stat analysis for the 308 processed cases.



