数据链接:
官方服务:
资源简介:
These were files created through the use of AntConc and AntConc specific tools with the use of my stopwords list.
应用场景:
创建时间:
2025-03-28
相关数据集
MCL - Multifunctional Computational Lexicon of Contemporary Portuguese
MCL is a 26,443 lemma Frequency Lexicon with 140,315 tokens, with the minimum lemma frequency of 6, extracted from CORLEX, a contemporary Portuguese corpus (16,210,438 words). CORLEX is a subcorpus of
DataCite Commons2022-06-01 更新50
Frequencies and Percentages of Language Categories Across Stapel's Publications.
Note: Table 1 is organized by descending LLR. LLR values of 10.83 and 15.13 equate to ***p<.001 and ****p<.0001, †p<.01 respectively [20]. Wmatrix categories were renamed for clarity: Amplifiers = “
NIAID Data Ecosystem30
Background data for: Advancing our understanding of dispersion measures in corpus research
<p><b>Dataset description</b></p> <p>This dataset contains background data and supplementary material for Sönning (forthcoming), a study that looks at the behavior of dis
DataverseNO2025-07-17 更新30
CCC1-Anindilyakwa - Anindilyakwa_Corpus
corpus files for ANNIS corpus viewer Note that there is a set of files for each corpus. The main basic text is in the file that includes 'text' in the title.. Language as given:
Research Data Australia40
DD1-1029 - Words from Texts 2
Elicitation about the lexical meaning of various words in the corpus, mainly from Jacob's transcripts. The recorder's maximum filesize was exceeded so it split the session into two recordings.. Langua
Research Data Australia30



