Dieser Ordner beinhaltet zehn Listen, die die Concgrams (frequente Wortkombinationen innerhalb einer bestimmten Wortspanne) jedes Subkorpus darstellen. Die Concgrams bestehen aus drei Wörtern (ein Sub
A collection of n-grams extracted from the Gos corpus of spoken Slovene (cf. http://eng.slovenscina.eu/korpusi/gos). Three sets of n-gram lists are provided for lowercased word...
Co-occurrence table created with the tool “TXM”. More information about TXM can be found at: Serge Heiden/Jean-Philippe Magué/Bénédicte Pincemin: TXM : Une plateforme logicielle open-source pour la te
Google n-gram data (for n=5) derived from "English Version 20090715" on the Google n-gram download page. This is a count of 5-gram strings of words from the 1500's, 1600's, and 1700's, for the sake of