Dieser Ordner beinhaltet zehn Listen, die die Concgrams (frequente Wortkombinationen innerhalb einer bestimmten Wortspanne) jedes Subkorpus darstellen. Die Concgrams bestehen aus drei Wörtern (ein Sub
Co-occurrence table created with the tool “TXM”. More information about TXM can be found at: Serge Heiden/Jean-Philippe Magué/Bénédicte Pincemin: TXM : Une plateforme logicielle open-source pour la te
Pickle file containing the PMI, calculated on the training corpus, between 2-event-tuples, using the standard PMI (called pmi_{standard} in the paper).
Dans le tableau Excel ci-dessous se trouvent les cooccurrents du mot « sikh » et de ses dérivés dans notre corpus, classés selon divers thèmes récurrents. Le chiffre accompagnant chaque cooccurrent in