遇见数据集

Derived features and metadata from a corpus of digital media content on public images of science in Bulgaria (2011–2024)

收藏
Zenodo2026-03-22 更新2026-05-29 收录
官方服务:

资源简介:

This dataset contains structured metadata and derived analytical features extracted from a large corpus of digital media content related to public images of science in Bulgaria, covering the period 2011–2024. The corpus includes online news articles and user-generated comments collected from digital media platforms. Due to copyright and legal restrictions, the dataset does not provide full original textual content. Instead, it contains metadata and computationally derived representations suitable for quantitative and qualitative analysis. The dataset includes variables such as publication date, source links, thematic classifications (topic modeling), extracted keywords (top words per topic), sentiment/tone indicators, and text length. These features enable the analysis of thematic structures, public discourse, and temporal dynamics in media representations of science. The dataset was created within the research project “Images of Science in the Digital World: Between Social Media and Internet Media”, implemented by the Institute of Philosophy and Sociology at the Bulgarian Academy of Sciences and funded by the Bulgarian National Science Fund (Contract No. КП 06-Н55/6, November 2021). The dataset is openly available for scientific research, educational use, and secondary analysis.

提供机构:
Zenodo
创建时间:
2026-03-22
二维码
社区交流群
二维码
科研交流群
商业服务