遇见数据集

gtfintechlab/all_annotated_sentences_25000

收藏
Hugging Face2025-05-15 更新2025-07-05 收录
官方服务:

资源简介:

该数据集是一个包含25,000个来自中央银行会议记录的句子的标注数据集,用于立场检测、时间分类和不确定性估计三个任务。每个句子都被标注了立场(鹰派、鸽派、中立、不相关)、时间属性(前瞻性、非前瞻性)和确定性(确定性、不确定性)。数据集分为训练集、测试集和验证集,每个集合包含17500个示例。

This dataset is an annotated collection of 25,000 sentences from central bank meeting minutes, designed for three tasks: Stance Detection, Temporal Classification, and Uncertainty Estimation. Each sentence is annotated with a stance (Hawkish, Dovish, Neutral, Irrelevant), a temporal attribute (Forward-looking, Not Forward-looking), and a certainty level (Certain, Uncertain). The dataset is split into training, test, and validation sets, each containing 17,500 examples.

提供机构:
gtfintechlab
二维码
社区交流群
二维码
科研交流群
商业服务