UnpredicTable数据集由从互联网表格中提取的少样本任务组成,旨在通过微调语言模型来提高其在少样本学习中的表现。数据集包含多个版本和子集,涵盖了广泛的任务类型,如多项选择、问答、文本分类等。数据集的创建基于WDC Web Table Corpus,并通过自动化的方式将表格转换为少样本学习任务。
To investigate the problem of classifying source code reviews, we have created a dataset suitable for evaluating and testing various methods for solving this problem. We combined four open datasets, a