遇见数据集

mteb/MTOPIntentClassification

收藏
Hugging Face2025-09-09 更新2025-10-25 收录
官方服务:

资源简介:

MTOPIntentClassification数据集是大规模文本嵌入基准(MTEB)的一部分,旨在用于文本分类任务。该数据集是多语言的,包括德语、英语、法语、印地语、西班牙语和泰语。数据集由人类标注,包含训练、验证和测试三个部分,每个部分的数据量和字节数不同。该数据集来源于mteb/mtop_intent数据集,是MTOP(多语言任务导向语义解析)项目的一部分。README中还提供了如何使用mteb库评估该数据集上的模型,以及引用数据集的说明。

The MTOPIntentClassification dataset is part of the Massive Text Embedding Benchmark (MTEB) and is designed for text classification tasks. It is a multilingual dataset featuring languages such as German, English, French, Hindi, Spanish, and Thai. The dataset is human-annotated and includes training, validation, and test splits with varying sizes and byte counts. It is sourced from the mteb/mtop_intent dataset and is part of the MTOP (Multilingual Task-Oriented Semantic Parsing) project. The README also includes instructions on how to evaluate models on this dataset using the mteb library and guidelines for citation.

提供机构:
mteb
二维码
社区交流群
二维码
科研交流群
商业服务