遇见数据集

amuvarma/humanml3d-flat-train-t5-toks-grouped-dupped-6k

收藏
Hugging Face2025-02-09 更新2025-02-15 收录
官方服务:

资源简介:

该数据集包含三个特征字段:input_ids、attention_mask和labels。input_ids和attention_mask是整数序列,分别可能表示输入文本的索引和注意力掩码;labels是整数序列,可能表示某种标签或分类。数据集分为训练集,共有6000个示例。数据集的总大小为22975478字节,下载大小为3631683字节。由于README中未提供详细描述,具体内容未知。

The dataset includes three feature fields: input_ids, attention_mask, and labels. input_ids and attention_mask are integer sequences which might represent input text indices and attention masks respectively; labels are integer sequences which might represent some kind of label or classification. The dataset is split into a training set with a total of 6000 examples. The overall size of the dataset is 22975478 bytes, with a download size of 3631683 bytes. As no detailed description is provided in the README, the specific content is unknown.

提供机构:
amuvarma
二维码
社区交流群
二维码
科研交流群
商业服务