遇见数据集

shivanikerai/ean_filter_data_v0.3

收藏
Hugging Face2025-12-12 更新2025-12-20 收录
官方服务:

资源简介:

--- dataset_info: features: - name: image dtype: image - name: caption dtype: string splits: - name: train num_bytes: 1203332097.0 num_examples: 10000 - name: test num_bytes: 1042178383.75 num_examples: 8930 download_size: 2199228199 dataset_size: 2245510480.75 configs: - config_name: default data_files: - split: train path: data/train-* - split: test path: data/test-* ---

This is a multimodal dataset containing images and their corresponding text descriptions, primarily intended for image caption generation or related tasks. The dataset includes 10,000 training samples and 8,930 test samples, where each sample consists of one image and a text description. The approximate download size of the dataset is 2.2 GB, with a total storage size of around 2.25 GB.

提供机构:
shivanikerai
二维码
社区交流群
二维码
科研交流群
商业服务