官方服务:
资源简介:
数据集中共有 8,142 个物种,有 437,513 个训练图像和 24,426 个验证图像。每张图像都有一个真实标签。
This dataset comprises 8,142 species, including 437,513 training images and 24,426 validation images. Each image is paired with a ground-truth label.
应用场景:
提供机构:
OpenDataLab创建时间:
2022-03-17
搜集汇总
数据集介绍

背景与挑战
背景概述
iNaturalist 2018是一个大规模图像分类数据集,包含8,142个物种,总计461,939张图像(437,513张训练图像和24,426张验证图像),每张图像都有真实标签,适用于物种分类和检测任务。该数据集由加州理工学院于2018年发布,常用于计算机视觉研究,支持生物多样性监测和自然场景识别。
以上内容由遇见数据集搜集并总结生成
相关数据集
Financial Document Image Dataset
A Collection of Personal Financial Documents from India
kaggle2023-05-12 更新340
Google Cloud Vision
Google Cloud Vision API — Powerful image analysis powered by Google's machine learning. Detect objects, read text (OCR), identify faces, landmarks, logos, and more. The same technology behind Google Photos and Google Lens. ## Why Use This API? - **Label Detection** — Identify objects, scenes, and activities in images - **OCR (Text Detection)** — Extract text from images in 100+ languages - **Face Detection** — Detect faces with emotion and landmark analysis - **Landmark Detection** — Identi...
RapidAPI2026-03-28 更新10
GLIB: image dataset
data/images: data/images/Base : 132 张游戏 1 和游戏 2 的屏幕截图,来自 466 份测试报告的 UI 显示问题。 data/images/Code:9,412 张游戏 1 和游戏 2 的屏幕截图,其中包含由我们的代码增强方法生成的 UI 显示问题。 data/images/Normal:随机遍历游戏场景收集到的game1和game2的7750张无UI显示问题的截图。 data/images/Rule(F) :7,750 张游戏 1 和游戏 2 的屏幕截图,其中 UI 显示问题由我们的 Rule(F) 增强方法生成。 data/images/Rule(R) :游戏 1 和游戏 2 的 7,750 个屏幕截图,其中 UI 显示问题由我们的 Rule(R) 增强方法生成。 data/images/testDataSet :来自 466 个测试报告(不包括 game1 和 game2)的 192 个带有 UI 显示问题的屏幕截图。 data/data_csv: data/data_csv/Base : 基线方法的数据集。 data/data_csv/Code :我们的代码增强方法的数据集。 data/data_csv/Rule(F) :我们的 Rule(F) 增强方法的数据集。 data/data_csv/Rule(R) :我们的 Rule(R) 增强方法的数据集。 data/data_csv/Code_plus_Rule(F) :我们的 Code&Rule(F) 增强方法的数据集。 data/data_csv/Code_plus_Rule(R) :我们的 Code&Rule(R) 增强方法的数据集。 data/data_csv/testDataSet : 测试数据集(来自 466 个测试报告的正常图像和真实故障图像)。
OpenDataLab2026-07-12 更新220
pouya-haghi/MSCOCO-1k
--- dataset_info: features: - name: image dtype: image - name: filepath dtype: string - name: sentids list: int32 - name: filename dtype: string - name: imgid dtype: int32 - name: split dtype: string - name: sentences struct: - name: tokens list: string - name: raw dtype: string - name: imgid dtype: int32 - name: sentid dtype: int32 - name: cocoid dtype: int32 splits: - name: test num_bytes: 169158987.72768652 num_examples: 1024 download_size: 167657377 dataset_size: 169158987.72768652 configs: - config_name: default data_files: - split: test path: data/test-* ---
Hugging Face2024-01-01 更新170
clothes-dataset
这是一个包含来自Aliexpress产品评论的照片数据集。共有379186张图片,数据量约30Gb。该数据集非常适合用于训练模型以找到相似的图像。数据集通过半手动方式进行了清理,部分图像被标记为好/坏。然后训练了一个模型来清理剩余的数据集。
github2023-10-06 更新280
centurion_reverse1999
这是一个名为'Dataset of Centurion/百夫长 (Reverse:1999)'的数据集,包含22张图片及其标签。这些图片主要描述了一个具有深色皮肤、长发、黑色头发等特征的角色。数据集提供了原始数据和经过裁剪的数据包,以及标签聚类的结果,帮助用户更好地理解和利用数据集中的图片。
Hugging Face2024-08-05 更新30
FRUIT VS VEGETABLE
BINARY CLASSIFICATION OF FRUITS AND VEGETABLES (550 IMAGES 512x512)
kaggle2020-10-31 更新80
A-Z Handwritten Alphabets in .csv format .csv 格式的 A-Z 手写字母数据集
.csv 格式的 AZ 手写字母数据集是一个大规模的英文手写字母图像集合,专为手写识别任务而设计。
超神经2024-06-21 更新330
植物OCR识别-动物OCR识别-菜品OCR识别-水果OCR识别-logoOCR识别
基于行业前沿的人工智能技术,为用户提供菜品、水果、logo、动植物、货币、地标、商品等的识别服务。广泛应用于数字营销、新零售、广告设计、园林景观等行业场景。
腾讯云市场70
Multimodal-Fatima/VQAv2_test_no_image
--- dataset_info: features: - name: question_type dtype: string - name: multiple_choice_answer dtype: string - name: answers_original list: - name: answer dtype: string - name: answer_confidence dtype: string - name: answer_id dtype: int64 - name: id_image dtype: int64 - name: answer_type dtype: string - name: question_id dtype: int64 - name: question dtype: string - name: id dtype: int64 - name: clip_tags_ViT_L_14 sequence: string - name: blip_caption dtype: string - name: LLM_Description_gpt3_downstream_tasks_visual_genome_ViT_L_14 sequence: string - name: DETA_detections_deta_swin_large_o365_coco_classes list: - name: attribute dtype: string - name: box sequence: float32 - name: label dtype: string - name: location dtype: string - name: ratio dtype: float32 - name: size dtype: string - name: tag dtype: string - name: Attributes_ViT_L_14_descriptors_text_davinci_003_full sequence: string - name: clip_tags_ViT_L_14_wo_openai sequence: string - name: clip_tags_ViT_L_14_with_openai sequence: string - name: clip_tags_LAION_ViT_H_14_2B_wo_openai sequence: string - name: clip_tags_LAION_ViT_H_14_2B_with_openai sequence: string - name: clip_tags_LAION_ViT_bigG_14_2B_wo_openai sequence: string - name: clip_tags_LAION_ViT_bigG_14_2B_with_openai sequence: string - name: Attributes_LAION_ViT_H_14_2B_descriptors_text_davinci_003_full sequence: string - name: Attributes_LAION_ViT_bigG_14_2B_descriptors_text_davinci_003_full sequence: string - name: DETA_detections_deta_swin_large_o365_coco_classes_caption_module_random list: - name: attribute dtype: string - name: box sequence: float64 - name: captions_module sequence: string - name: captions_module_filter sequence: string - name: label dtype: string - name: location dtype: string - name: ratio dtype: float64 - name: size dtype: string - name: tag dtype: string - name: answers sequence: string splits: - name: test num_bytes: 21976237587 num_examples: 447793 download_size: 5670512625 dataset_size: 21976237587 --- # Dataset Card for "VQAv2_test_no_image" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
Hugging Face2023-05-13 更新150
mnist(MNIST, FashionMNIST, EMNIST)
处理好的torchvision.datasets中的mnist(MNIST, FashionMNIST, EMNIST)数据集,下载下来可直接使用,方便国内同学。这些数据集很常用,作为新手可以让你专注搭建模型。
github2024-04-11 更新1370
datacomp/imagenet-1k-random-50.0-frac-1over2
该数据集包含图像和对应的标签,主要用于训练目的。数据集包含640,583个训练样本,总大小为51,659,759,699.125字节。数据集的下载大小为51,650,685,014字节。数据文件路径在默认配置中指定为train-*。
Hugging Face2025-01-08 更新130
RefOI Dataset
RefOI数据集是由美国密歇根大学等机构创建,包含约1487张图片,每张图片都标注有3个书面和2个口头指代表达式。该数据集的创建旨在解决现有数据集存在的数据泄露问题以及包含过多由大型语言模型生成的描述,缺乏真实人类对话中使用的口头语言的问题。
arXiv2025-04-23 更新290
Weather Image Dataset: Sunny, Overcast, Rainy
A small dataset of 510 labeled sky images for weather classification.
kaggle2025-08-30 更新180
Untitled Item
The MIDOG++ dataset represents an extension of the data set used in the MIDOG 2021 and 2022 challenges. We provide region of interest images from 503 histological specimens of seven different tumor types with variable morphology: breast carcinoma, lung carcinoma, lymphosarcoma, neuroendocrine tumor, cutaneous mast cell tumor, cutaneous melanoma, and (sub)cutaneous soft tissue sarcoma. The human and canine samples were processed and stained at different human and veterinary pathology laboratories with standard H&E dye and digitized with different digital whole slide image scanners. We provide labels for 11,937 mitotic figures that have been differentiated against 14,351 imposter cells in a blinded consensus by two pathologists and a final decision by a third pathologist for disagreed labels.
DataCite Commons2025-06-01 更新50
cjensen/celeb-identities
--- dataset_info: features: - name: image dtype: image - name: label dtype: class_label: names: '0': Carrot_Top '1': Chris_Hemsworth '2': Gru '3': Michael_Jordan '4': Mother_Teresa '5': Winona_Ryder splits: - name: train num_bytes: 8636520.0 num_examples: 18 download_size: 8635182 dataset_size: 8636520.0 --- # Dataset Card for "celeb-identities" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
Hugging Face2023-05-13 更新90
Waste classification
The dataset contains images for classification and detection.
kaggle2023-11-26 更新120
fw407/cifar10
这是一个包含图像和对应分类标签的数据集,共有10个分类,分别是飞机、汽车、鸟、猫、鹿、狗、青蛙、马、船和卡车。数据集分为训练集和测试集,训练集包含50000个示例,测试集包含10000个示例。
Hugging Face2025-10-23 更新30



