遇见数据集

sergey-ilinskiy-1970/rvl_cdip

收藏
Hugging Face2026-05-17 更新2026-05-31 收录
官方服务:

资源简介:

RVL-CDIP(Ryerson Vision Lab Complex Document Information Processing)数据集包含400,000张灰度图像,分为16个类别,每个类别有25,000张图像。数据集包括320,000张训练图像、40,000张验证图像和40,000张测试图像。图像的尺寸调整为其最大维度不超过1000像素。该数据集主要用于文档图像分类任务,目标是将给定文档分类为16个类别之一(如信件、表格、电子邮件等)。所有文档和类别均使用英语作为主要语言。数据集是IIT-CDIP集合的一个标记子集,适用于训练卷积神经网络进行文档分析。

The RVL-CDIP (Ryerson Vision Lab Complex Document Information Processing) dataset consists of 400,000 grayscale images in 16 classes, with 25,000 images per class. There are 320,000 training images, 40,000 validation images, and 40,000 test images. The images are sized so their largest dimension does not exceed 1000 pixels. The dataset is used for image classification tasks to classify documents into one of 16 document types (e.g., letter, form, email). All classes and documents use English as their primary language. It is a labelled subset of the IIT-CDIP collection, useful for training CNNs for document analysis.

提供机构:
sergey-ilinskiy-1970
二维码
社区交流群
二维码
科研交流群
商业服务