redactable-llm/synth-text-recognition
收藏官方服务:
资源简介:
--- dataset_info: features: - name: image dtype: image - name: label dtype: string splits: - name: train num_bytes: 12173747703 num_examples: 7224600 - name: val num_bytes: 1352108669.283 num_examples: 802733 - name: test num_bytes: 1484450563.896 num_examples: 891924 download_size: 12115256620 dataset_size: 15010306936.179 task_categories: - image-to-text language: - en size_categories: - 1M<n<10M pretty_name: MJSynth --- # Dataset Card for "Synth-Text Recognition" This is the dataset for text recognition on document images, synthetically generated, covering 90K English words. It includes training, validation and test splits.
提供机构:
redactable-llm


