MIRAGE
收藏资源简介:
MIRAGE是一个多模态基础模型,用于全面分析光学相干断层扫描(OCT)和扫描激光眼底照相术(SLO)图像。该数据集包含来自维也纳医科大学Macula Clinic的Vienna Imaging Biomarker Eye Study(VIBES)注册中心的261,184个配对的多模态视网膜图像样本,包括OCT和SLO图像,以及通过自动方法生成的视网膜层标签。该模型使用配对的多模态掩码自动编码(MAE)方法进行预训练,旨在从同一图像的掩码版本中重建所有输入模态。该数据集旨在为开发用于OCT和SLO图像分析的强大AI系统提供基础。
MIRAGE is a multimodal foundation model designed for comprehensive analysis of optical coherence tomography (OCT) and scanning laser ophthalmoscopy (SLO) images. This dataset comprises 261,184 paired multimodal retinal image samples sourced from the Vienna Imaging Biomarker Eye Study (VIBES) registry at the Macula Clinic of the Medical University of Vienna, including OCT and SLO images, as well as retinal layer annotations generated via automated methods. This model is pre-trained using paired multimodal masked autoencoding (MAE) methodology, with the objective of reconstructing all input modalities from the masked versions of the corresponding images. This dataset aims to provide a foundational resource for developing robust AI systems for OCT and SLO image analysis.
MIRAGE 数据集概述
数据集简介
- 名称: MIRAGE (Multimodal foundation model for comprehensive retinal OCT image analysis)
- 类型: 多模态视网膜图像分析基础模型
- 数据形式: 光学相干断层扫描(OCT)、扫描激光检眼镜(SLO)图像及视网膜层自动生成标签
- 用途: 疾病分期、诊断、视网膜层和病变分割等任务
关键特性
-
模型架构:
- 基于MultiMAE架构和Vision Transformer(ViT)
- 提供两种规模: MIRAGE-Base和MIRAGE-Large
-
评估基准:
- 包含19个任务(来自14个公开数据集和2个私有数据集)
- 涵盖OCT和SLO分类与分割任务
技术细节
- 预训练: 采用多任务学习策略的多模态自监督学习
- 系统要求:
- 操作系统: Linux
- Python版本: 3.10.x
- PyTorch版本: 2.5.1 (CUDA 11.8)
可用资源
-
模型权重:
-
评估基准数据集:
- 公开可用数据集及数据划分
- 分割基准文档: segmentation_benchmark.md
- 分类基准文档: classification_benchmark.md
使用说明
- 快速开始: 使用prepare_env.py脚本设置环境
- 推理: 提供mirage_wrapper.py脚本进行单样本推理
- 调优: 提供分类和分割任务的调优代码
许可信息
- 许可证: CC-BY-NC-ND 4.0
- 限制: 仅限非商业学术研究使用
引用格式
bibtex @article{morano2025mirage, title={{MIRAGE}: A multimodal foundation model and benchmark for comprehensive retinal {OCT} image analysis}, author={José Morano and Botond Fazekas and Emese Sükei and Ronald Fecso and Taha Emre and Markus Gumpinger and Georg Faustmann and Marzieh Oghbaie and Ursula Schmidt-Erfurth and Hrvoje Bogunović}, journal={Preprint}, year={2025} }

- 1MIRAGE: Multimodal foundation model and benchmark for comprehensive retinal OCT image analysis维也纳医科大学 · 2025年



