Precomputed Embeddings for Spectraverse
收藏官方服务:
资源简介:
Precomputed embeddings for the Spectraverse dataset. Data is split in train/test set such that no molecules in the test as the same formulae than a molecule in the train test. Spectra are encoded with DreAMS Molecules are encoded with ChemBERTa (Derify/ChemBERTa_augmented_pubchem_13m)
针对Spectraverse数据集的预计算嵌入向量(embeddings)。 数据集已按训练集与测试集划分,且测试集中的所有分子均不会与训练集中的分子共享相同的分子式。 光谱数据采用DreAMS进行编码。 分子数据采用ChemBERTa (Derify/ChemBERTa_augmented_pubchem_13m)进行编码。
提供机构:
Zenodo创建时间:
2026-03-28



