遇见数据集

Precomputed Embeddings for Spectraverse

收藏
Zenodo2026-03-31 更新2026-05-26 收录
官方服务:

资源简介:

Precomputed embeddings for the Spectraverse dataset. Data is split in train/test set such that no molecules in the test as the same formulae than a molecule in the train test. Spectra are encoded with DreAMS Molecules are encoded with ChemBERTa (Derify/ChemBERTa_augmented_pubchem_13m)

针对Spectraverse数据集的预计算嵌入向量(embeddings)。 数据集已按训练集与测试集划分,且测试集中的所有分子均不会与训练集中的分子共享相同的分子式。 光谱数据采用DreAMS进行编码。 分子数据采用ChemBERTa (Derify/ChemBERTa_augmented_pubchem_13m)进行编码。

提供机构:
Zenodo
创建时间:
2026-03-28
二维码
社区交流群
二维码
科研交流群
商业服务