ddidacus/nuevamol-chembl-eval-set
收藏资源简介:
该数据集是一个包含化学结构数据的训练集,主要用于化学或生物信息学任务。数据集中包含三个字段:smiles(表示化学结构的字符串)、score(表示评分或属性的浮点数)和task(表示任务类型的字符串)。数据集共有1700个样本,所有数据均用于训练分割,总大小约为132.9 KB,下载大小约为40.1 KB。数据文件以train-*格式存储,适用于机器学习和模型训练应用。
This dataset is a training set containing chemical structure data, primarily used for chemistry or bioinformatics tasks. It includes three fields: smiles (a string representing chemical structures), score (a float representing scores or properties), and task (a string indicating the task type). The dataset consists of 1700 samples, all allocated to the train split, with a total size of approximately 132.9 KB and a download size of approximately 40.1 KB. The data files are stored in the train-* format and are suitable for machine learning and model training applications.




