遇见数据集

Dataset for drevalpy

收藏
Zenodo2025-05-28 更新2026-05-26 收录
官方服务:

资源简介:

All preprocessing details can be found in https://github.com/JudithBernett/preprocess_drp_data. News: Drug IDs and names have been harmonized over datasets: Every drug is now identified by its pubchem ID. If it had no pubchem ID, the drug name is in the respective field. Fingerprints and DIPK MolGNet features have now been generated for all drugs with SMILES. They are filtered for each dataset such that the respective folder only contains features for drugs contained in the dataset. All drugs from CCLE, CTRPv1, and CTRPv2 had SMILES. For GDSC1 and GDSC2, there are some missing values. The cell line ID is now always called cellosaurus_id, the drug ID pubchem_id The drug features are now identified over the pubchem_id (column names in fingerprint files, suffix in MolGNet files) The old response files have been removed, now, for each dataset, the response file is [dataset].csv which also contains the curve-curated measurements Unnamed: 0 column removed from copy number variation

提供机构:
Zenodo
创建时间:
2025-02-20
二维码
社区交流群
二维码
科研交流群
商业服务