遇见数据集

DisoNuc: Disordered Nucleic Acid Regions Database

收藏
Zenodo2026-08-05 更新2026-08-13 收录
官方服务:

资源简介:

The format of the CSV download file is as follows: ID: Starts with "DN", followed by 5-digits number, e.g., DN00001. Length: Disordered DNA/RNA sequence length. Sequence: Disordered DNA/RNA sequence. Type: Type of nucleic acid, i.e., DNA or RNA. PDB Sources: The corresponding PDB sources of each disordered DNA/RNA sequence. Number of PDBs: The number of corresponding PDB sources for each disordered DNA/RNA sequence. The format of the JSON download file is as follows: ID/Region ID: Starts with "DN", followed by 5-digits number, e.g., DN00001. Length: Disordered DNA/RNA sequence length. Sequence: Disordered DNA/RNA sequence. Type: Type of nucleic acid, i.e., DNA or RNA. PDB Sources: The corresponding PDB sources of each disordered DNA/RNA sequence. PDB Chain: PDB and chain ID. Start: Start position of the disordered region in the PDB chain. End: End position of the disordered region in the PDB chain. Synthetic: The sample used in the experiment was synthetically produced or not. Organism: The organism name. Taxon ID: The NCBI taxonomy ID. Strain: The particular strain of the organism used. ATCC: The strain ID from American Type Culture Collection. Number of PDBs: The number of corresponding PDB sources for each disordered DNA/RNA sequence. Nucleotide Composition: Nucleotide composition. A: Number of Adenine (A). T: Number of Thymine (T). U: Number of Uracil (U). G: Number of Guanine (G). C: Number of Cytosine (C). Binding Induced Structure Informaion: List of all homologous occurrences of this disordered region, identified from a structured PDB sequence search, together with their binding interactions. Region ID: Starts with "DN", followed by 5-digits number, e.g., DN00001. PDB Chain with Interaction Data: PDB and chain ID of the homologous occurrences of this disordered region, identified from a structured PDB sequence search. Chain Length: Length of the PDB chain. Region Position Start: Start position of the region in the PDB chain. Region Position End: End position of the region in the PDB chain. Number of Binding Nucleobases: The number of binding nucleobases on the region in the PDB chain. Binding Coverage: The fraction or percentage of actual binding nucleobases on the region in the PDB chain. Binding Annotations: Binding annotation information for the region in the PDB chain. Region Position: Each position of the region in the PDB chain. Nucleobase: Each nucleobase of the region in the PDB chain. Binding or not: 1 for binding nucleobase; 0 for non-binding nucleobase. Binding Molecules: All molecules that bind to each region position, formatted as: <chain ID>_<molecules type>_<position in chain>.

提供机构:
Zenodo
创建时间:
2026-08-05
二维码
社区交流群
二维码
科研交流群
商业服务