Initialization files and data supporting https://arxiv.org/abs/1506.09008
收藏资源简介:
###Preamble### This upload contains initialization files and data for simulations reported in:<br> https://arxiv.org/abs/1506.09008: Coarse-grained modelling of strong DNA bending II: Cyclization The initialization files allow a user to repeat the reported simulations using the oxDNA model. oxDNA is available for download from:<br> https://dna.physics.ox.ac.uk/index.php/Main_Page. <br> The use and meaning of the input and output files are documented extensively on this wiki. <br> ###Organisation### A pdf copy of the main text and supplementary material of the relevant paper are provided as main.pdf and SI.pdf in the head directory. Simulations are organised by system type within subdirectories. ################# The "Basic cyclization" folder contains the files for sequence-independent simulations of cyclization, for varying duplex and single-stranded overhang lengths. Folder DXXCYY corresponds to simulation of a cyclization system with Nd = XX and Nbp=YY, with the meaning of these symbols given in the paper. These simulations underlie:<br> - The black data points in Fig. 3, 7 and S4 (alongside data from the simulations in the "Dimerization" folder), <br> - The data in Fig. 4 and S5.<br> - The data points connected by solid lines in Fig. 5.<br> - The black data points in Fig. 6.<br> - The black data points in Fig S2(a), and the data in Fig. S2(b)<br> - The black data points in Fig. S3(a) ################# The "Dimerization" folder contains the files for the simulations of dimerization, at the two reference concentrations 336nM and 2690nM. For simulations at 336nM, the folder bimolecularXX contains files for the sequence-independent simulation with XX = Nd1+Ns, and bimolecularXX_seq contains the sequence-dependent variants. For simulations at 2690nM, the folder bimolecularXX contains files for the sequence-independent simulation with XX = Nd1+Nd2+Ns, and bimolecularXX_seq contains the sequence-dependent variants. Note that the folders bimolecular73 and bimolecular101_seq were accidentally deleted, and the output data is missing. Input files for bimolecular73 have been recreated (the bimolecular101_seq has not been recreated because these simulations were of very limited importance for the paper). These simulations provide:<br> - The data points in Fig. 3, 7 and S4 (alongside data from the simulations in the "Basic cyclization" folder and the "Perturbations" folder), <br> - The data points connected by dashed lines in Fig. 5.<br> - The grey data points in Fig. 6.<br> - The data in Table S2. ################# The "Perturbations" folder contains the files for cyclization simulations that do not correspond to sequence-averaged, defect-free systems.<br> - Folders DXXCYY_seq and DXXCYY_* contain files relevant to simulations of cyclization with Nd = XX and Nbp=YY, incorporating sequence-dependence. These simulations underlie the brown/blue/mauve data points in Fig. 3, 7 and S2b (alongside data from the simulations in the "Dimerization" folder).<br> - Folder D5969mm contains files relevant to the simulation of a mismatch containing-sytem, with data reported on Table S3 and Figure S3 (intact simulations can be found in the "Basic cyclization" folder). <br> - Folders D87C97, D87C97n1, D87C97n2 and D87C97nn contains files relevant to the comparison of a system with no nicks, a nick in position 1, a nick in position 2, and a double nick, respectively. Data are reported in Table S3 and Fig. S3(b). <br> ###Content### For each system, a "closed1" and an "open1" folder are present. These correspond to the two windows of umbrella sampling that were performed separately. Within each folder are the necessary initialization files to run the simulations exactly as reported in the paper, simply by calling oxDNA from within the folder, using "inputVMMC" as the input file. Also included are output files for a single realisation of the simulation. The meaning of these files are outlined at https://dna.physics.ox.ac.uk/index.php/Main_Page. Note that the results in the paper were all obtained from 5 independent replicas, using different initial conditions and different seeds. These can be (statistically) recreated simply by drawing random starting configurations from the single available traj_hist file. <br> ###A note on topology### Many of these simulations were performed with "unique" topology, which prevents non-native base pairing. In the topology files, instead of indicating the base type with a letter (A, C, G or T) in the second column, an integer 0 < n or n > 10 is used instead.<br> - If n modulo 4 =0, the base is treated as possessing the interaction strengths of A but will only bind to a base with a type m = 3-n.<br> - If n modulo 4 =1, the base is treated as possessing the interaction strengths of C but will only bind to a base with a type m = 3-n.<br> - If n modulo 4 =2, the base is treated as possessing the interaction strengths of G but will only bind to a base with a type m = 3-n.<br> - If n modulo 4 =3, the base is treated as possessing the interaction strengths of T but will only bind to a base with a type m = 3-n. In addition, please note that the topology files for the dimerization simulations at 336nM were set out slightly strangely, in that base IDs are not assigned contiguously to contiguous sequences of bases in a strand at some points. Nonetheless, the connectivity specified by these topology files is correct. <br>
### 前言 ### 本上传文件包含以下报道的模拟所用的初始化文件与数据:https://arxiv.org/abs/1506.09008:《强DNA弯曲的粗粒化建模II:环化》 用户可使用oxDNA模型(oxDNA)复现本报道中的模拟,该模型的下载地址为:https://dna.physics.ox.ac.uk/index.php/Main_Page。该维基页面已对输入、输出文件的用途与含义进行了详尽说明。 ### 组织结构 ### 主目录下提供了相关论文的正文与补充材料的PDF副本,分别命名为main.pdf和SI.pdf。模拟按系统类型划分至各个子目录中。 ################# "Basic cyclization(基础环化)"文件夹包含了序列无关的环化模拟文件,用于研究不同双链与单链悬垂长度。文件夹DXXCYY对应Nd=XX、Nbp=YY的环化系统模拟,相关符号的含义详见论文。这些模拟支撑了如下结果: - 图3、图7与图S4中的黑色数据点(与"Dimerization(二聚化)"文件夹中的模拟数据共同组成) - 图4与图S5中的数据 - 图5中以实线连接的数据点 - 图6中的黑色数据点 - 图S2(a)中的黑色数据点与图S2(b)中的数据 - 图S3(a)中的黑色数据点 ################# "Dimerization(二聚化)"文件夹包含了两种参考浓度(336nM与2690nM)下的二聚化模拟文件。对于336nM浓度下的模拟,文件夹bimolecularXX对应XX=Nd1+Ns的序列无关模拟,bimolecularXX_seq则对应序列相关的变体。对于2690nM浓度下的模拟,文件夹bimolecularXX对应XX=Nd1+Nd2+Ns的序列无关模拟,bimolecularXX_seq则对应序列相关的变体。请注意,文件夹bimolecular73与bimolecular101_seq被意外删除,导致其输出数据丢失。目前已重新生成bimolecular73的输入文件,但由于bimolecular101_seq的模拟对论文的价值有限,故未对其进行重建。 这些模拟提供了如下结果: - 图3、图7与图S4中的数据点(与"Basic cyclization"及"Perturbations(扰动)"文件夹中的模拟数据共同组成) - 图5中以虚线连接的数据点 - 图6中的灰色数据点 - 表S2中的数据 ################# "Perturbations(扰动)"文件夹包含了非序列平均、无缺陷系统的环化模拟文件。 - 文件夹DXXCYY_seq与DXXCYY_*包含了Nd=XX、Nbp=YY的环化模拟相关文件,该模拟纳入了序列相关性。这些模拟支撑了图3、图7与图S2(b)中的棕/蓝/紫数据点(与"Dimerization"文件夹中的模拟数据共同组成)。 - 文件夹D5969mm包含了含错配系统的模拟相关文件,相关数据见表S3与图S3(完整模拟可在"Basic cyclization"文件夹中找到)。 - 文件夹D87C97、D87C97n1、D87C97n2与D87C97nn分别对应无切口、第1位置有切口、第2位置有切口与双切口的系统对比。相关数据见表S3与图S3(b)。 ### 内容 ### 针对每个系统,均设有"closed1"与"open1"文件夹,二者分别对应单独执行的两次伞形采样窗口。每个文件夹内均包含复现论文中所述模拟所需的全部初始化文件,只需在该文件夹内调用oxDNA,并以"inputVMMC"作为输入文件即可运行。此外还包含了单次模拟实现的输出文件。这些文件的含义详见https://dna.physics.ox.ac.uk/index.php/Main_Page。 请注意,论文中的所有结果均来自5组独立复现,采用了不同的初始条件与随机种子。只需从现有的单个traj_hist文件中抽取随机初始构型,即可(从统计意义上)复现这些结果。 ### 关于拓扑的说明 ### 本数据集的多数模拟采用了"unique"拓扑,该拓扑可防止非天然碱基配对。在拓扑文件中,第二列未使用字母(A、C、G或T)表示碱基类型,而是使用整数0<n或n>10。规则如下: - 若n mod 4 = 0,则该碱基被视为具有A的相互作用强度,但仅能与类型为m=3-n的碱基结合 - 若n mod 4 = 1,则该碱基被视为具有C的相互作用强度,但仅能与类型为m=3-n的碱基结合 - 若n mod 4 = 2,则该碱基被视为具有G的相互作用强度,但仅能与类型为m=3-n的碱基结合 - 若n mod 4 = 3,则该碱基被视为具有T的相互作用强度,但仅能与类型为m=3-n的碱基结合 此外请注意,336nM浓度下二聚化模拟的拓扑文件格式略有异常:在部分位置,碱基ID并未按照碱基链的连续序列进行连续分配。尽管如此,这些拓扑文件所指定的碱基连接关系仍是正确的。



