遇见数据集

DNA / Aptamer dataset

收藏
DataCite Commons2025-05-01 更新2025-05-17 收录
官方服务:

资源简介:

The data for the DNA Aptamers was web-scraped from the online database Aptagen (https://www.aptagen.com/) using an in-house python script. After cleaning the data and removing modified bases and duplicates the Dataframe included a 238 unique aptamer sequences. The data for the DNA sequences was web-scraped from the online database (http://ndbserver.rutgers.edu) using an in-house python script, After downloading the data into a dataframe, cleaning and removing modified bases and duplicates the Dataframe included a 4885 of unique sequences.

本数据集所使用的DNA适配子(DNA Aptamers)数据,通过自研Python脚本从在线数据库Aptagen(https://www.aptagen.com/)网页爬取获得。经数据清洗、移除修饰碱基与重复序列后,该数据框(DataFrame)共包含238条唯一适配子序列。本次数据集的DNA序列数据,则通过自研Python脚本从在线数据库(http://ndbserver.rutgers.edu)抓取获取。将数据下载至数据框后,经清洗、移除修饰碱基与重复序列步骤,该数据框共包含4885条唯一DNA序列。

提供机构:
Mendeley
创建时间:
2020-08-13
搜集汇总
背景与挑战
背景概述
该数据集整合了两个来源的DNA序列数据:DNA Aptamers部分包含238条独特序列,从Aptagen网站爬取;DNA序列部分包含4885条独特序列,从另一个在线数据库爬取。所有数据均通过内部脚本获取,并经过清洗、去除修饰碱基和重复项的处理,确保了数据的独特性和质量,适用于生物信息学分析或分子设计研究。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务