遇见数据集

DNA-LLM/train_generated_seq

收藏
Hugging Face2024-12-06 更新2024-12-14 收录
官方服务:

资源简介:

该数据集包含多个特征字段,包括索引、序列、域、基础值、波长、第二波长、第二波长余弦值、振幅和振幅比率。数据集仅包含训练集,训练集包含168,000个样本,总大小为2,763,936,000字节。数据集的下载大小为293,916,957字节。

The dataset includes multiple feature fields such as index, sequence, domain, base, wavelength, second wavelength, second wavelength cosine, amplitude, and amplitude ratio. The dataset contains only a training set with 168,000 samples and a total size of 2,763,936,000 bytes. The download size of the dataset is 293,916,957 bytes.

提供机构:
DNA-LLM
二维码
社区交流群
二维码
科研交流群
商业服务