# s1-vis-mid-resize Original dataset structure preserved, filtered by token length and image quality ## Dataset Description This dataset was processed using the [data-preproc](https://github.com/o
This python script was used to split one file with 1000 simulated datasets into 1000 files with one dataset each. This script was used after simulating data in ms and seqgen and before summary statist