Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Eight independent datasets of simulated paired-end reads for evaluating the accuracy of SARS-CoV-2 lineage abundance estimates from amplicon-based and whole genome-based sequencing. The data includes 200 bp and 400 bp reads from the Netherlands and Texas, with varying amplicon counts and coverages, and each dataset contains 20 sets of reads generated with different random seeds. Jasper van Bemmelen from Delft University of Technology created this dataset, which is available via paperswithcode under an Open Access license.
Source genomes require separate access via GISAID using the provided accession IDs.