Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Long-read proteogenomics integrates matched long-read RNA sequencing and mass spectrometry proteomics to characterize protein isoforms. The dataset contains the complete output from executing the open-source Nextflow workflow on Jurkat cell samples, including a 6.5 GB BAM file of full-length non-concatemer reads. This approach enables the discovery of novel protein isoforms and provides a classification scheme and inference algorithm for isoforms intractable to mass spectrometry alone.
The primary data file (jurkat.flnc.bam) was split into 13 parts for distribution and requires rejoining before use, as described in the source instructions. License is listed as Open Access (green).