How quickly can an SRA submission proceed directly from S3
How long would 10TB of RNA-Seq data take to submit to NCBI SRA if it's already in AWS S3?
sra
• 1,661 views
•
link
written
by
Jeremy Leipzig
0 answers
No answers yet.
Log in to answer this question.
More posts like this
-
Converting SRA Data to FASTQ and Saving Directly to S3
written by Soohyun •I am currently downloading SRA data stored in AWS S3. I would like to download SRA data from S3 and convert it to FASTQ format, …
-
Fetch Fastq files directly for SRA data
written by Chintan •I'm trying to fetch fastq files for SRA data hosted by NCBI from AWS data exchange for a project, but the only way I can …
-
bioinformatics pipelines on GCP or AWS
written by BogdanDear all, would you please let me know if there are bioinformatics pipelines (RNA-seq, ChIP-seq, ATAC-seq, etc) already available on GCP or AWS ? it …
-
GEO submission when I have raw data in SRA
written by fifty_fifty •I am trying to submit my scRNA-seq data to GEO. [GEO submission guidelines][1] state that I should upload metadata, raw and processed data. And they …
-
How to download NCBI fastq data to local machine where access type is "Use Cloud Data Delivery"
written by daewowo •I have setup AWS and Google cloud accounts with billing. But both are overwhelmingly complicated for a simple user. What I would like to do …
-
RNA seq data analysis
written by arinataj1396 •I need films to teach me the steps of RNA seq data analysis I dont know how I can download the data from SRA bank …
-
NCBI SRA AWS AMI
written by noodleI came across the below NCBI website with instructions to access the SRA dataset through AWS, but it seems the AMI they reference no longer …
-
Tutorial: NCBI submissions: SRA and TSA
written by saraoppenheim •We have just published a very detailed guide to submitting RNA-Seq data to NCBI's SRA, and transcriptome assemblies to TSA. We have often felt the …
-
Forum: What server do you use?
written by caggtaagtatHi there, the cluster at my university often makes me wait for days until my jobs get from the queue to execution. I was therefore …
-
Which repository is more appropriate to submit for cancer cell line (bulk) RNA-seq data ?
written by artaI would like to submit our RNA-seq data (2 cell lines with triplicates) to repository to make it public. I know there are some options …
Have you already got in touch with SRA help desk, if you are truly planning to upload that much data? They can probably suggest efficient/non-standard way.
They should post here then. Isn't the point of biostars to disseminate this type of siloed knowledge?
Need to upload several TB of data is uncommon so while someone who has done this in the past may post you would save time by proactively contacting SRA help desk. While I see some NCBI folks answer questions on biostars I don't think I have ever seen anyone from SRA team here.
If you do get generally usable info please post that. My guess is solutions at this scale may be tailored for specific situations.
My primary concern here is how to upload the metadata and processed files matched the actual fastq files. SRA can use Aspera so for the fastq files it's probably just go and wait. This terrible metadata spreadsheet from GEO is already a pita for a few dozen samples. Leave alone hundreds or thousands.
I'm not sure Aspera is necessary for an S3-to-S3 transfer