Home - SRA - NCBI
NCBI SRA provides sequencing-study records and raw data.
Search the literature, test an assumption, plan an archive and check a file. Calculations and file checks run in your browser.
Build a transparent query and open the original research indexes.
Search results open at the provider. They have not been screened by Death X. Dates apply to publications, not our review date.
Estimate the number of nucleotides needed to encode a digital file under explicit assumptions.
Total nucleotides across the specified copies
Nucleotides = MB × 1,000,000 × 8 ÷ effective bits per nucleotide × copies. Round up to a whole nucleotide per copy before multiplying by copies. The example rate is an assumption that includes coding overhead. This is not a cost, durability or biological-memory estimate. Read the storage research
Estimate storage for photos, recordings and other files. Adjust the assumptions to match your actual media.
Decimal GB/TB. Excludes version history, transcoding, indexes and future growth. Three copies here is an editable example, not a preservation guarantee.
A practical checklist inspired by digital-preservation and biobanking guidance.
Saved only in this browser. A planning aid, not certification. NDSA guidance · NCI governance
Calculate a SHA-256 checksum to record a file’s exact bytes. The file stays on this device.
Recalculate later and compare with your trusted original checksum. A match verifies byte equality, not provenance, authorship or scientific validity. Read about integrity manifests
Explore nine claims across the Archive, Culture, Grove, Reef and memory research. Each assessment links to its evidence and states what remains unresolved.
Sequence data, recorded experience and autobiographical memory are different forms of information. Evidence for one does not establish preservation of the others.
Repositories, software and APIs maintained by their original providers.
NCBI SRA provides sequencing-study records and raw data.
dbGaP links genotype and phenotype studies with controlled-access research data.
EMBL-EBI introduces the European Nucleotide Archive's sequence holdings.
GA4GH refget supports unambiguous reference-sequence identifiers and retrieval.
ENA documents programmatic sequence-record retrieval.
GA4GH catalogs genomic data, security and interoperability standards.
GA4GH Starter Kit demonstrates genomic data-service implementations.
NCBI BLAST compares biological sequences with reference databases.
Galaxy supports web-based, reproducible biomedical analyses.
BioStudies links data and supporting materials for biological studies.
NCBI Datasets packages genomes, annotations and metadata.
Ensembl supplies genome annotation and comparative-genomics resources.
UCSC supports genome visualization, track queries and coordinate conversion.
MetaboLights holds metabolomics studies, raw measurements and metadata.
PRIDE archives mass-spectrometry evidence for proteins and peptides.
PRIDE documents project, file and metadata queries.
QIIME 2 provides reproducible microbiome and multi-omics analysis.
Bioconductor supplies open-source biological data-analysis tools.
Europe PMC documents publication and preprint search.
Crossref exposes bibliographic metadata, identifiers and publication updates.
OpenAlex supports discovery across works, authors, institutions and topics.