Search | re3data.org

Filter

Reset all

Subjects

Content Types

Countries

AID systems

API

Data access

Data access restrictions

Database access

Database access restrictions

Database licenses

Data licenses

Data upload

Data upload restrictions

Enhanced publication

Institution responsibility type

Institution type

Keywords

Metadata standards

PID systems

Provider types

Quality management

Repository languages

Software

Syndications

Repository types

Versioning

Toogle short help

* at the end of a keyword allows wildcard searches
" quotes can be used for searching phrases
+ represents an AND search (default)
| represents an OR search
- represents a NOT operation
( and ) implies priority
~N after a word specifies the desired edit distance (fuzziness)
~N after a phrase specifies the desired slop amount

← Previous
1
2 (current)
3
4
Next →

Found 97 result(s)

eFish

Genomic Database Repository

Subject(s)

Content type(s)

Country

The project aims to examine and index the genomic diversity through the generation of complete mitochondrial and nuclear genome sequences of sharks and rays of the Pacific Rim. There is a huge diversity of elasmobranch fishes in this region, but many species are under threat because of poor management and conservation measures in many countries. It is absolutely critical that species’ identities are correct for conservation and fisheries management purposes. This project will provide this clarity of identity for both charismatic and commercially important species through the inclusion of ‘genetypes’ (ie., BioVouchers) and the application of genetic tools that utilize whole mitochondrial and nuclear genome sequences.

Sequence Read Archive

SRA

Subject(s)

Content type(s)

Country

United States

The Sequence Read Archive stores the raw sequencing data from such sequencing platforms as the Roche 454 GS System, the Illumina Genome Analyzer, the Applied Biosystems SOLiD System, the Helicos Heliscope, and the Complete Genomics. It archives the sequencing data associated with RNA-Seq, ChIP-Seq, Genomic and Transcriptomic assemblies, and 16S ribosomal RNA data.

SISSA Open Data

Scuola Internazionale Superiore di Studi Avanzati Open Data

Subject(s)

Content type(s)

Country

Italy

SISSA Open Data is the Sissa repository for the research data managment. It is an institutional repository that captures, stores, preserves, and redistributes the data of the SISSA scientific community in digital form. SISSA Open Data is managed by the SISSA Library as a service to the SISSA scientific community.

PiroplasmaDB

Subject(s)

Content type(s)

Country

A genome database for the genus Piroplasma. PiroplasmaDB is a member of pathogen-databases that are housed under the NIAID-funded EuPathDB Bioinformatics Resource Center (BRC) umbrella.

FlyReactome

a curated knowledgebase of drosophila melanogaster pathways

Subject(s)

Content type(s)

Country

<<<!!!<<< This repository is no longer available. This record is out dated >>>!!!>>>- eir fields, maintained by the FlyReactome staff.

Bioconductor

Open Source Software for Bioinformatics

Subject(s)

Content type(s)

Country

United States

Bioconductor provides tools for the analysis and comprehension of high-throughput genomic data. Bioconductor uses the R statistical programming language, and is open source and open development. It has two releases each year, and an active user community. Bioconductor is also available as an AMI (Amazon Machine Image) and a series of Docker images.

MicrosporidiaDB

Subject(s)

Content type(s)

Country

MicrosporidiaDB belongs to the EuPathDB family of databases and is an integrated genomic and functional genomic database for the phylum Microsporidia. In its first iteration (released in early 2010), MicrosporidiaDB contains the genomes of two Encephalitozoon species (see below). MicrosporidiaDB integrates whole genome sequence and annotation and will rapidly expand to include experimental data and environmental isolate sequences provided by community researchers. The database includes supplemental bioinformatics analyses and a web interface for data-mining.

NCBI Probe

Probe

Subject(s)

Content type(s)

Country

United States

<<<!!!<<< NCBI has retired the Probe Database >>>!!!>>>

TrichDB

Trichomonas vaginalis Sequence Database

Subject(s)

Content type(s)

Raw data

Country

TrichDB integrated genomic resources for the eukaryotic protist pathogens Trichomonas vaginalis.

AspGD

Aspergillus Genome Database

Subject(s)

Content type(s)

Country

United States

>>>>!!!!<<<< AspGD data are being integrated into FungiDB. Please click here for additional details http://fungidb.org/ . Discussion of how to maximize the value of FungiDB for the Aspergillus research community will be a major topic at the upcoming AsperFest12 meeting at Asilomar (March 16-17, 2015). >>>>!!!!<<<< AspGD is an organized collection of genetic and molecular biological information about the filamentous fungi of the genus Aspergillus. Among its many species, the genus contains an excellent model organism (A. nidulans, or its teleomorph Emericella nidulans), an important pathogen of the immunocompromised (A. fumigatus), an agriculturally important toxin producer (A. flavus), and two species used in industrial processes (A. niger and A. oryzae). AspGD contains information about genes and proteins of multiple Aspergillus species; descriptions and classifications of their biological roles, molecular functions, and subcellular localizations; gene, protein, and chromosome sequence information; tools for analysis and comparison of sequences; and links to literature information; as well as a multispecies comparative genomics browser tool (Sybil) for exploration of orthology and synteny across multiple sequenced Aspergillus species.

IMGT/GENE-DB

ImMunoGeneTics / Gene-DB

Subject(s)

Content type(s)

Country

IMGT/GENE-DB is the IMGT genome database for IG and TR genes from human, mouse and other vertebrates. IMGT/GENE-DB provides a full characterization of the genes and of their alleles: IMGT gene name and definition, chromosomal localization, number of alleles, and for each allele, the IMGT allele functionality, and the IMGT reference sequences and other sequences from the literature. IMGT/GENE-DB allele reference sequences are available in FASTA format (nucleotide and amino acid sequences with IMGT gaps according to the IMGT unique numbering, or without gaps).

National Omics Data Encyclopedia

NODE

Subject(s)

Content type(s)

Country

China

NODE (The National Omics Data Encyclopedia) provides an integrated, compatible, comparable, and scalable multi-omics resource platform that supports flexible data management and effective data release. NODE uses a hierarchical data architecture to support storage of muti-omics data including sequencing data, MS based proteomics data, MS or NMR based metabolomics data, and fluorescence imaging data. Launched in early 2017, NODE has collected and published over 900 terabytes of omics data for researchers from China and all over the world in last three years, 22% of which contains multiple omics data. NODE provides functions around the whole life cycle of omics data, from data archive, data requests/responses to data sharing, data analysis, data review and publish.

OrthoMCL

Ortholog Groups of Protein Sequences

Subject(s)

Content type(s)

Country

United States

OrthoMCL is a genome-scale algorithm for grouping orthologous protein sequences. It provides not only groups shared by two or more species/genomes, but also groups representing species-specific gene expansion families. So it serves as an important utility for automated eukaryotic genome annotation. OrthoMCL starts with reciprocal best hits within each genome as potential in-paralog/recent paralog pairs and reciprocal best hits across any two genomes as potential ortholog pairs. Related proteins are interlinked in a similarity graph. Then MCL (Markov Clustering algorithm,Van Dongen 2000; www.micans.org/mcl) is invoked to split mega-clusters. This process is analogous to the manual review in COG construction. MCL clustering is based on weights between each pair of proteins, so to correct for differences in evolutionary distance the weights are normalized before running MCL.

STRING

Known and Predicted Protein-Protein Interactions

Subject(s)

Content type(s)

Country

STRING is a database of known and predicted protein interactions. The interactions include direct (physical) and indirect (functional) associations; they are derived from four sources: - Genomic Context - High-throughput Experiments - (Conserved) Coexpression - Previous Knowledge STRING quantitatively integrates interaction data from these sources for a large number of organisms, and transfers information between these organisms where applicable.

FungiDB

The Fungal and Oomycete Genomics Resource

Subject(s)

Content type(s)

Country

United States

FungiDB belongs to the EuPathDB family of databases and is an integrated genomic and functional genomic database for the kingdom Fungi. FungiDB was first released in early 2011 as a collaborative project between EuPathDB and the group of Jason Stajich (University of California, Riverside). At the end of 2015, FungiDB was integrated into the EuPathDB bioinformatic resource center. FungiDB integrates whole genome sequence and annotation and also includes experimental and environmental isolate sequence data. The database includes comparative genomics, analysis of gene expression, and supplemental bioinformatics analyses and a web interface for data-mining.

Genome-Scale Metabolic Network DataBase

GSMNDB

Subject(s)

Content type(s)

Country

China

<<<!!!<<< 2019-12-23: the repository is offline >>>!!!>>> Introduction of genome-scale metabolic network: The completion of genome sequencing and subsequent functional annotation for a great number of species enables the reconstruction of genome-scale metabolic networks. These networks, together with in silico network analysis methods such as the constraint based methods (CBM) and graph theory methods, can provide us systems level understanding of cellular metabolism. Further more, they can be applied to many predictions of real biological application such as: gene essentiality analysis, drug target discovery and metabolic engineering

Joint Genome Institute Data Portal

JGI Data Portal

Subject(s)

Content type(s)

Country

United States

The U.S. Department of Energy (DOE) Joint Genome Institute (JGI) is a DOE Office of Science User Facility located at Lawrence Berkeley National Laboratory (Berkeley Lab). All data generated by the DOE Joint Genome Institute is available through this repository once the data are published or public.

Genomes OnLine Database

GOLD

Subject(s)

Content type(s)

Country

United States

GOLD is currently the largest repository for genome project information world-wide. The accurate and efficient genome project tracking is a vital criterion for launching new genome sequencing projects, and for avoiding significant overlap between various sequencing efforts and centers.

CanGEM

Cancer GEnome Mine

Subject(s)

Content type(s)

Country

Finland

>>>!!!<<<As stated 2017-05-23 Cancer GEnome Mine is no longer available >>>!!!<<< Cancer GEnome Mine is a public database for storing clinical information about tumor samples and microarray data, with emphasis on array comparative genomic hybridization (aCGH) and data mining of gene copy number changes.

GEISHA

Gallus Expression in Situ Hybridization Analysis

Subject(s)

Content type(s)

Country

United States

GEISHA is the online repository of in situ hybridization and corresponding metadata for genes expressed in the chicken embryo during the first six days of development.

Clinical Proteomic Tumor Analysis Consortium Data Portal

CPTAC Data Portal

Subject(s)

Content type(s)

Country

United States

The CPTAC Data Portal is the centralized repository for the dissemination of proteomic data collected by the Proteome Characterization Centers (PCCs) for the CPTAC program. The portal also hosts analyses of the mass spectrometry data (mapping of spectra to peptide sequences and protein identification) from the PCCs and from a CPTAC-sponsored common data analysis pipeline (CDAP).

arrayMap

Subject(s)

Content type(s)

Country

Switzerland

arrayMap is a repository of cancer genome profiling data. Original) from primary repositories (e.g. NCBI GEO, EBI ArrayExpress, TCGA) is re-processed and annotated for metadata. Unique visualization of the processed data allows critical evaluation of data quality and genome information. Structured metadata provides easy access to summary statistics, with a focus on copy number aberrations in cancer entities.

UCLanData

Subject(s)

Content type(s)

Country

United Kingdom

Repository of research data sets created under the auspices of the University of Central Lancashire

EcoGene

Escherichia coli strain K12 genome database

Subject(s)

Content type(s)

Country

United States

<<<!!!<<< The repository is no longer available. 2019-12-02: no more access to EcoGene >>>!!!<<<

The Cell Map

TheCellMap

Subject(s)

Content type(s)

Country

TheCellMap.org serves as a central repository for storing and analyzing quantitative genetic interaction data produced by genome-scale Synthetic Genetic Array (SGA) experiments with the budding yeast Saccharomyces cerevisiae. In particular, TheCellMap.org allows users to easily access, visualize, explore, and functionally annotate genetic interactions, or to extract and reorganize subnetworks, using data-driven network layouts in an intuitive and interactive manner.

← Previous
1
2 (current)
3
4
Next →

Current projects
EOSC FAIR-IMPACT

re3data COREF

To the extent possible under law, re3data.org has waived all copyright and related or neighboring rights to the database entries of re3data.org.
Except where otherwise noted, content on this site is licensed under a Creative Commons Attribution 4.0 International License .
Cite this service: re3data.org - Registry of Research Data Repositories. https://doi.org/10.17616/R3D last accessed: 2024-04-25