Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://exac.broadinstitute.org/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 9, 2023. An aggregated data platform for genome sequencing data created by a coalition of investigators seeking to aggregate and harmonize exome sequencing data from a variety of large-scale sequencing projects, and to make summary data available for the wider scientific community. The data set provided on this website spans 61,486 unrelated individuals sequenced as part of various disease-specific and population genetic studies. They have removed individuals affected by severe pediatric disease, so this data set should serve as a useful reference set of allele frequencies for severe disease studies. All of the raw data from these projects have been reprocessed through the same pipeline, and jointly variant-called to increase consistency across projects. They ask that you not publish global (genome-wide) analyses of these data until after the ExAC flagship paper has been published, estimated to be in early 2015. If you''re uncertain which category your analyses fall into, please email them. The aggregation and release of summary data from the exomes collected by the Exome Aggregation Consortium has been approved by the Partners IRB (protocol 2013P001477, Genomic approaches to gene discovery in rare neuromuscular diseases).
Proper citation: ExAc (RRID:SCR_004068) Copy
A curated database that provides comprehensive integrated biological information for Saccharomyces cerevisiae along with search and analysis tools to explore these data. SGD allows researchers to discover functional relationships between sequence and gene products in fungi and higher organisms. The SGD also maintains the S. cerevisiae Gene Name Registry, a complete list of all gene names used in S. cerevisiae which includes a set of general guidelines to gene naming. Protein Page provides basic protein information calculated from the predicted sequence and contains links to a variety of secondary structure and tertiary structure resources. Yeast Biochemical Pathways allows users to view and search for biochemical reactions and pathways that occur in S. cerevisiae as well as map expression data onto the biochemical pathways. Literature citations are provided where available.
Proper citation: SGD (RRID:SCR_004694) Copy
A database of protein families, each represented by multiple sequence alignments and hidden Markov models (HMMs). Users can analyze protein sequences for Pfam matches, view Pfam family annotation and alignments, see groups of related families, look at the domain organization of a protein sequence, find the domains on a PDB structure, and query Pfam by keywords. There are two components to Pfam: Pfam-A and Pfam-B. Pfam-A entries are high quality, manually curated families that may automatically generate a supplement using the ADDA database. These automatically generated entries are called Pfam-B. Although of lower quality, Pfam-B families can be useful for identifying functionally conserved regions when no Pfam-A entries are found. Pfam also generates higher-level groupings of related families, known as clans (collections of Pfam-A entries which are related by similarity of sequence, structure or profile-HMM).
Proper citation: Pfam (RRID:SCR_004726) Copy
https://sites.google.com/site/jpopgen/dbNSFP
A database for functional prediction and annotation of all potential non-synonymous single-nucleotide variants (nsSNVs) in the human genome. Version 2.0 is based on the Gencode release 9 / Ensembl version 64 and includes a total of 87,347,043 nsSNVs and 2,270,742 essential splice site SNVs. It compiles prediction scores from six prediction algorithms (SIFT, Polyphen2, LRT, MutationTaster, MutationAssessor and FATHMM), three conservation scores (PhyloP, GERP++ and SiPhy) and other related information including allele frequencies observed in the 1000 Genomes Project phase 1 data and the NHLBI Exome Sequencing Project, various gene IDs from different databases, functional descriptions of genes, gene expression and gene interaction information, etc. Some dbNSFP contents (may not be up-to-date though) can also be accessed through variant tools, ANNOVAR, KGGSeq, UCSC Genome Browser''s Variant Annotation Integrator, Ensembl Variant Effect Predictor and HGMD.
Proper citation: dbNSFP (RRID:SCR_005178) Copy
A clade oriented, community curated database containing genomic, genetic, phenotypic and taxonomic information for plant genomes. Genomic information is presented in a comparative format and tied to important plant model species such as Arabidopsis. SGN provides tools such as: BLAST searches, the SolCyc biochemical pathways database, a CAPS experiment designer, an intron detection tool, an advanced Alignment Analyzer, and a browser for phylogenetic trees. The SGN code and database are developed as an open source project, and is based on database schemas developed by the GMOD project and SGN-specific extensions.
Proper citation: SGN (RRID:SCR_004933) Copy
Database of the international consortium working together to mutate all protein-coding genes in the mouse using a combination of gene trapping and gene targeting in C57BL/6 mouse embryonic stem (ES) cells. Detailed information on targeted genes is available. The IKMC includes the following programs: * Knockout Mouse Project (KOMP) (USA) ** CSD, a collaborative team at the Children''''s Hospital Oakland Research Institute (CHORI), the Wellcome Trust Sanger Institute and the University of California at Davis School of Veterinary Medicine , led by Pieter deJong, Ph.D., CHORI, along with K. C. Kent Lloyd, D.V.M., Ph.D., UC Davis; and Allan Bradley, Ph.D. FRS, and William Skarnes, Ph.D., at the Wellcome Trust Sanger Institute. ** Regeneron, a team at the VelociGene division of Regeneron Pharmaceuticals, Inc., led by David Valenzuela, Ph.D. and George D. Yancopoulos, M.D., Ph.D. * European Conditional Mouse Mutagenesis Program (EUCOMM) (Europe) * North American Conditional Mouse Mutagenesis Project (NorCOMM) (Canada) * Texas A&M Institute for Genomic Medicine (TIGM) (USA) Products (vectors, mice, ES cell lines) may be ordered from the above programs.
Proper citation: International Knockout Mouse Consortium (RRID:SCR_005574) Copy
http://swissregulon.unibas.ch/fcgi/sr/swissregulon
A database of genome-wide annotations of regulatory sites. The predictions are based on Bayesian probabilistic analysis of a combination of input information including: * Experimentally determined binding sites reported in the literature. * Known sequence-specificities of transcription factors. * ChIP-chip and ChIP-seq data. * Alignments of orthologous non-coding regions. Predictions were made using the PhyloGibbs, MotEvo, IRUS and ISMARA algorithms developed in their group, depending on the data available for each organism. Annotations can be viewed in a Gbrowse genome browser and can also be downloaded in flat file format.
Proper citation: SwissRegulon (RRID:SCR_005333) Copy
The TIGR database is a collection of plant transcript sequences. Transcript assemblies are searchable using BLAST and accession number. The construction of plant transcript assemblies (TAs) is similar to the TIGR gene indices. The sequences that are used to build the plant TAs are expressed transcripts collected from dbEST (ESTs) and the NCBI GenBank nucleotide database (full length and partial cDNAs). "Virtual" transcript sequences derived from whole genome annotation projects are not included. All plant species for which more than 1,000 ESTs or cDNA sequences are available are included in this project. TAs are clustered and assembled using the TGICL tool (Pertea et al., 2003), Megablast (Zhang et al., 2000) and the CAP3 assembler (Huang and Madan, 1999). TGICL is a wrapper script which invokes Megablast and CAP3. Sequences are initially clustered based on an all-against-all comparisons using Megablast. The initial clusters are assembled to generate consensus sequences using CAP3. Assembly criteria include a 50 bp minimum match, 95% minimum identity in the overlap region and 20 bp maximum unmatched overhangs. Any EST/cDNA sequences that are not assembled into TAs are included as singletons. All singletons retain their GenBank accession numbers as identifiers. Plant TA identifiers are of the form TAnumber_taxonID, where number is a unique numerical identifier of the transcript assembly and taxonID represents the NCBI taxon id. In order to provide annotation for the TAs, each TA/singleton was aligned to the UniProt Uniref database. For release 1 TAs, a masked version of the Uniref90 database was used. For release 2 and onwards, a masked version of the UniRef100 database is used. Alignments were required to have at least 20% identity and 20% coverage. The annotation for the protein with the best alignment to each TA or singleton was used as the annotation for that sequence. Additionally, the relative orientation of each TA/singleton to the best matching protein sequence was used to determine the orientation of each TA/singleton. Some sequences did not have alignments to the protein database that met our quality criteria, and those sequences have neither annotation nor orientation assignments. The release number for the plant TAs refers to the release version for a particular species. For the initial build, all TA sets are of version 1. Subsequent TA updates for new releases will be carried out when the percentage increase of the EST and cDNA counts exceeds 10% of the previous release and when the increase contains more than 1,000 new sequences. New releases will also include additional plant species with more than 1,000 EST or cDNA sequences that have become publicly available.
Proper citation: TIGR Plant Transcript Assembly database (RRID:SCR_005470) Copy
A next-generation web-based application that aims to provide an integrated solution for both visualization and analysis of deep-sequencing data, along with simple access to public datasets.
Proper citation: Systems Transcriptional Activity Reconstruction (RRID:SCR_005622) Copy
http://supfam.org/SUPERFAMILY/
SUPERFAMILY is a database of structural and functional protein annotations for all completely sequenced organisms. The SUPERFAMILY annotation is based on a collection of hidden Markov models, which represent structural protein domains at the SCOP superfamily level. A superfamily groups together domains which have an evolutionary relationship. The annotation is produced by scanning protein sequences from over 1,700 completely sequenced genomes against the hidden Markov models.
Proper citation: SUPERFAMILY (RRID:SCR_007952) Copy
Database to explore known and predicted interactions of chemicals and proteins. It integrates information about interactions from metabolic pathways, crystal structures, binding experiments and drug-target relationships. Inferred information from phenotypic effects, text mining and chemical structure similarity is used to predict relations between chemicals. STITCH further allows exploring the network of chemical relations, also in the context of associated binding proteins. Each proposed interaction can be traced back to the original data sources. The database contains interaction information for over 68,000 different chemicals, including 2200 drugs, and connects them to 1.5 million genes across 373 genomes and their interactions contained in the STRING database.
Proper citation: Search Tool for Interactions of Chemicals (RRID:SCR_007947) Copy
http://www.bioinfodatabase.com/pint/
A protein-protein interactions thermodynamic database which contains data of several thermodynamic parameters along with sequence and structural information experimental conditions and literature information. Each entry contains numerical data for features of the interacting proteins such as the free energy change, dissociation constant, association constant, enthalpy change, and heat capacity change. PINT includes: the name and source of the proteins involved in binding, SWISS-PROT and Protein Data Bank (PDB) codes, secondary structure and solvent accessibility of residues at mutant positions, measuring methods, and experimental conditions such as buffers, ions and additives, and literature information. PINT is cross-linked with other related databases such as PIR, SWISS-PROT, PDB and the NCBI PUBMED literature database.
Proper citation: PINT (RRID:SCR_007856) Copy
A database of mRNA polyadenylation sites. PolyA_DB version 1 contains human and mouse poly(A) sites that are mapped by cDNA/EST sequences. PolyA_DB version 2 contains poly(A) sites in human, mouse, rat, chicken and zebrafish that are mapped by cDNA/EST and Trace sequences. Sequence alignments between orthologous sites are available. PolyA_SVM predicts poly(A) sites using 15 cis elements identified for human poly(A) sites.
Proper citation: PolyA DB (RRID:SCR_007867) Copy
A web analysis system and resource, which provides comprehensive information on piRNAs in the widely studied mammals. It compiles all the possible clusters of piRNAs and also depicts piRNAs along with the associated genomic elements like genes and repeats on a genome wide map. piRNABank mainly provides data onnamely Human, Mouse, Rat, Zebrafish, Platypus and a fruit fly, Drosophila.Search options have been designed to query and obtain useful data from this online resource. It also facilitates abstraction of sequences and structural features from piRNA data. piRNABank provides the following features: * Simple search * Search piRNA clusters * Search homologous piRNAs * piRNA visualization map * Analysis tools, THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: piRNABank (RRID:SCR_007858) Copy
Resource for reuse, sharing and meta-analysis of expression profiling data. Database and set of tools for meta analysis, reuse and sharing of genomics data. Targeted at analysis of gene expression profiles. Users can search, access and visualize coexpression and differential expression results.
Proper citation: Gemma (RRID:SCR_008007) Copy
http://www.grt.kyushu-u.ac.jp/spad/
It is divided to four categories based on extracellular signal molecules (Growth factor, Cytokine, and Hormone) and stress, that initiate the intracellular signaling pathway. SPAD is compiled in order to describe information on interaction between protein and protein, protein and DNA as well as information on sequences of DNA and proteins. There are multiple signal transduction pathways: cascade of information from plasma membrane to nucleus in response to an extracellular stimulus in living organisms. Extracellular signal molecule binds specific intracellular receptor, and initiates the signaling pathway. Now, there is a large amount of information about the signaling pathway which controls the gene expression and cellular proliferation. We have developed an integrated database SPAD to understand the overview of signaling transduction.
Proper citation: Signaling Pathway Database (RRID:SCR_008243) Copy
A horizontally and vertically structured database that pulls scientific and medical information and describes it consistently using the Ingenuity Ontology. The Knowledge Base pulls information from journals, public molecular content databases, and textbooks. Data is curated and and integrated into the Knowledge Base .
Proper citation: Ingenuity Pathways Knowledge Base (RRID:SCR_008117) Copy
http://www.bioinformatics2.wsu.edu/cgi-bin/Athena/cgi/home.pl
Athena is a web-based application that warehouses disparate datatypes related to the control of gene expression. Athena provides several features to enable exploration of the regulatory mechanisms of Arabidopsis gene control. The first main tool we provide is visualization of promoter domains of selected genes. Database crossreference for these transcription factors is provided as well as a statistical test for enrichment of binding activity within the set of selected promoters. The data mining tools in Athena allow for selection of sets of genes based on two different factors. -Genes can be select by specifying a set of binding factors whose putative sites must be present within all of those genes'' promoter regions. -Alternatively, genes can be selected using Gene Ontology annotations. Both GO (Gene Ontology) Slim terms and Gene Ontology terms are available. One can select a set of genes by either choosing a union of the genes annotated by a selected set of Slim terms or Gene Ontology terms. The selected gene''s putative binding factors are listed, including enrichment data. Furthermore, enriched presence of Gene Ontology terms is given. The analysis suite provides both enhanced data mining tools for selecting genes as well as several data displays., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Athena (RRID:SCR_008110) Copy
WHOSIS, the WHO Statistical Information System, is an interactive database bringing together core health statistics for the 193 WHO Member States. It comprises more than 100 indicators, which can be accessed by way of a quick search, by major categories, or through user-defined tables. The data can be further filtered, tabulated, charted and downloaded. The data are also published annually in the World Health Statistics Report released in May. The WHO Statistical Information System is the guide to health and health-related epidemiological and statistical information available from the World Health Organization. Most WHO technical programs make statistical information available, and they will be linked from here. Sponsors: WHOSIS is supported by the World Health Organization. Note: The WHO Statistical Information System (WHOSIS) has been incorporated into the Global Health Observatory (GHO) to provide you with more data, more tools, more analysis and more reports.
Proper citation: World Health Organization Statistical Information System (RRID:SCR_008250) Copy
A database, catalog and index to the collections of the National Agricultural Library, as well as a primary public source for world-wide access to agricultural information. This database resource covers materials in all formats and periods, including printed works from as far back as the 15th century. AGRICOLA is a bibliographic database of citations to the agricultural literature created by the National Agricultural Library and its cooperators. The records describe publications and resources encompassing all aspects of agriculture and allied disciplines, including animal and veterinary sciences, entomology, plant sciences, forestry, aquaculture and fisheries, farming and farming systems, agricultural economics, extension and education, food and human nutrition, and earth and environmental sciences. Although the NAL Catalog (AGRICOLA) does not contain the text of the materials it cites, thousands of its records are linked to full-text documents online, with new links added daily. The NAL Catalog (AGRICOLA) is organized into two bibliographic data sets: *The NAL Online Public Access Catalog (AGRICOLA NAL) contains citations to books, audiovisuals, serials, and other materials, most of which are in the Library''s collection. (The Catalog does contain some records for items not held at NAL.) *The Article Citation Database (AGRICOLA IND) includes citations, many with abstracts, to journal articles (see Journals Indexed in AGRICOLA), book chapters, reports, and reprints, selected primarily from the materials found in the NAL Catalog.
Proper citation: AGRICOLA (RRID:SCR_008158) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the RRID Resources search. From here you can search through a compilation of resources used by RRID and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that RRID has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on RRID then you can log in from here to get additional features in RRID such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into RRID you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within RRID that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.