Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Datasets and tools for comparative analysis and annotation of all publicly available genomes from three domains of life in a uniquely integrated context. Plasmids that are not part of a specific microbial genome sequencing project and phage genomes are also included in order to increase its genomic context for comparative analysis. The user interface (see User Interface Map) allows navigating the microbial genome data space along its three key dimensions (genes, genomes, and functions), and groups together the main comparative analysis tools. Microbial genome data analysis in IMG usually starts with the definition of an analysis context in terms of selected genomes, functional annotations, and/or genes, followed by the individual or comparative analysis of genomes, functional annotations, or genes.
Proper citation: IMG (RRID:SCR_007733) Copy
It contains predictions of precursor miRNA genes covering several animal genomes combining orthology and a Support Vector Machine. We provide homology extended alignments of already known miRBase families and putative miRNA families exclusively predicted by our SVM and orthology pipeline. The current release of miROrtho covers 46 animal genomes. We provide homology extended alignments of already known miRBase families and putative miRNA families exclusively predicted by our SVM and orthology pipeline.
Proper citation: miROrtho: the catalogue of animal microRNA genes (RRID:SCR_007797) Copy
http://biobases.ibch.poznan.pl/ncRNA/
It is intended to provide information on the sequences and functions of transcripts which do not code for proteins, but perform regulatory roles in the cell. Currently, the database includes over 30,000 individual sequences from 99 species of Bacteria, Archaea and Eukaryota. The primary source of sequences included in the database was the GenBank. Additional annotation information for mouse and human ncRNAs was derived from FANTOM3 database and H-inviational Integrated Database of Annotated Human Genes version 3.4, respectively. Genome mapping information was derived from tha data available at the UCSC Genome Browser site. The sequences and annotations of small cytoplasmic RNAs from bacteria, for which annotation is lacking in the genome sequences, were derived from the Rfam database. The microRNAs or snoRNAs which were available in previous editions, as well as other housekeeping (infrastructural) RNAs (e.g. rRNA, tRNA, snRNA, SRP RNA) are not included in our database to avoid redundancy with more specialized databases which emerged in recent years.
Proper citation: Noncoding RNA database (RRID:SCR_007815) Copy
http://www.cmbi.ru.nl/phylopat
A database of phylogenetic patterns of evolution between 46 different species. PhyloPat uses the latest release of EnsMart (release 52), and their one-to-one, one-to-many and many-to-many orthologies. First, we stored all of the Ensembl IDs within the 46 species, and the orthologies between them. Second, we determined the evolutionary order of the studied species using the NCBI Taxonomy database. The phylogenetic tree of these species can be viewed here. Third, we used this phylogenetic tree as a starting point for building our phylogenetic lineages. For each gene in the first species (S. cerevisiae), we looked for orthologs in the other species. All orthologs were added to the phylogenetic lineage, and in the next round were checked for orthologs themselves, until no more orthologies were found for any of the genes. This process was repeated for all genes in all species that were not connected to any phylogenetic lineage yet. The complete phylogenetic lineage determination generated 329,998 phylogenetic lineages, consisting of 973,821 genes. These lineages can be queried here by phylogenetic patterns, MySQL regular expressions or simply a list of Ensembl/EMBL/EntrezGene/HGNC IDs. Output can be given in HTML, Excel or plain text format.
Proper citation: PhyloPat (RRID:SCR_007851) Copy
http://phylomedb.bioinfo.cipf.es
Database for phylomes, that is, complete collections of phylogenetic trees for all proteins encoded in a given genome. It aims at providing a repository of high-quality phylogenies and alignments for proteins encoded in model species. To derive a phylome, each protein encoded in a given genome is used as a seed to retrieve its homologs in other complete genomes. These sequences are aligned and processed to derive reliable phylogenies using several phylogenetic methods. Besides providing the evolutionary history of the gene families, phylomeDB includes phylogeny based predictions of orthology and paralogy relationships., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: PhylomeDB (RRID:SCR_007850) Copy
A publicly available database resource containing the assembled partial genomes for ~700 eukaryotic organisms. Partial genomes are generated from expressed sequence tag datasets containing more than 1000 sequences. PartiGeneDB allows users to view sets of genes and identify genes of interest in organisms for which a full genome is not currently available. PartiGeneDB is automatically updated to include new organism datasets as they are generated. PartiGeneDB provides four portals of entry into the database. It is hosted and supported by the Hospital for Sick Children, Toronto. In addition to providing a comprehensive resource facilitating comparative analyses, PartiGeneDB allows researchers to access the partial genomes of organisms that may not be available elsewhere. However, we recommend and encourage users interested in exploring datasets from a single organism in more depth, that you visit the specific web sites associated with the sequencing effort associated with that organism .
Proper citation: PartiGeneDB (RRID:SCR_007848) Copy
MetaCyc is a database of nonredundant, experimentally elucidated metabolic pathways. MetaCyc contains more than 1,200 pathways from more than 1,600 different organisms, and is curated from the scientific experimental literature. MetaCyc contains pathways involved in both primary and secondary metabolism, as well as associated compounds, enzymes, and genes.
Proper citation: MetaCyc (RRID:SCR_007778) Copy
An information resource for peptidases (also termed proteases, proteinases and proteolytic enzymes) and the proteins that inhibit them. The MEROPS database uses an hierarchical, structure-based classification of the peptidases. In this, each peptidase is assigned to a Family on the basis of statistically significant similarities in amino acid sequence, and families that are thought to be homologous are grouped together in a Clan. There is a Summary page for each family and clan, and these have indexes. Each of the Summary pages offers links to supplementary pages. About 3000 individual peptidases and inhibitors are included in the database, and there is a Summary page describing each one. You can navigate to this by any of several routes. There are indexes of Name, MEROPS Identifier and source Organism on the menu bar. Each Summary page describes the classification and nomenclature of the peptidase or inhibitor, and provides links to supplementary pages showing sequence identifiers, the structure if known, literature references and more.
Proper citation: MEROPS (RRID:SCR_007777) Copy
LOCATE is a curated database that houses data describing the membrane organization and subcellular localization of proteins from the RIKEN FANTOM4 mouse and human protein sequence set. The membrane organization is predicted by the high-throughput, computational pipeline MemO. The subcellular locations were determined by a high-throughput, immunofluorescence-based assay and by manually reviewing peer-reviewed publications.
Proper citation: LOCATE: subcellular localization database (RRID:SCR_007763) Copy
An integrated and comprehensive database of virulence factors for bacterial pathogens (also including Chlamydia and Mycoplasma). VFDB is a platform for further study of comparative pathogenomics. Major features include tabular comparison of pathogenomic composition in terms of virulence, multiple alignments and statistic analysis of homologous virulence genes, and graphical comparison of pathogenomic organization of VFs. Category: Genomics Databases (non-vertebrate) Subcategory: Prokaryotic genome databases
Proper citation: VFDB - Virulence Factors of Bacterial Pathogens (RRID:SCR_007969) Copy
This database functions both as a website where researchers can look for information on their targets of interest; and as a tool for prioritization of targets in whole genomes. Using the database as a tool, researchers can quickly prioritize a genome of interest by performing any number of individual queries on a species of interest, then assigning numerical weights to each query (in the history page) to finally obtain a ranked list of genes by combining the weighted queries. This site is part of a WHO/TDR project seeking to exploit the availability of diverse datasets to facilitate the identification and prioritization of drug targets in pathogens causing neglected diseases.
Proper citation: TDR Targets Database (RRID:SCR_007963) Copy
http://bioafrica.mrc.ac.za/rnavirusdb/
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 19, 2016. It is a database and web application describing the genome organization and providing analytical tools for the 938 known species of RNA virus. It can identify submitted nucleotide sequences, can place them into multiple whole-genome alignments (in species where more than one isolate has been fully sequenced) and contains translated genome sequences for all species. It has been created for two main purposes: to facilitate the comparative analysis of RNA viruses and to become a hub for other, more specialised virus Web sites.
Proper citation: RNA Virus Database (RRID:SCR_007899) Copy
Alternative splicing essentially increases the diversity of the transcriptome and has important implications for physiology, development and the genesis of diseases. This resource uses a different approach to investigate alternative splicing (instead of the conventional case-by case fashion) and integrates all transcripts derived from a gene into a single splicing graph. ASG is a database of splicing graphs for human genes, using transcript information from various major sources (Ensembl, RefSeq, STACK, TIGR and UniGene). Each transcript corresponds to a path in the graph, and alternative splicing is displayed by bifurcations. This representation preserves the relationships between different splicing variants and allows us to investigate systematically all possible putative transcripts. Web interface allows users to display the splicing graphs, to interactively assemble transcripts and to access their sequences as well as neighboring genomic regions. ASG also provide for each gene, an exhaustive pre-computed catalog of putative transcriptsin total more than 1.2 million sequences. It has found that ~65 of the investigated genes show evidence for alternative splicing, and in 5 of the cases, a single gene might produce over 100 transcripts.
Proper citation: Alternate splicing gallery (RRID:SCR_008129) Copy
http://www-personal.umich.edu/~jianghui/rseq/
A software toolkit for RNA sequence data analysis. It contains programs that cover several aspects of RNA-Seq data analysis such as read quality assessment, reference sequence generation, sequence mapping, and gene and isoform expressions estimations.
Proper citation: rSeq (RRID:SCR_000562) Copy
https://services.healthtech.dtu.dk/services/DictyOGlyc-1.1/
Server that produces neural network predictions for GlcNAc O-glycosylation sites in Dictyostelium discoideum proteins.
Proper citation: DictyOGlyc (RRID:SCR_001600) Copy
http://www.glycosciences.de/modeling/glyprot/
Web-based tool that enables meaningful N-glycan conformations to be attached to all the spatially accessible potential N-glycosylation sites of a known three-dimensional (3D) protein structure. The 3D structure of protein is required as input. Potential N-glysylations site are automatically detected. The attached glycan are constructed with SWEET-II, http://www.glycosciences.de/modeling/sweet2/doc/index.php
Proper citation: GlyProt (RRID:SCR_001560) Copy
Database of experimentally verified phosphorylation sites in eukaryotic proteins. Entries are manually curated with links to literature references, information about structure, interaction partners and sub-cellular compartment tissues, and sequences from the UniProt database.
Proper citation: Phospho.ELM (RRID:SCR_001109) Copy
http://hipipe.ncgm.sinica.edu.tw/
Tool that provides high performance NGS (next-generation sequencing) data analysis pipelines so that researchers with minimum IT or bioinformatics knowledge can perform common analyses on NGS data. 3 TB of storage space is reserved for each task.
Proper citation: HiPipe (RRID:SCR_001215) Copy
A database dedicated to the collection and classification of mobile genetic elements (MGEs) from various sources, comprising all known phage genomes, plasmids and transposons. In addition to provide information on the full genomes and genetic entities, it aims at building a comprehensive classification of the functional modules of MGE's at the protein, gene, and higher levels. Prophinder, a tool dedicated to the detection of prophages in sequenced bacterial genomes, is available on ACLAME.
Proper citation: A Classification of Mobile genetic Elements (RRID:SCR_001694) Copy
Database for conserved sequence motifs identified by genome scale motif discovery, similarity, clustering, co-occurrence and coexpression calculations. Sequence inputs include low-coverage genome sequence data and ENCODE data. The database offers information on atomic motifs, motif groups and patterns. In promoter-based cisRED databases, sequence search regions for motif discovery extend from 1.5 Kb upstream to 200b downstream of a transcription start site, net of most types of repeats and of coding exons. Many transcription factor binding sites are located in such regions. For each target gene's search region, a base set of probabilistic ab initio discovery tools is used, in parallel, to find over-represented atomic motifs. Discovery methods use comparative genomics with over 40 vertebrate input genomes. In ChIP-seq-based cisRED databases, sequence search regions for motif discovery correspond to significant peaks that represent genome-wide sites of protein-DNA binding. Because such peaks occur in a wide range of genic and intergenic locations, ChIP-seq and promoter-based databases are complementary. Currently, motif discovery for ChIP-seq data uses scan-based approaches that make more explicit use of sets of sequences known to be functional transcription factor binding sites, and that consider a wide range of levels of conservation. For the human STAT1 ChIP-seq database search regions in the target species (human) was selected +/- 300 bp around the ChIP-seq peak maximum. Repeats and coding regions were masked. Multiple sequence alignment were used to assemble orthologous input sequences from other species.
Proper citation: cisRED: cis-regulatory element (RRID:SCR_002098) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the RRID Resources search. From here you can search through a compilation of resources used by RRID and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that RRID has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on RRID then you can log in from here to get additional features in RRID such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into RRID you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within RRID that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.