Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
The web portal provides comprehensive local database of human genome variants with a user-friendly web page that provides a one-stop annotating and funtonal prediction service which is both convenient and up-to-date. A query can be accepted as either a dbSNP Id or a chromosomal location and our system will instantly provide all the annotation information in an interactive LD panel. The system can also simultaneously prioritize this variant based on additive effect mode by corresponding annotation information and evaluate the variant effect that is then displayed in a prioritization tree. Furthermore, cohort sequencing continuously produces lots of un-annotated variants such as rare variants or de novo variants, and our system can even fit this data by accepting genomic coordinates (hg19) to offer maximal annotations. Main Functions Over 40 up-to-date annotation items for human single nucleotide variations; Functional prediction for different types of variants; Dynamic LD panel for both HapMap and 1000 Genomes Project populations; Prioritization score and tree viewer based on variant functional model.
Proper citation: SNVrap (RRID:SCR_010512) Copy
http://www.cbil.upenn.edu/cgi-bin/tess/tess
TESS is a web tool for predicting transcription factor binding sites in DNA sequences. It can identify binding sites using site or consensus strings and positional weight matrices from the TRANSFAC, JASPAR, IMD, and our CBIL-GibbsMat database. You can use TESS to search a few of your own sequences or for user-defined CRMs genome-wide near genes throughout genomes of interest. Search for CRMs Genome-wide: TESS now has the ability to search whole genomes for user defined CRMs. Try a search in the AnGEL CRM Searches section of the navigation bar.. You can search for combinations of consensus site sequences and/or PWMs from TRANSFAC or JASPAR. Search DNA for Binding Sites: TESS also lets you search through your own sequence for TFBS. You can include your own site or consensus strings and/or weight matrices in the search. Use the Combined Search under ''Site Searches'' in the menu or use the box for a quick search. TESS assigns a TESS job number to all sequence search jobs. The job results are stored on our server for a period of time specified in the search submit form. During this time you may recall the search results using the form on this page. TESS can also email results to you as a tab-delimited file suitable for loading into a spreadsheet program. Query for Transcription Factor Info: TESS also has data browsing and querying capabilities to help you learn about the factors that were predicted to bind to your sequence. Use the Query TRANSFAC or Query Matrices links above or use the search interface provided from the home page.
Proper citation: TESS: Transcription Element Search System (RRID:SCR_010739) Copy
http://bio-bwa.sourceforge.net/
Software for aligning sequencing reads against large reference genome. Consists of three algorithms: BWA-backtrack, BWA-SW and BWA-MEM. First for sequence reads up to 100bp, and other two for longer sequences ranged from 70bp to 1Mbp.
Proper citation: BWA (RRID:SCR_010910) Copy
https://www.genome-cloud.com/user/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 29, 2019. A cloud platform for next-generation sequencing analysis and storage. Services include: * g-Analysis: Automated genome analysis pipelines at your fingertips * g-Cluster: Easy-of-use and cost-effective genome research infrastructure * g-Storage: A simple way to store, share and protect data * g-Insight: Accurate analysis and interpretation of biological meaning of genome data
Proper citation: GenomeCloud (RRID:SCR_011886) Copy
Web based instant protein network modeler for newly sequenced species. Web server designed to instantly construct genome scale protein networks using protein sequence data. Provides network visualization, analysis pages and solution for instant network modeling of newly sequenced species.
Proper citation: JiffyNet (RRID:SCR_011954) Copy
Catalog of published genome-wide association studies. Genome-wide set of genetic variants in different individuals to see if any variant is associated with trait and disease. Database of genome-wide association study (GWAS) publications including only those attempting to assay single nucleotide polymorphisms (SNPs). Publications are organized from most to least recent date of publication. Studies are identified through weekly PubMed literature searches, daily NIH-distributed compilations of news and media reports, and occasional comparisons with an existing database of GWAS literature (HuGE Navigator). Works with HANCESTRO ancestry representation.
Proper citation: GWAS: Catalog of Published Genome-Wide Association Studies (RRID:SCR_012745) Copy
Integrated database resource consisting of 16 main databases, broadly categorized into systems information, genomic information, and chemical information. In particular, gene catalogs in completely sequenced genomes are linked to higher-level systemic functions of cell, organism, and ecosystem. Analysis tools are also available. KEGG may be used as reference knowledge base for biological interpretation of large-scale datasets generated by sequencing and other high-throughput experimental technologies.
Proper citation: KEGG (RRID:SCR_012773) Copy
A high-quality integrated knowledge resource specialized in the immunoglobulins (IG) or antibodies, T cell receptors (TR), major histocompatibility complex (MHC) of human and other vertebrate species, and in the immunoglobulin superfamily (IgSF), MHC superfamily (MhcSF) and related proteins of the immune system (RPI) of vertebrates and invertebrates, serving as the global reference in immunogenetics and immunoinformatics. IMGT provides a common access to sequence, genome and structure Immunogenetics data, based on the concepts of IMGT-ONTOLOGY and on the IMGT Scientific chart rules. IMGT works in close collaboration with EBI (Europe), DDBJ (Japan) and NCBI (USA). IMGT consists of sequence databases, genome database, structure database, and monoclonal antibodies database, Web resources and interactive tools.
Proper citation: IMGT - the international ImMunoGeneTics information system (RRID:SCR_012780) Copy
http://www-sequence.stanford.edu/group/candida/
The Stanford Genome Technology Center began a whole genome shotgun sequencing of strain SC5314 of Candida albicans. After reaching its original goal of 1.5X mean coverage of the haploid genome (16Mb) in summer, 1998, Stanford was awarded a supplemental grant to continue sequencing up to a coverage of 10X, performing as much assembly of the sequence as possible, using recognizable genes as nucleation points. Candida albicans is one of the most commonly encountered human pathogens, causing a wide variety of infections ranging from mucosal infections in generally healthy persons to life-threatening systemic infections in individuals with impaired immunity. Oral and esophogeal Candida infections are frequently seen in AIDS patients. Few classes of drugs are effective against these fungal infections, and all of them have limitations with regard to efficacy and side-effects.
Proper citation: Sequencing of Candida Albicans (RRID:SCR_013437) Copy
Functional genomic database for malaria parasites. Database for Plasmodium spp. Provides resource for data analysis and visualization in gene-by-gene or genome-wide scale. PlasmoDB 5.5 contains annotated genomes, evidence of transcription, proteomics evidence, protein function evidence, population biology and evolution data. Data can be queried by selecting from query grid or drop down menus. Results can be combined with each other on query history page. Search results can be downloaded with associated functional data and registered users can store their query history for future retrieval or analysis.Key community database for malaria researchers, intersecting many types of laboratory and computational data, aggregated by gene.
Proper citation: PlasmoDB (RRID:SCR_013331) Copy
Database for ESTs (Expressed Sequence Tags), consensus sequences, bacterial artificial chromosome (BAC) clones, BES (BAC End Sequences). They have generated 69,545 ESTs from 6 full-length cDNA libraries (Porcine Abdominal Fat, Porcine Fat Cell, Porcine Loin Muscle, Liver and Pituitary gland). They have also identified a total of 182 BAC contigs from chromosome 6. It is very valuable resources to study porcine quantitative trait loci (QTL) mapping and genome study. Users can explore genomic alignment of various data types, including expressed sequence tags (ESTs), consensus sequences, singletons, QTL, Marker, UniGene and BAC clones by several options. To estimate the genomic location of sequence dataset, their data aligned BES (BAC End Sequences) instead of genomic sequence because Pig Genome has low-coverage sequencing data. Sus scrofa Genome Database mainly provide comparative map of four species (pig, cattle, dog and mouse) in chromosome 6.
Proper citation: PiGenome (RRID:SCR_013394) Copy
http://bioinformatics.psb.ugent.be/ENIGMA/
A software tool to extract gene expression modules from perturbational microarray data, based on the use of combinatorial statistics and graph-based clustering. The modules are further characterized by incorporating other data types, e.g. GO annotation, protein interactions and transcription factor binding information, and by suggesting regulators that might have an effect on the expression of (some of) the genes in the module. Version : ENIGMA 1.1 used GO annotation version : Aug 29th 2007
Proper citation: ENIGMA (RRID:SCR_013400) Copy
http://fnih.org/work/past-programs/genetic-association-information-network-gain
The Genetic Association Information Network (GAIN) supports a series of Genome-Wide Association Studies (GWAS) designed to identify specific points of DNA variation associated with the occurrence of a particular common disease. Initially focusing on six major common diseases, GAIN focused on combining the results with clinical data to create a significant new resource for genetic researchers.
Proper citation: Genetic Association Information Network (GAIN) (RRID:SCR_013703) Copy
A SEED-quality automated service that annotates complete or nearly complete bacterial and archaeal genomes across the entire phylogenetic tree. RAST can also be used to analyze draft genomes.
Proper citation: RAST Server (RRID:SCR_014606) Copy
http://www.vicbioinformatics.com/software.prokka.shtml
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for the rapid annotation of prokaryotic genomes. It produces GFF3, GBK and SQN files that are ready for editing in Sequin and ultimately submitted to Genbank/DDJB/ENA. A typical 4 Mbp genome can be fully annotated in less than 10 minutes on a quad-core computer, and scales well to 32 core SMP systems., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Prokka (RRID:SCR_014732) Copy
https://www.encodeproject.org/
Consortium to build comprehensive parts list of functional elements in human genome. This includes elements that act at protein and RNA levels, and regulatory elements that control cells and circumstances in which gene is active. Data from 2012-present.
Proper citation: Encode (RRID:SCR_015482) Copy
http://compbio.dfci.harvard.edu/tgi/
THIS RESOURCE IS NO LONGER IN SERVICE, documented May 10, 2017. A pilot effort that has developed a centralized, web-based biospecimen locator that presents biospecimens collected and stored at participating Arizona hospitals and biospecimen banks, which are available for acquisition and use by researchers. Researchers may use this site to browse, search and request biospecimens to use in qualified studies. The development of the ABL was guided by the Arizona Biospecimen Consortium (ABC), a consortium of hospitals and medical centers in the Phoenix area, and is now being piloted by this Consortium under the direction of ABRC. You may browse by type (cells, fluid, molecular, tissue) or disease. Common data elements decided by the ABC Standards Committee, based on data elements on the National Cancer Institute''s (NCI''s) Common Biorepository Model (CBM), are displayed. These describe the minimum set of data elements that the NCI determined were most important for a researcher to see about a biospecimen. The ABL currently does not display information on whether or not clinical data is available to accompany the biospecimens. However, a requester has the ability to solicit clinical data in the request. Once a request is approved, the biospecimen provider will contact the requester to discuss the request (and the requester''s questions) before finalizing the invoice and shipment. The ABL is available to the public to browse. In order to request biospecimens from the ABL, the researcher will be required to submit the requested required information. Upon submission of the information, shipment of the requested biospecimen(s) will be dependent on the scientific and institutional review approval. Account required. Registration is open to everyone.. Documented on August 19,2019.The goal of The Gene Index Project is to use the available Expressed Sequence Transcript (EST) and gene sequences, along with the reference genomes wherever available, to provide an inventory of likely genes and their variants and to annotate these with information regarding the functional roles played by these genes and their products. The promise of genome projects has been a complete catalog of genes in a wide range of organisms. While genome projects have been successful in providing reference genome sequences, the problem of finding genes and their variants in genomic sequence remains an ongoing challenge. TGI has created an inventory that contains genes and their variants together with description. In addition, this resource is attempting to use these catalogs to find links between genes and pathways in different species and to provide lists of features within completed genomes that can aid in the understanding of how gene expression is regulated. DATABASES *Eukaryotic Gene Orthologues (formerly known as TOGA - TIGR Orthologous Gene Alignment): Eukaryotic Gene Orthologues (EGO) at DFGI are generated by pair-wise comparison between the Tentative Consensus (TC) sequences that comprise the Dana Farber Gene Indices from individual organisms. The reciprocal pairs of the best match were clustered into individual groups and multiple sequence alignments were displayed for each group. *GeneChip Oncology Database (GCOD):Cancer gene expression database is a collection of publicly available microarray expression data on Affymetrix GeneChip Arrays related to human cancers. Currently only datasets with available raw data (Affymetrix .CEL files) are processed. All processed datasets were subjected to extensive manual curation, uniform processing and consistent quality control. You can browse the experiments in our collection, perform statistical analysis, and download processed data; or to search gene expression profiles using Entrez gene symbol, Unigene ID, or Affymetrix probeset ID. *Gene Indices: As of July 1, 2008, there are 111 publicly available gene indices. They are separated into 4 categories for better organization and easier access. Animal: 41, Plant: 45, Protist: 15, Fungal: 10 *Genomic Maps: Human, mouse, rat, chicken, drosophila melanogaster, zebrafish, mosquito, caenorhabditis elegans, Arabidopsis thaliana, rice, yeast, fission yeast Dana-Farber Cancer Institute (DFCI) Gene Indices Software Tools: *TGI Clustering tools (TGICL): a software system for fast clustering of large EST datasets. *GICL: this package contains the scripts and all the necessary pre-compiled binaries for 32bit Linux systems. *clview: an assembly file viewer. *SeqClean:a script for automated trimming and validation of ESTs or other DNA sequences by screening for various contaminants, low quality and low-complexity sequences. *cdbfasta/cdbyank: fast indexing/retrieval of fasta records from flat file databases. *DAS/XML Genomic Viewer The Genomic viewer borrows modules from http://www.biodas.org (lstein (at) cshl.org) & http://webreference.com.
Proper citation: Gene Index Project (RRID:SCR_002148) Copy
http://mips.gsf.de/genre/proj/yeast/index.jsp
The MIPS Comprehensive Yeast Genome Database (CYGD) aims to present information on the molecular structure and functional network of the entirely sequenced, well-studied model eukaryote, the budding yeast Saccharomyces cerevisiae. In addition, the data of various projects on related yeasts are used for comparative analysis.
Proper citation: CYGD - Comprehensive Yeast Genome Database (RRID:SCR_002289) Copy
http://microbes.ucsc.edu/cgi-bin/hgGateway?db=neisMeni_MC58_1
Portal contains detailed information for Neisseria meningitidis MC58. Information include DNA molecule summary, primary annotation summary, and taxonomy. It is a tool that allows the researcher to access all of the bacterial genome sequences completed to date. Users may access information on all of the bacterial genomes or any subset of them. Information in the website about its DNA molecule includes: total number of DNA molecules, total size of all DNA molecules, number of primary annotation coding bases, and number of G + C bases. Its primary annotation summary include: total genes, protein coding genes, tRNA genes, and rRNA genes. Sponsors: The CMR was previously funded by two grants, one from the U.S. Department of Energy (DOE) and one from the National Science Foundation (NSF). It is currently partially funded by a Microbial Sequence Center (MSC) grant from the National Institute of Allergy and Infectious Diseases (NIAID)
Proper citation: Neisseria meningitidis MC58 Genome Page (RRID:SCR_002200) Copy
A comprehensive collection of experimentally determined and computationally predicted CCCTC-binding factor (CTCF) binding sites (CTCFBS) from the literature. The database is designed to facilitate the studies on insulators and their roles in demarcating functional genomic domains. The CTCFBS Prediction Tool allows users to scan sequences for the single best match to CTCF position weight matrices. Currently (March 2014), the database contains almost 15 million experimentally determined CTCF binding sites across several species. CTCF binding sites were collected from published papers containing CTCF binding sites identified using ChIPSeq or similar methods, data from the ENCODE project, and a set of approximately 100 manually curated binding sites identified by low-throughput experiments. Users can browse insulator sequence features, function annotations, genomic contexts including histone methylation profiles, flanking gene expression patterns and orthologous regions in other mammalian genomes. Users can also retrieve data by text search, sequence search and genomic range search.
Proper citation: CTCFBSDB (RRID:SCR_002279) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the RRID Resources search. From here you can search through a compilation of resources used by RRID and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that RRID has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on RRID then you can log in from here to get additional features in RRID such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into RRID you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within RRID that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.