Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://jjwanglab.org:8080/gwasdb/
Combines collections of genetic variants (GVs) from GWAS and their comprehensive functional annotations, as well as disease classifications. Used to maximize utilility of GWAS data to gain biological insights through integrative, multi-dimensional functional annotation portal. In addition to all GVs annotated in NHGRI GWAS Catalog, we manually curate GVs that are marginally significant (P value < 10-3) by looking into supplementary materials of each original publication and provide extensive functional annotations for these GVs. GVs are manually classified by diseases according to Disease Ontology Lite and HPO (Human Phenotype Ontology) for easy access. Database can also conduct gene based pathway enrichment and PPI network association analysis for those diseases with sufficient variants. SOAP services are available. You may Download GWASdb SNP. (This file contains all of the significant SNP in GWASdb. In the pvalue column, 0 means this P-value is not reported in the study but it is significant SNP. In the source column, GWAS:A represents the original data in GWAS catalog, while GWAS:B is our curation data which P-value < 10-3)
Proper citation: GWASdb (RRID:SCR_006015) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on July 7, 2022. Federation of International Mouse Resources (FIMRe) is a collaborating group of Mouse Repository and Resource Centers worldwide whose collective goal is to archive and provide strains of mice as cryopreserved embryos and gametes, ES cell lines, and live breeding stock to the research community. Goals of the Federation of International Mouse Resources: * Coordinate repositories and resource centers to: ** archive valuable genetically defined mice and ES cell lines being created worldwide ** meet research demand for these genetically defined mice and ES cell lines * Establish consistent, highest quality animal health standards in all resource centers * Provide genetic verification and quality control for genetic background and mutations * Provide resource training to enhance user ability to utilize cryopreserved resources
Proper citation: Federation of International Mouse Resources (RRID:SCR_006137) Copy
http://202.38.126.151:8080/SDisease/
Curated database of experimentally supported data of RNA Splicing mutation and disease. The RNA Splicing mutations include cis-acting mutations that disrupt splicing and trans-acting mutations that affecting RNA-dependent functions that cause disease. Information such as EntrezGeneID, gene genomic sequence, mutation (nucleotide substitutions, deletions and insertions), mutation location within the gene, organism, detailed description of the splicing mutation and references are also given. Users are able to submit new entries to the database. This database integrating RNA splicing and disease associations would be helpful for understanding not only the RNA splicing but also its contribution to disease. In SpliceDisease database, they manually curated 2337 splicing mutation disease entries involving 303 genes and 370 diseases, which have been supported experimentally in 898 publications. The SpliceDisease database provides information including the change of the nucleotide in the sequence, the location of the mutation on the gene, the reference PubMed ID and detailed description for the relationship among gene mutations, splicing defects and diseases. They standardized the names of the diseases and genes and provided links for these genes to NCBI and UCSC genome browser for further annotation and genomic sequences. For the location of the mutation, they give direct links of the entry to the respective position/region in the genome browser.
Proper citation: SpliceDisease (RRID:SCR_006130) Copy
http://www.mousephenotype.org/impress
Contains standardized phenotyping protocols essential for the characterization of mouse phenotypes. IMPReSS holds definitions of the phenotyping Pipelines and mandatory and optional Procedures and Parameters carried out and data collected by international mouse clinics following the protocols defined. This allows data to be comparable and shareable and ontological annotations permit interspecies comparison which may help in the identification of phenotypic mouse-models of human diseases. The IMPC (International Mouse Phenotyping Consortium) core pipeline describes the phenotype pipeline that has been agreed by the research institutions. IMPReSS has a SOAP web service machine interface. The WSDL can be accessed here: http://www.mousephenotype.org/impress/soap/server?wsdl
Proper citation: Impress (RRID:SCR_006160) Copy
Clearinghouse and exchange portal for gene variant (mutation) data produced by diagnostics laboratories, offering users a portal through which to announce, discover and acquire a comprehensive listing of observed neutral and disease-causing gene variants in patients and unaffected individuals. Cafe Variome is not a ''''database'''' for the hosting/display/release of data, but a shop window for finding data. As such, it holds only core info for each record, and uses this merely to enable holistic searching across resources. Diagnostics laboratories routinely assess DNA samples from patients with various inherited disorders, and so produce a great wealth of data on the genetic basis of disease. Unfortunately, those data are not usually shared with others. To address this gross deficiency, a novel system has been developed that aims to facilitate the automated transfer of diagnostic laboratory data to the wider community, via an internet based Cafe for routinely exchanging genetic variation data. The flow of research data concerning the genetic basis of health and disease is critical to understanding and developing treatments for a range of genetic diseases. Overall, the project aims to lower the barriers and provide incentives for a willing community to share data, and thereby facilitate the broader exploitation of diagnostic laboratory data. Cafe Variome aims to address the above data flow problems by: # Minimizing the effort required to publish variant data # Ensuring attribution for data creators working in diagnostic laboratories Key elements of the project strategy are: * Data publication will be automated by endowing standard analysis tools used by laboratories with an online data submission function. Submissions will be received by a central Internet depot, which will serve as a place where published datasets are advertised, and subsequently discovered by diverse 3rd parties. * Each dataset will be unambiguously linked with the data submitter''''s identity, and systems devised to facilitate citation of published variant datasets so they can be cited in the literature. Data creators will thus be credited for their contributions. Data submitters can use Cafe Variome to simply announce or publicize their data to the world. To enable this, only core, non-identifiable data is submitted to the central repository, enabling users to search and discover records of interest in the source repository. The data are not automatically handed on to the user (unless intended by the submitters). Hence, the concept is used to deal with the challenge of maximally sharing data whilst fully respecting ethico-legal considerations.
Proper citation: cafe variome (RRID:SCR_006162) Copy
http://www.lji.org/faculty-research/scientific-cores/functional-genomics/#overview
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on July 5, 2024. Core facility that combines large-scale automation and high-throughput capabilities with gene disruption techniques to pinpoint the function of individual genes and find new ways to disrupt genetic triggers of disease. The research capabilities are aimed towards finding new treatments for immune-related diseases.
Proper citation: La Jolla Institute for Allergy and Immunology Functional Genomics Core Facility (RRID:SCR_014836) Copy
http://www.salk.edu/science/core-facilities/integrative-genomics-and-bioinformatics-core/
Core facility established to assist the Salk community with integrating genomics data into their research. The primary focus of the core is to provide analysis support for next-generation sequencing applications.
Proper citation: Salk Institute Razavi Newman Integrative Genomics and Bioinformatics Core Facility (IGC) (RRID:SCR_014842) Copy
http://www.scienceexchange.com/facilities/university-of-utah
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on April 15,2024. Labs and facilities of the University of Utah, which include: Microarray and Genomic Analysis Core Facility, Flow Cytometry Core Facility, Mutation Generation and Detection Facility, and the Transgenic and Gene Targeting Core.
Proper citation: University of Utah Labs and Facilities (RRID:SCR_001042) Copy
Multicenter observational study designed to identify genetic determinants of diabetic nephropathy. It is conducted in eleven U.S. clinical centers and a coordinating center, and with four ethnic groups (European Americans, African Americans, Mexican Americans, and American Indians). Two strategies are used to localize susceptibility genes: a family-based linkage study and a case-control study using mapping by admixture linkage disequilibrium (MALD). In the family-based study, probands with diabetic nephropathy are recruited with their parents and selected siblings. Linkage analyses will be conducted to identify chromosomal regions containing genes that influence the development of diabetic nephropathy or related quantitative traits such as serum creatinine concentration, urinary albumin excretion, and plasma glucose concentrations. Regions showing evidence of linkage will be examined further with both genetic linkage and association studies to identify genes that influence diabetic nephropathy or related traits. Two types of MALD studies are being done. One is a case-control study of unrelated individuals of Mexican American heritage in which both cases and controls have diabetes, but only the case has nephropathy. The other is a case-control study of African American patients with nephropathy (cases) and their spouses (controls) unaffected by diabetes and nephropathy; offspring are genotyped when available to provide haplotype data. The specific goals of this program: * Delineate genomic regions associated with the development and progression of renal disease(s) * Evaluate whether there is a genetic link between diabetic nephropathy and diabetic retinopathy * Improve outcomes * Provide protection for people at risk and slow the progression of renal disease * Help establish a resource for genetic studies of kidney disease and diabetic complications by creating a repository of genetic samples and a database * Encourage studies of the genetics of progressive renal disease
Proper citation: Family Investigation of Nephropathy of Diabetes (RRID:SCR_001525) Copy
http://www.sci.unisannio.it/docenti/rampone/
Data set of Homo Sapiens Exons, Introns and Splice regions extracted from GenBank Rel.123 with an aim of giving standardized material to train and to assess the prediction accuracy of computational approaches for gene identification and characterization. From the complete GenBank (Primate Sequences Division) Rel.123 (162,557 entries), entries of Human Nuclear DNA including Complete CDS and more than one Exon have been selected, and 4523 exons and 3802 introns have been extracted from these entries. Details about extracted exons and introns are reported (Locus, number, Start and End position in the entry, sequence, length, G+C content, presence of not AGCT data (nucleotide scan check)). Statistics are also reported (overall nucleotides, average G+C content, nucleotide scan check results, number of not GT starting / AG ending introns, minimum / maximum / average length, length standard deviation). 3799+3799 donor and acceptor sites, as windows of 140 nucleotides around each splice site have been extracted. After discarding sequences not including canonical GTAG junctions (65+74), including insufficient data (not enough material for a 140 nucleotide window) (686+589), including not AGCT bases (29+30), and redundant (218+226) there are 2796+ 2880 windows. Finally, there are 271,937 + 332,296 windows of false splice sites, selected by searching canonical GTAG pairs in not splicing positions. The false sites in a range of +/- 60 from a true splice site are marked as proximal.
Proper citation: HS3D - Homo Sapiens Splice Sites Dataset (RRID:SCR_002939) Copy
Curated lists of genes associated to speech / language phenotypes and structural or functional abnormalities observed in patient populations. Entrez ID gene information, as well as gene expression profiles from the Allen Brain Atlas are available. You can also download expression data for a given gene in JSON or XML format.
Proper citation: Speech Language Disorders Database (RRID:SCR_003655) Copy
http://archive.ics.uci.edu/ml/datasets/EEG+Database
Data set from a large study to examine EEG correlates of genetic predisposition to alcoholism. It contains measurements from 64 electrodes placed on the scalp sampled at 256 Hz (3.9-msec epoch) for 1 second. There were two groups of subjects: alcoholic and control. Each subject was exposed to either a single stimulus (S1) or to two stimuli (S1 and S2) which were pictures of objects chosen from the 1980 Snodgrass and Vanderwart picture set. When two stimuli were shown, they were presented in either a matched condition where S1 was identical to S2 or in a non-matched condition where S1 differed from S2. There were 122 subjects and each subject completed 120 trials where different stimuli were shown. The electrode positions were located at standard sites (Standard Electrode Position Nomenclature, American Electroencephalographic Association 1990). Zhang et al. (1995) describes in detail the data collection process. There are three versions of the EEG data set. * The Small Data Set (smni97_eeg_data.tar.gz) contains data for the 2 subjects, alcoholic a_co2a0000364 and control c_co2c0000337. For each of the 3 matching paradigms, c_1 (one presentation only), c_m (match to previous presentation) and c_n (no-match to previous presentation), 10 runs are shown. * The Large Data Set (SMNI_CMI_TRAIN.tar.gz and SMNI_CMI_TEST.tar.gz) contains data for 10 alcoholic and 10 control subjects, with 10 runs per subject per paradigm. The test data used the same 10 alcoholic and 10 control subjects as with the training data, but with 10 out-of-sample runs per subject per paradigm. * The Full Data Set contains all 120 trials for 122 subjects. The entire set of data is about 700 MBytes.
Proper citation: EEG Database (RRID:SCR_001581) Copy
The EBI genomes pages give access to a large number of complete genomes including bacteria, archaea, viruses, phages, plasmids, viroids and eukaryotes. Methods using whole genome shotgun data are used to gain a large amount of genome coverage for an organism. WGS data for a growing number of organisms are being submitted to DDBJ/EMBL/GenBank. Genome entries have been listed in their appropriate category which may be browsed using the website navigation tool bar on the left. While organelles are all listed in a separate category, any from Eukaryota with chromosome entries are also listed in the Eukaryota page. Within each page, entries are grouped and sorted at the species level with links to the taxonomy page for that species separating each group. Within each species, entries whose source organism has been categorized further are grouped and numbered accordingly. Links are made to: * taxonomy * complete EMBL flatfile * CON files * lists of CON segments * Project * Proteomes pages * FASTA file of Proteins * list of Proteins
Proper citation: EBI Genomes (RRID:SCR_002426) Copy
http://www.linked-neuron-data.org/
Neuroscience data and knowledge from multiple scales and multiple data sources that has been extracted, linked, and organized to support comprehensive understanding of the brain. The core is the CAS Brain Knowledge base, a very large scale brain knowledge base based on automatic knowledge extraction and integration from various data and knowledge sources. The LND platform provides services for neuron data and knowledge extraction, representation, integration, visualization, semantic search and reasoning over the linked neuron data. Currently, LND extracts and integrates semantic data and knowledge from the following resources: PubMed, INCF-CUMBO, Allen Reference Atlas, NIF, NeuroLex, MeSH, DBPedia/Wikipedia, etc.
Proper citation: Linked Neuron Data (RRID:SCR_003658) Copy
http://bpg.utoledo.edu/~afedorov/lab/eid.html
Data sets of protein-coding intron-containing genes that contain gene information from humans, mice, rats, and other eukaryotes, as well as genes from species whose genomes have not been completely sequenced. This is a comprehensive and convenient dataset of sequences for computational biologists who study exon-intron gene structures and pre-mRNA splicing. The database is derived from GenBank release 112, and it contains protein-coding genes that harbor introns, along with extensive descriptions of each gene and its DNA and protein sequences, as well as splice motif information. They have created subdatabases of genes whose intron positions have been experimentally determined. The collection also contains data on untranslated regions of gene sequences and intron-less genes. For species with entirely sequenced genomes, species-specific databases have been generated. A novel Mammalian Orthologous Intron Database (MOID) has been introduced which includes the full set of introns that come from orthologous genes that have the same positions relative to the reading frames.
Proper citation: EID: Exon-Intron Database (RRID:SCR_002469) Copy
http://aws.amazon.com/1000genomes/
A dataset containing the full genomic sequence of 1,700 individuals, freely available for research use. The 1000 Genomes Project is an international research effort coordinated by a consortium of 75 companies and organizations to establish the most detailed catalogue of human genetic variation. The project has grown to 200 terabytes of genomic data including DNA sequenced from more than 1,700 individuals that researchers can now access on AWS for use in disease research free of charge. The dataset containing the full genomic sequence of 1,700 individuals is now available to all via Amazon S3. The data can be found at: http://s3.amazonaws.com/1000genomes The 1000 Genomes Project aims to include the genomes of more than 2,662 individuals from 26 populations around the world, and the NIH will continue to add the remaining genome samples to the data collection this year. Public Data Sets on AWS provide a centralized repository of public data hosted on Amazon Simple Storage Service (Amazon S3). The data can be seamlessly accessed from AWS services such Amazon Elastic Compute Cloud (Amazon EC2) and Amazon Elastic MapReduce (Amazon EMR), which provide organizations with the highly scalable compute resources needed to take advantage of these large data collections. AWS is storing the public data sets at no charge to the community. Researchers pay only for the additional AWS resources they need for further processing or analysis of the data. All 200 TB of the latest 1000 Genomes Project data is available in a publicly available Amazon S3 bucket. You can access the data via simple HTTP requests, or take advantage of the AWS SDKs in languages such as Ruby, Java, Python, .NET and PHP. Researchers can use the Amazon EC2 utility computing service to dive into this data without the usual capital investment required to work with data at this scale. AWS also provides a number of orchestration and automation services to help teams make their research available to others to remix and reuse. Making the data available via a bucket in Amazon S3 also means that customers can crunch the information using Hadoop via Amazon Elastic MapReduce, and take advantage of the growing collection of tools for running bioinformatics job flows, such as CloudBurst and Crossbow.
Proper citation: 1000 Genomes Project and AWS (RRID:SCR_008801) Copy
http://www.sgn.cornell.edu/bulk/input.pl?modeunigene
Allows users to download Unigene or BAC information using a list of identifiers or complete datasets with FTP., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Sol Genomics Network - Bulk download (RRID:SCR_007161) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11, 2023. Archiving services, insertional site analysis, pharmacology and toxicology resources, and reagent repository for academic investigators and others conducting gene therapy research. Databases and educational resources are open to everyone. Other services are limited to gene therapy investigators working in academic or other non-profit organizations. Stores reserve or back-up clinical grade vector and master cell banks. Maintains samples from any gene therapy related Pharmacology or Toxicology study that has been submitted to FDA by U.S. academic investigator that require storage under Good Laboratory Practices. For certain gene therapy clinical trials, FDA has required post-trial monitoring of patients, evaluating clinical samples for evidence of clonal expansion of cells. To help academic investigators comply with this FDA recommendation, the NGVB offers assistance with clonal analysis using LAM-PCR and LM-PCR technology.
Proper citation: National Gene Vector Biorepository (RRID:SCR_004760) Copy
http://linux1.softberry.com/spldb/SpliceDB.html
Database of canonical and non-canonical mammalian splice sites. The information about verified splice site sequences for canonical and non-canonical sites is presented with the supporting evidence. Weight matrices were built for the major splice groups, which can be incorporated into gene prediction programs.
Proper citation: SpliceDB (RRID:SCR_006262) Copy
http://www.grt.kyushu-u.ac.jp/spad/
It is divided to four categories based on extracellular signal molecules (Growth factor, Cytokine, and Hormone) and stress, that initiate the intracellular signaling pathway. SPAD is compiled in order to describe information on interaction between protein and protein, protein and DNA as well as information on sequences of DNA and proteins. There are multiple signal transduction pathways: cascade of information from plasma membrane to nucleus in response to an extracellular stimulus in living organisms. Extracellular signal molecule binds specific intracellular receptor, and initiates the signaling pathway. Now, there is a large amount of information about the signaling pathway which controls the gene expression and cellular proliferation. We have developed an integrated database SPAD to understand the overview of signaling transduction.
Proper citation: Signaling Pathway Database (RRID:SCR_008243) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the RRID Resources search. From here you can search through a compilation of resources used by RRID and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that RRID has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on RRID then you can log in from here to get additional features in RRID such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into RRID you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within RRID that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.