Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
DPVweb provides a central source of information about viruses, viroids and satellites of plants, fungi and protozoa. Comprehensive taxonomic information, including brief descriptions of each family and genus, and classified lists of virus sequences are provided. The database also holds detailed, curated, information for all sequences of viruses, viroids and satellites of plants, fungi and protozoa that are complete or that contain at least one complete gene. For comparative purposes, it also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA genome. The start and end positions of each feature (gene, non-translated region and the like) have been recorded and checked for accuracy. As far as possible, nomenclature for genes and proteins are standardized within genera and families. Sequences of features (either as DNA or amino acid sequences) can be directly downloaded from the website in FASTA format. The sequence information can also be accessed via client software for PC computers (freely downloadable from the website) that enable users to make an easy selection of sequences and features of a chosen virus for further analyses. The public sequence databases contain vast amounts of data on virus genomes but accessing and comparing the data, except for relatively small sets of related viruses can be very time consuming. The procedure is made difficult because some of the sequences on these databases are incorrectly named, poorly annotated or redundant. The NCBI Reference Sequence project (1) provides a comprehensive, integrated, non-redundant set of sequences, including genomic DNA, transcript (RNA) and protein products, for major research organisms. This now includes curated information for a single sequence of each fully sequenced virus species. While this is a welcome development, it can only deal with complete sequences. An important feature of DPV is the opportunity to access genes (and other features) of multiple sequences quickly and accurately. Thus, for example, it is easy to obtain the nucleotide or amino acid sequences of all the available accessions of the coat protein gene of a given virus species or for a group of viruses. To increase its usefulness further, DPVweb also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA (ssDNA) genome. Sponsors: This site is supported by the Association of Applied Biologists and the Zhejiang Academy of Agricultural Sciences, Hangzhou, People''s Republic of China.
Proper citation: Descriptions of Plant Viruses (RRID:SCR_006656) Copy
Portal to the PSORT family of computer programs for the prediction of protein localization sites in cells, as well as other datasets and resources relevant to localization prediction. The standalone versions are available for download for larger analyses.
Proper citation: Psort (RRID:SCR_007038) Copy
Comprehensive set of protein domain families automatically generated from UniProt Knowledge Database. Automated clustering of homologous domains generated from global comparison of all available protein sequences., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: ProDom (RRID:SCR_006969) Copy
http://www.genoscope.cns.fr/externe/tetraodon/
The initial objective of Genoscope was to compare the genomic sequences of this fish to that of humans to help in the annotation of human genes and to estimate their number. This strategy is based on the common genetic heritage of the vertebrates: from one species of vertebrate to another, even for those as far apart as a fish and a mammal, the same genes are present for the most part. In the case of the compact genome of Tetraodon, this common complement of genes is contained in a genome eight times smaller than that of humans. Although the length of the exons is similar in these two species, the size of the introns and the intergenic sequences is greatly reduced in this fish. Furthermore, these regions, in contrast to the exons, have diverged completely since the separation of the lineages leading to humans and Tetraodon. The Exofish method, developed at Genoscope, exploits this contrast such that the conserved regions which can be identified by comparing genomic sequences of the two species, correspond only to coding regions. Using preliminary sequencing results of the genome of Tetraodon in the year 2000, Genoscope evaluated the number of human genes at about 30,000, whereas much higher estimations were current. The progress of the annotation of the human genome has since supported the Genoscope hypothesis, with values as low as 22,000 genes and a consensus of around 25,000 genes. The sequencing of the Tetraodon genome at a depth of about 8X, carried out as a collaboration between Genoscope and the Whitehead Institute Center for Genome Research (now the Broad Institute), was finished in 2002, with the production of an assembly covering 90 of the euchromatic region of the genome of the fish. This has permitted the application of Exofish at a larger scale in comparisons with the genome of humans, but also with those of the two other vertebrates sequenced at the time (Takifugu, a fish closely related to Tetraodon, and the mouse). The conserved regions detected in this way have been integrated into the annotation procedure, along with other resources (cDNA sequences from Tetraodon and ab initio predictions). Of the 28,000 genes annotated, some families were examined in detail: selenoproteins, and Type 1 cytokines and their receptors. The comparison of the proteome of Tetraodon with those of mammals has revealed some interesting differences, such as a major diversification of some hormone systems and of the collagen molecules in the fish. A search for transposable elements in the genomic sequences of Tetraodon has also revealed a high diversity (75 types), which contrasts with their scarcity; the small size of the Tetraodon genome is due to the low abundance of these elements, of which some appear to still be active. Another factor in the compactness of the Tetraodon genome, which has been confirmed by annotation, is the reduction in intron size, which approaches a lower limit of 50-60 bp, and which preferentially affects certain genes. The availability of the sequences from the genomes of humans and mice on one hand, and Takifugu and Tetraodon on the other, provide new opportunities for the study of vertebrate evolution. We have shown that the level of neutral evolution is higher in fish than in mammals. The protein sequences of fish also diverge more quickly than those of mammals. A key mechanism in evolution is gene duplication, which we have studied by taking advantage of the anchoring of the majority of the sequences from the assembly on the chromosomes. The result of this study speaks strongly in favor of a whole genome duplication event, very early in the line of ray-finned fish (Actinopterygians). An even stronger evidence came from synteny studies between the genomes of humans and Tetraodon. Using a high-resolution synteny map, we have reconstituted the genome of the vertebrate which predates this duplication - that is, the last common ancestor to all bony vertebrates (most of the vertebrates apart from cartilaginous fish and agnaths like lamprey). This ancestral karyotype contains 12 chromosomes, and the 21 Tetraodon chromosomes derive from it by the whole genome duplication and a surprisingly small number of interchromosomal rearrangements. On the contrary, exchanges between chromosomes have been much more frequent in the lineage that leads to humans. Sponsors: The project was supported by the Consortium National de Recherche en Genomique and the National Human Genome Research Institute.
Proper citation: Tetraodon Genome Browser (RRID:SCR_007079) Copy
http://medgen.ugent.be/rtprimerdb/
Database for primer and probe sequences used in real-time PCR assays employing popular chemistries (SYBR Green I, Taqman, Hybridization Probes, Molecular Beacon) to prevent time-consuming primer design and experimental optimization, and to introduce a certain level of uniformity and standardization among different laboratories. Researchers are encouraged to submit their validated primer and probe sequence, so that other users can benefit from their expertise. The database can be queried using the official gene name or symbol, Entrez or Ensembl Gene identifier, SNP identifier, or oligonucleotide sequence. Different options make it possible to restrict a query to a particular application (Gene Expression Quantification/Detection, DNA Copy Number Quantification/Detection, SNP Detection, Mutation Analysis, Fusion Gene Quantification/Detection, Chromatin immunoprecipitation (ChIP)), organism (Human, Mouse, Rat, and others) or detection chemistry.
Proper citation: RTPrimerDB- The Real-Time PCR and Probe Database (RRID:SCR_007106) Copy
http://weizhong-lab.ucsd.edu/cd-hit/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: CD-HIT (RRID:SCR_007105) Copy
http://www.scied.com/pr_cmbas.htm
A software system to assist with cloning simulation, enzyme operations, and graphic map drawing. Clone Manager can also be used as a way to view or edit sequence files, find open reading frames, translate genes, or find genes or text in files. Clone Manager Professional is an upgraded version of Clone Manager Basic., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Clone Manager Software (RRID:SCR_014521) Copy
http://aem.asm.org/content/71/12/8228.full
THIS RESOURCE IS NO LONGER IN SERVICE, documented Setember 8, 2016. A suite of tools for the comparison of microbial communities using phylogenetic information. It takes as input a single phylogenetic tree that contains sequences derived from at least two different environmental samples and a file describing which sequences came from which sample.
Proper citation: Unifrac (RRID:SCR_014616) Copy
http://www.cbcb.umd.edu/software/metapath
A statistical package for comparing metagenomic data-sets at the pathway level. It relies on a combination of metagenomic sequence data and prior metabolic pathway knowledge, which is pulled from KEGG.
Proper citation: Metapath (RRID:SCR_014621) Copy
https://sourceforge.net/projects/soapdenovo2/files/GapCloser/
Module of SOAPdenovo2 commonly used independently to close gaps in genome assemblies.
Proper citation: GapCloser (RRID:SCR_015026) Copy
Database for ESTs (Expressed Sequence Tags), consensus sequences, bacterial artificial chromosome (BAC) clones, BES (BAC End Sequences). They have generated 69,545 ESTs from 6 full-length cDNA libraries (Porcine Abdominal Fat, Porcine Fat Cell, Porcine Loin Muscle, Liver and Pituitary gland). They have also identified a total of 182 BAC contigs from chromosome 6. It is very valuable resources to study porcine quantitative trait loci (QTL) mapping and genome study. Users can explore genomic alignment of various data types, including expressed sequence tags (ESTs), consensus sequences, singletons, QTL, Marker, UniGene and BAC clones by several options. To estimate the genomic location of sequence dataset, their data aligned BES (BAC End Sequences) instead of genomic sequence because Pig Genome has low-coverage sequencing data. Sus scrofa Genome Database mainly provide comparative map of four species (pig, cattle, dog and mouse) in chromosome 6.
Proper citation: PiGenome (RRID:SCR_013394) Copy
http://chgv.org/GenicIntolerance/
A gene-based score intended to help in the interpretation of human sequence data. The score is designed to rank genes in terms of whether they have more or less common functional genetic variation relative to the genome wide expectation given the amount of apparently neutral variation the gene has. A gene with a positive score has more common functional variation, and a gene with a negative score has less and is referred to as intolerant.
Proper citation: Residual Variation Intolerance Score (RVIS) (RRID:SCR_013850) Copy
http://bio-bwa.sourceforge.net/
Software for aligning sequencing reads against large reference genome. Consists of three algorithms: BWA-backtrack, BWA-SW and BWA-MEM. First for sequence reads up to 100bp, and other two for longer sequences ranged from 70bp to 1Mbp.
Proper citation: BWA (RRID:SCR_010910) Copy
http://www.glycosciences.de/tools/linucs/
Service that directly converts the commonly used extended representation of complex carbohydrates into the preferred canonical description or into its inverted form. Input: A structure using the extended, non-graphic nomenclature (in ASCII writing) to describe complex carbohydrates as recommended by IUPAC. Output: A linear, unique notation. The source code (written in C), will be distributed so that software developers can easily implement their algorithm within their own application. LINUCS was chosen to fulfill to following conditions: * Input of extended, non-graphic nomenclature to describe carbohydrate structures. * Resulting linear code is closely related to notations and abbreviations recommended by IUPAC. * Number of additional rules to define the priority of the branches is low * Extended nomenclature of complex carbohydrates contains all information to define the hierarchy. * LINUCS is applicable to all types of carbohydrates (macrocyclic system are currently not implemented) . * Remaining unassigned linkage information are tolerated
Proper citation: LINUCS (RRID:SCR_001571) Copy
A curated collection of chaperonin sequence data collected from public databases or generated by a network of collaborators exploiting the cpn60 target in clinical, phylogenetic and microbial ecology studies. The database contains all available sequences for both group I and group II chaperonins. Users can search the database by Chaperonin type, group (I or II), BLAST, or other options, and can also enter and analyze FASTA sequences.
Proper citation: cpnDB: A Chaperonin Database (RRID:SCR_002263) Copy
http://ww2.sanbi.ac.za/Dbases.html
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 23, 2016. The STACKdb is knowledgebase generated by processing EST and mRNA sequences obtained from GenBank through a pipeline consisting of masking, clustering, alignment and variation analysis steps. The STACK project aims to generate a comprehensive representation of the sequence of each of the expressed genes in the human genome by extensive processing of gene fragments to make accurate alignments, highlight diversity and provide a carefully joined set of consensus sequences for each gene. The STACK project is comprised of the STACKdb human gene index, a database of virtual human transcripts, as well as stackPACK, the tools used to create the database. STACKdb is organized into 15 tissue-based categories and one disease category. STACK is a tool for detection and visualization of expressed transcript variation in the context of developmental and pathological states. The data system organizes and reconstructs human transcripts from available public data in the context of expression state. The expression state of a transcript can include developmental state, pathological association, site of expression and isoform of expressed transcript. STACK consensus transcripts are reconstructed from clusters that capture and reflect the growing evidence of transcript diversity. The comprehensive capture of transcript variants is achieved by the use of a novel clustering approach that is tolerant of sub-sequence diversity and does not rely on pairwise alignment. This is in contrast with other gene indexing projects. STACK is generated at least four times a year and represents the exhaustive processing of all publicly available human EST data extracted from GenBank. This processed information can be explored through 15 tissue-specific categories, a disease-related category and a whole-body index
Proper citation: Sequence Tag Alignment and Consensus Knowledgebase Database (RRID:SCR_002156) Copy
The Hepatitis C Virus (HCV) Database Project strives to present HCV-associated genetic and immunologic data in a user-friendly way, by providing access to the central database via web-accessible search interfaces and supplying a number of analysis tools.
Proper citation: HCV Databases (RRID:SCR_002863) Copy
Professionally curated repository for genetics, genomics and related data resources for soybean that contains the most current genetic, physical and genomic sequence maps integrated with qualitative and quantitative traits. SoyBase includes annotated Williams 82 genomic sequence and associated data mining tools. The genetic and sequence views of the soybean chromosomes and the extensive data on traits and phenotypes are extensively interlinked. This allows entry to the database using almost any kind of available information, such as genetic map symbols, soybean gene names or phenotypic traits. The repository maintains controlled vocabularies for soybean growth, development, and traits that are linked to more general plant ontologies. Contributions to SoyBase or the Breeder''s Toolbox are welcome.
Proper citation: SoyBase (RRID:SCR_005096) Copy
Issue
Software package for analysis of brain imaging data sequences. Sequences can be a series of images from different cohorts, or time-series from same subject. Current release is designed for analysis of fMRI, PET, SPECT, EEG and MEG.
Proper citation: SPM (RRID:SCR_007037) Copy
https://github.com/cmayer/BaitFisher-package
Software toolkit for multispecies target DNA enrichment probe design. It consists of two programs: BaitFisher and BaitFilter, which are designed to construct hybrid enrichment baits for multiple sequence alignments or annotated features in multiple sequence alignments.
Proper citation: Baitfisher (RRID:SCR_015985) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the RRID Resources search. From here you can search through a compilation of resources used by RRID and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that RRID has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on RRID then you can log in from here to get additional features in RRID such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into RRID you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within RRID that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.