Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://probeexplorer.cicancer.org/principal.php
Probe Explorer is an open access web-based bioinformatics application designed to show the association between microarray oligonucleotide probes and transcripts in the genomic context, but flexible enough to serve as a simplified genome and transcriptome browser. Coordinates and sequences of the genomic entities (loci, exons, transcripts), including vector graphics outputs, are provided for fifteen metazoa organisms and two yeasts. Alignment tools are used to built the associations between Affymetrix microarrays probe sequences and the transcriptomes (for human, mouse, rat and yeasts). Search by keywords is available and user searches and alignments on the genomes can also be done using any DNA or protein sequence query. Platform: Online tool
Proper citation: ProbeExplorer (RRID:SCR_007116) Copy
http://www.broadinstitute.org/annotation/tetraodon/
This database have been funded by the National Human Genome Research Institute (NHGRI) to produce shotgun sequence of the Tetraodon nigriviridis genome. The strategy involves Whole Genome Shotgun (WGS) sequencing, in which sequence from the entire genome is generated. Whole genome shotgun libraries were prepared from Tetraodon genomic DNA obtained from the laboratory of Jean Weissenbach at Genoscope. Additional sequence data of approximately 2.5X coverage of Tetraodon has also been generated by Genoscope in plasmid and BAC end reads. Broad and Genoscope intend to pool their data and generate whole genome assemblies. Tetraodon nigroviridis is a freshwater pufferfish of the order Tetraodontiformes and lives in the rivers and estuaries of Indonesia, Malaysia and India. This species is 20-30 million years distant from Fugu rubripes, a marine pufferfish from the same family. The gene repertoire of T. nigroviridis is very similar to that of other vertebrates. However, its relatively small genome of 385 Mb is eight times more compact than that of human, mostly because intergenic and intronic sequences are reduced in size compared to other vertebrate genomes. These genome characteristics along with the large evolutionary distance between bony fish and mammals make Tetraodon a compact vertebrate reference genome - a powerful tool for comparative genetics and for quick and reliable identification of human genes.
Proper citation: Tetraodon nigroviridis Database (RRID:SCR_007123) Copy
http://www.sugp.caltech.edu/SpBase/
SpBase is designed to present the results of the genome sequencing project for the purple sea urchin. The sequences and annotations emerging from this effort are organized in a database that provides the research community access to those data not normally presented through National Center for Biotechnology Information and other large databases. Additionally, the unique information on that links gene identities and sequences to the plate and well location to the library filters from the Sea Urchin genome Resource will also be presented. The software used to organize and present the sea urchin genome comes from GMOD, a collection of open source software tools for creating and managing genome-scale biological databases. That sea urchins eggs and embryos have long remained a popular research subject for cell and developmental biologists is one rationale for sequencing the genome. In addition, studies of embryonic development in the California Purple Sea Urchin, Strongylocentrotus purpuratus , have paralleled the emergence of molecular techniques ranging from the characterization of genomic repeat sequences in the 1970''s to the elucidation of gene regulatory networks in recent times. The parent of this site, SUGP, was meant to provide a focal point for the exchange of genomic information as the genome of the Purple sea urchin was being sequenced. Over these past years it has served as a repository for small sequencing projects and a source of sequence information useful for gene discovery projects. Here one could find information on macro-array libraries of cDNAs from the purple sea urchin and genomic DNA from several species. In addition, a Sequence Tag Connector (STC) collection has been assembled from 5% of the genome sequence and a very extensive repeat sequence catalog prepared. All of the sequence data that we maintained at SUGP was incorporated into the new SPBase. Of course, it is all in public sequence databases such as the National Center for Biological Information as well. Some additional sequence information is available at the Resource Center of the German Human Genome Project. With the publication of The Genome of the Sea Urchin Strongylocentrotus purpuratus by The Sea Urchin Genome Sequencing Consortium a link to the first 9941 gene annotations are now publicly available. The effort to sequence the whole purple sea urchin genome was a cooperative one that included contributions from the Sea Urchin Genome Facility here at the Center for Computational Regulatory Genomics, Beckman Institute, Caltech, and support from the Human Genome Research Institute of the National Institutes of Health. The sequencing was done at the Baylor College of Medicine, Human Genome Sequencing Center, Houston, Texas. Funding was approved based on an initiative submitted by the Sea Urchin Genome Advisory Committee.
Proper citation: SpBase - Strongylocentrotus purpuratus: the Sea Urchin Genome Database (RRID:SCR_007441) Copy
A database of human, chimpanzee, mouse, and rat proteases and protease inhibitors, as well as as the growing number of hereditary diseases caused by mutations in protease genes. Analysis of the human and mouse genomes has allowed us to annotate 581 human, 580 chimpanzee, 667 mouse, and 655 rat protease genes. Proteases are classified in five different classes according to their mechanism of catalysis. Proteases are a diverse and important group of enzymes representing >2% of the human, chimpanzee, mouse and rat genomes. This group of enzymes is implicated in numerous physiological processes. The importance of proteases is illustrated by the existence of 99 different hereditary diseases due to mutations in protease genes. Furthermore, proteases have been implicated in multiple human pathologies, including vascular diseases, rheumatoid arthritis, neurodegenerative processes, and cancer. During the last ten years, our laboratory has identified and characterized more than 60 human protease genes. Due to the importance of proteolytic enzymes in human physiology and pathology, we have recently introduced the concept of Degradome, as the complete repertoire of proteases expressed by a tissue or organism. Thanks to the recent completion of the human, chimpanzee, mouse, and rat genome sequencing projects, we were able to analyze and compare for the first time the complete protease repertoire in those mammalian organisms, as well as the complement of protease inhibitor genes. This webpage also contains the Supplementary Material of Human and mouse proteases: a comparative genomic approach Nat Rev Genet (2003) 4: 544-558, Genome sequence of the brown Norway rat yields insights into mammalian evolution Nature (2004) 428: 493-521, A genomic analysis of rat proteases and protease inhibitors Genome Res. (2004) 14: 609-622, and Comparative genomic analysis of human and chimpanzee proteases Genomics (2005) 86: 638-647.
Proper citation: Mammalian Degradome Database (RRID:SCR_007624) Copy
http://www.ncbi.nlm.nih.gov/genomes/GenomesHome.cgi?taxid=2759&hopt=html
Curated sequence data and related information on organelles from NCBI Refseq for the community to use as a standard. The animal mitochondrial records are considered reviewed; that is, they have been manually curated by the NCBI staff. Other mitochondrial and chloroplast genome records are provisional and are presented with varying levels of review compared to the primary record used to build the RefSeq. Additionally, protein clusters for the metazoan and plastid genomes proteins can be reviewed with Entrez Protein Clusters.
Proper citation: Organelle Genome Resources (RRID:SCR_007838) Copy
Alternative splicing essentially increases the diversity of the transcriptome and has important implications for physiology, development and the genesis of diseases. This resource uses a different approach to investigate alternative splicing (instead of the conventional case-by case fashion) and integrates all transcripts derived from a gene into a single splicing graph. ASG is a database of splicing graphs for human genes, using transcript information from various major sources (Ensembl, RefSeq, STACK, TIGR and UniGene). Each transcript corresponds to a path in the graph, and alternative splicing is displayed by bifurcations. This representation preserves the relationships between different splicing variants and allows us to investigate systematically all possible putative transcripts. Web interface allows users to display the splicing graphs, to interactively assemble transcripts and to access their sequences as well as neighboring genomic regions. ASG also provide for each gene, an exhaustive pre-computed catalog of putative transcriptsin total more than 1.2 million sequences. It has found that ~65 of the investigated genes show evidence for alternative splicing, and in 5 of the cases, a single gene might produce over 100 transcripts.
Proper citation: Alternate splicing gallery (RRID:SCR_008129) Copy
https://wiki.cgb.indiana.edu/display/DGC/Home
The Daphnia Genomics Consortium (DGC) is an international network of investigators committed to mounting the freshwater crustacean Daphnia as a model system for ecology, evolution and the environmental sciences. Along with research activities, the DGC is: (1) coordinating efforts towards developing the Daphnia genomic toolbox, which will then be available for use by the general community; (2) facilitating collaborative cross-disciplinary investigations; (3) developing bioinformatic strategies for organizing the rapidly growing genome database; and (4) exploring emerging technologies to improve high throughput analyses of molecular and ecological samples. If we are to succeed in creating a new model system for modern life-sciences research, it will need to be a community-wide effort. Research activities of the DGC are primarily focused on creating genomic tools and information. When completed, the current projects will offer a first view of the Daphnia genome''s topography, including regions of high and low recombination, the distribution of transposable, repetitive and regulatory elements, the size and structure of genes and of their neighborhoods. This information is crucial in formulating testable hypotheses relating genetics and demographics to the evolutionary potential or constraints of natural populations. Projects aiming to compile identifiable genes with their function are also underway, together with robust methods to verify these findings. Finally, these tools are being tested, by exploring their uses in key ecological and toxicological investigations. Each project benefits from the leadership and expertise of many individuals. For further details, begin by contacting the project directors. The DGC consists of biologists from a broad spectrum of subdisciplines, including limnology, ecotoxicology, quantitative and population genetics, systematics, molecular biology and evolution, developmental biology, genomics and bioinformatics. In many regards, the rapid early success of the consortium results from its grass-roots origin promoting an international composition, under a cooperative model, with significant scientific breadth. We hold to this approach in building this network and encourage more people to participate. All the while, the DGC is structured to effectively reach specific goals. The consortium includes an advisory board (composed of experts of the various subdisciplines), whose responsibility is to act as the research community''s agent in guiding the development of Daphnia genomic resources. The advisors communicate directly to DGC members, who are either contributing genomic tools or actively seeking funds for this function. The consortium''s main body (given the widespread interest in applying genomic tools in environmental studies) are the affiliates, who make use of these tools for their research and who are soliciting support.
Proper citation: Daphnia genomics consortium (RRID:SCR_008148) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 29, 2016. The BayGenomics gene-trap resource provides researchers with access to thousands of mouse embryonic stem (ES) cell lines harboring characterized insertional mutations in both known and novel genes. The major goal of BayGenomics is to identify genes relevant to cardiovascular and pulmonary disease.
Proper citation: BayGenomics (RRID:SCR_008168) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented on August 20,2019.The COG-database has become a powerful tool in the field of comparative genomics. The construction of this data-base is based on sequence homologies of proteins from different completely sequenced genomes. Highly homologous proteins are assigned to clusters of orthologous groups. The updated collection of orthologous protein sets for prokaryotes and eukaryotes is expected to be a useful platform for functional annotation of newly sequenced genomes, including those of complex eukaryotes, and genome-wide evolutionary studies. The availability of multiple, essentially complete genome sequences of prokaryotes and eukaryotes spurred both the demand and the opportunity for the construction of an evolutionary classification of genes from these genomes. Such a classification system based on orthologous relationships between genes appears to be a natural framework for comparative genomics and should facilitate both functional annotation of genomes and large-scale evolutionary studies. Here is a major update of the previously developed system for delineation of Clusters of Orthologous Groups of proteins (COGs) from the sequenced genomes of prokaryotes and unicellular eukaryotes and the construction of clusters of predicted orthologs for 7 eukaryotic genomes, which we named KOGs after eukaryotic orthologous groups. The COG collection currently consists of 138,458 proteins, which form 4873 COGs and comprise 75% of the 185,505 (predicted) proteins encoded in 66 genomes of unicellular organisms. The eukaryotic orthologous groups (KOGs) include proteins from 7 eukaryotic genomes: three animals (the nematode Caenorhabditis elegans, the fruit fly Drosophila melanogaster and Homo sapiens), one plant, Arabidopsis thaliana, two fungi (Saccharomyces cerevisiae and Schizosaccharomyces pombe), and the intracellular microsporidian parasite Encephalitozoon cuniculi. The current KOG set consists of 4852 clusters of orthologs, which include 59,838 proteins, or approximately 54% of the analyzed eukaryotic 110,655 gene products. Compared to the coverage of the prokaryotic genomes with COGs, a considerably smaller fraction of eukaryotic genes could be included into the KOGs; addition of new eukaryotic genomes is expected to result in substantial increase in the coverage of eukaryotic genomes with KOGs. Examination of the phyletic patterns of KOGs reveals a conserved core represented in all analyzed species and consisting of approximately 20% of the KOG set. This conserved portion of the KOG set is much greater than the ubiquitous portion of the COG set (approximately 1% of the COGs). In part, this difference is probably due to the small number of included eukaryotic genomes, but it could also reflect the relative compactness of eukaryotes as a clade and the greater evolutionary stability of eukaryotic genomes.
Proper citation: Phylogenetic Clusters of Orthologous Groups Ranking (RRID:SCR_008223) Copy
http://www.nisc.nih.gov/projects/comp_seq.html
Generates data for use in developing and refining computational tools for comparing genomic sequence from multiple species. The NISC Comparative Sequencing Program's goal is to establish a data resource consisting of sequences for the same set of targeted genomic regions derived from multiple animal species. The broader program includes plans for a diverse set of analytical studies using the generated sequence and the publication of a series of papers describing the results of those analysis in peer-reviewed journals in a timely fashion. Experimentally, this project involves the shotgun sequencing of mapped BAC clones. For each BAC, an assembly is first performed when a sufficient number of sequence reads have been generated to provide full shotgun coverage of the clone. At that time, the assembled sequence is submitted to the HTGS division of GenBank. Subsequent refinements of the sequence, including the generation of higher-accuracy finished sequence, results in the updating of the sequence record in GenBank. By immediately submitting our BAC-derived sequences to GenBank, it makes their data available as a public service to allow colleagues to speed up their research, consistent with the now well-established routine of sequencing centers participating in the Human Genome Project. However, at the same time, it has made considerable investment in acquiring these mapping and sequence data, including sizable efforts of graduate students, postdoctoral fellows, and other trainees. Furthermore, in most cases, large data sets involving multiple BAC sequences from multiple species must first be generated, often taking many months to accumulate, before the planned analysis can be performed and the resulting papers written and submitted for publication.
Proper citation: Comparative Vertebrate Sequencing (RRID:SCR_008213) Copy
The aim of the PEROXISOME database (PeroxisomeDB) is to gather, organize and integrate curated information on peroxisomal genes, their encoded proteins, their molecular function and metabolic pathway they belong to, and their related disorders. PeroxisomeDB contains the complete peroxisomal proteome of Homo sapiens (encoded by 85 genes) and Saccharomyces cerevisiae (encoded by 61 genes). Now, we have included 34 new organism genomes with the acquisition of 2426 new peroxisomal homolog proteins. PeroxisomeDB 2.0 integrates the peroxisomal metabolome of whole microbody family by the new incorporation of the glycosome proteomes of trypanosomatids and the glyoxysome proteome of Arabidopsis thaliana. The site also provides a Peroxisome Metabolome of peroxisomal genes and proteins, their molecular interactions and metabolic pathways, tools for comparative genomics, predictive tools. Sponsors: Preoxisome Database is funded by Institut de Gntique et deBiologie Molculaire et Cellulaire.
Proper citation: Peroxisome Database (RRID:SCR_008352) Copy
http://cgap.nci.nih.gov/Chromosomes/Mitelman
The web site includes genomic data for humans and mice, including transcript sequence, gene expression patterns, single-nucleotide polymorphisms, clone resources, and cytogenetic information. Descriptions of the methods and reagents used in deriving the CGAP datasets are also provided. An extensive suite of informatics tools facilitates queries and analysis of the CGAP data by the community. One of the newest features of the CGAP web site is an electronic version of the Mitelman Database of Chromosome Aberrations in Cancer. The data in the Mitelman Database is manually culled from the literature and subsequently organized into three distinct sub-databases, as follows: -The sub-database of cases contains the data that relates chromosomal aberrations to specific tumor characteristics in individual patient cases. It can be searched using either the Cases Quick Searcher or the Cases Full Searcher. -The sub-database of molecular biology and clinical associations contains no data from individual patient cases. Instead, the data is pulled from studies with distinct information about: -Molecular biology associations that relate chromosomal aberrations and tumor histologies to genomic sequence data, typically genes rearranged as a consequence of structural chromosome changes. -Clinical associations that relate chromosomal aberrations and/or gene rearrangements and tumor histologies to clinical variables, such as prognosis, tumor grade, and patient characteristics. It can be searched using the Molecular Biology and Clinical (MBC) Associations Searcher -The reference sub-database contains all the references culled from the literature i.e., the sum of the references from the cases and the molecular biology and clinical associations. It can be searched using the Reference Searcher. CGAP has developed six web search tools to help you analyze the information within the Mitelman Database: -The Cases Quick Searcher allows you to query the individual patient cases using the four major fields: aberration, breakpoint, morphology, and topography. -The Cases Full Searcher permits a more detailed search of the same individual patient cases as above, by including more cytogenetic field choices and adding search fields for patient characteristics and references. -The Molecular Biology Associations Searcher does not search any of the individual patient cases. It searches studies pertaining to gene rearrangements as a consequence of cytogenetic aberrations. -The Clinical Associations Searcher does not search any of the individual patient cases. It searches studies pertaining to clinical associations of cytogenetic aberrations and/or gene rearrangements. -The Recurrent Chromosome Aberrations Searcher provides a way to search for structural and numerical abnormalities that are recurrent, i.e., present in two or more cases with the same morphology and topography. -The Reference Searcher queries only the references themselves, i.e., the references from the individual cases and the molecular biology and clinical associations. Sponsors: This database is sponsored by the University of Lund, Sweden and have support from the Swedish Cancer Society and the Swedish Children''s Cancer Foundation
Proper citation: Mitelman Database of Chromosome Aberrations in Cancer (RRID:SCR_012877) Copy
http://www.informatics.jax.org/
Community model organism database for laboratory mouse and authoritative source for phenotype and functional annotations of mouse genes. MGD includes complete catalog of mouse genes and genome features with integrated access to genetic, genomic and phenotypic information, all serving to further the use of the mouse as a model system for studying human biology and disease. MGD is a major component of the Mouse Genome Informatics.Contains standardized descriptions of mouse phenotypes, associations between mouse models and human genetic diseases, extensive integration of DNA and protein sequence data, normalized representation of genome and genome variant information. Data are obtained and integrated via manual curation of the biomedical literature, direct contributions from individual investigators and downloads from major informatics resource centers. MGD collaborates with the bioinformatics community on the development and use of biomedical ontologies such as the Gene Ontology (GO) and the Mammalian Phenotype (MP) Ontology.
Proper citation: Mouse Genome Database (RRID:SCR_012953) Copy
http://cello.life.nctu.edu.tw/
A subCELlular LOcalization predictor based on a multi-class support vector machine (SVM) classification system. CELLO uses 4 types of sequence coding schemes: the amino acid composition, the di-peptide composition, the partitioned amino acid composition and the sequence composition based on the physico-chemical properties of amino acids. They combine votes from these classifiers and use the jury votes to determine the final assignment.
Proper citation: CELLO (RRID:SCR_011968) Copy
https://cran.r-project.org/web/packages/LDheatmap/index.html
Software application that plots measures of pairwise linkage disequilibria for SNPs (entry from Genetic Analysis Software)
Proper citation: LDHEATMAP (RRID:SCR_006312) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on May 12,2023. An integrated packages of tools for microarray data analysis. GEPAS provides a web-based interface that offers diverse analysis options from the early step of preprocessing (normalization of Affymetrix and two-color microarray experiments and other preprocessing options), to the final step of the functional profiling of the experiment (using Gene Ontology, pathways, PubMed abstracts etc.), which include different possibilities for clustering, gene selection, class prediction and array-comparative genomic hybridization management.
Proper citation: Gene Expression Profile Analysis Suite (RRID:SCR_008341) Copy
http://wpicr.wpic.pitt.edu/WPICCompGen/hclust/hclust.htm
Software application that is a simple clustering method that can be used to rapidly identify a set of tag SNP's based upon genotype data (entry from Genetic Analysis Software), THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: HCLUST (RRID:SCR_009154) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on May 12,2023. Database of expression patterns of C. elegans promoter::GFP constructs. A text description of the observed pattern is provided, indicating the stage(s) and tissue(s) in which GFP is expressed. Also available for some strains are the corresponding 2D and 3D images. Investigators may browse the entire list, search by gene name, tissue, stage, and pattern. Search results may be downloaded in .csv and .txt formats. All of the strains in the expression pattern database are displayed in the browse page. The records are organized by gene; information such as locus name, genomic location (WormBase), the presence of images and videos, and the actual expression pattern are shown in a tabular format.
Proper citation: Expression Patterns for C. elegans promoter GFP fusions (RRID:SCR_001619) Copy
http://hanalyzer.sourceforge.net/
An open-source data integration system designed to assist biologists in explaining the results observed in genome-scale experiments as well as generating new hypotheses. It combines information extraction techniques, semantic data integration, and reasoning and facilitates network visualization. The Hanalyzer source code and binaries are available for download.
Proper citation: Hanalyzer (RRID:SCR_000923) Copy
A genomics data analysis platform which generates decision models for healthcare organizations and medical research. This service is meant to utilize data through machine learning methods.
Proper citation: iOMICS (RRID:SCR_000239) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.