Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://microbialgenomics.energy.gov/index.shtml
Through its Microbial Genome Program (MGP) and its Genomics:GTL (GTL) program, DOEs Office of Biological and Environmental Research (BER) has sequenced more than 485 microbial genomes and 30 microbial communities having specialized biological capabilities. Identifying these genes will help investigators discern how gene activities in whole living systems are orchestrated to solve myriad life challenges. The MGP was begun in 1994 as a spinoff from the Human Genome Program. The goal of the program was to sequence the genomes of a number of nonpathogenic microbes that would be useful in solving DOE''s mission challenges in environmental-waste cleanup, energy production, carbon cycling, and biotechnology. Past projects include microbial genome program, microbial cell project, and the Laboratory Science Program at the DOE Joint Genome Institute. The two ongoing projects are Genomics: GTL program and Community Sequencing Program at the DOE Joint Genome Institute. Sponsors: Site sponsored by the U.S. Department of Energy Office of Science, Office of Biological and Environmental Research, THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Microbial Genomics Program (RRID:SCR_008140) Copy
http://mips.helmholtz-muenchen.de/genre/proj/mpcdb/
A database of manually annotated mammalian protein complexes. To obtain a high-quality dataset, information was extracted from individual experiments described in the scientific literature. Data from high-throughput experiments was not included.
Proper citation: Mammalian Protein Complex Data Base (RRID:SCR_008209) Copy
http://www.sanger.ac.uk/Projects/C_elegans/index.shtml
The Sanger Institute and the Genome Sequencing Center at the Washington University School of Medicine, St. Louis have collaborated to sequence the genomes of both C. elegans and C. briggsae. The completed C. elegans genome sequence is represented by over 3,000 individual clone sequences which can be accessed through this site (or through WormBase). These sequences are submitted to EMBL whenever the sequence or annotation changes (e.g. modification to gene structures) and these submissions are then mirrored to GenBank and DDBJ. These sequences (along with ESTs and proteins) can be searched on our C. elegans BLAST server. WormBase is the repository of mapping, sequencing and phenotypic information for C. elegans. The worm informatics group at the Sanger Institute play a key role in assembling the whole database. They also curate and develop some of the constituent databases that comprise WormBase.
Proper citation: Caenorhabditis Genome Sequencing Projects (RRID:SCR_008155) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 15, 2013. Doodle is a database that was developed to store and distribute information about the protein oligomerization domains that are encoded by various genomes. The protein oligomerization domains described here were found using the lambda repressor fusion system. Doodle uses a schema that is based on EnsEMBL, while also utilizing bioperl modules to both store and retrieve data. The frontend was developed entirely in perl, while the backend utilizes MySQL. GMOD was used to develop the genomic view.
Proper citation: Database of oligomerization domains from lambda experiments (RRID:SCR_008107) Copy
http://www.ebi.ac.uk/genomes/plasmid.html
The Plasmid Genome Database aims to collate biological and genomic data for all bacterial plasmids in the hopes of enabling rapid, interrogation of both meta- and genomic data. Data maintained includes access to all plasmid genomes and information on core genomic features obtained from parsing the original EMBL/DDBJ/NCBI submission. In addition a suite of third party analyses has been performed for each genome to supplement the original annotation. This site also links to Genome Atlases provided by the Centre for Biological Sequence Analysis (CBS). The motivation behind the construction of this site derived from observations from genome sequencing projects: the abundance and inferred importance of the horizontal gene pool (HGP) in bacterial adaptation and evolution. In so far as plasmids are autonomously replicating, extrachromosomal elements they are a readily identifiable and accessible component of the HGP. Also plasmids have been identified in almost all bacterial divisions, ranging in size from less than 2 kbp to > 1.5 Mbp and as such represent a defined, yet diverse and complex sample of genes in the HGP.
Proper citation: Plasmid Genome Database (RRID:SCR_008228) Copy
http://mips.gsf.de/services/genomes/uwe25/
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 15, 2013. This is the official database of the environmental chlamydia genome project. This resource provides access to finished sequence for Parachlamydia-related symbiont UWE25 and to a wide range of manual annotations, automatical analyses and derived datasets. Functional classification and description has been manually annotated according to the Annotation guidelines. Chlamydiae are the major cause of preventable blindness and sexually transmitted disease. Genome analysis of a chlamydia-related symbiont of free-living amoebae revealed that it is twice as large as any of the pathogenic chlamydiae and had few signs of recent lateral gene acquisition. We showed that about 700 million years ago the last common ancestor of pathogenic and symbiotic chlamydiae was already adapted to intracellular survival in early eukaryotes and contained many virulence factors found in modern pathogenic chlamydiae, including a type III secretion system. Ancient chlamydiae appear to be the originators of mechanisms for the exploitation of eukaryotic cells. Environmental chlamydiae have recently been recognized as obligate endosymbionts of free-living amoebae and have been implicated as potential human pathogens. Environmental chlamydiae form a deep branching evolutionary lineage within the medically important order Chlamydiales. Despite their high diversity and ubiquitous distribution in clinical and environmental samples only limited information about genetics and ecology of these microorganisms is available. The Parachlamydia-related Acanthamoeba symbiont UWE25 was therefore selected as representative environmental chlamydia strain for whole genome sequencing. Comparative genome analysis was performed using PEDANT and simap. Sponsors: The environmental chlamydia genome project was funded by the bmb+f (German Federal Ministry of Education and Research) and is part of the Competence Network PathoGenoMiK.
Proper citation: Protochlamydia amoebophila UWE25 (RRID:SCR_008222) Copy
http://chromium.lovd.nl/LOVD2/home.php?select_db=CDKN2A
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 23, 2016. The CDKN2A Database presents the germline and somatic variants of the CDKN2A tumor suppressor gene recorded in human disease through June 2003, annotated with evolutionary, structural, and functional information, in a format that allows the user to either download it or manipulate it for their purposes online. The goal is to provide a database that can be used as a resource by researchers and geneticists and that aids in the interpretation of CDKN2A missense variants. Most online mutation databases present flat files that cannot be manipulated, are often incomplete, and have varying degrees of annotation that may or may not help to interpret the data. They hope to use CDKN2A as a prototype for integrating computational and laboratory data to help interpret variants in other cancer-related genes and other single nucleotide polymorphisms (SNPs) found throughout the genome. Another goal of the lab is to interpret the functional and disease significance of missense variants in cancer susceptibility genes. Eventually, these results will be relevant to the interpretation of single nucleotide polymorphisms (SNPs) in general. The CDKN2A locus is a valuable model for assessing relationships among variation, structure, function, and disease because: Variants of this gene are associated with hereditary cancer: Familial Melanoma (and related syndromes); somatic alterations play a role in carcinogenesis; allelic variants occur whose functional consequences are unknown; reliable functional assays exist; and crystal structure is known. All variants in the database are recorded according to the nomenclature guidelines as outlined by the Human Genome Variation Society. This database is currently designed for research purposes only and is not yet recommended as a clinical resource. Many of the mutations reported here have not been tested for disease association and may represent normal, non-disease causing polymorphisms.
Proper citation: CDKN2A Database (RRID:SCR_008179) Copy
http://ccr.coriell.org/Sections/Collections/HuREF/?SsId=78
The Human Reference Genetic Material Repository makes available DNA from a single individual, J. Craig Venter, whose genome has been sequenced and assembled. The DNA samples are prepared from a lymphoblastoid cell line established at Coriell Cell Repositories from a sample of peripheral blood. The DNA samples are available in 50 microgram aliquots. The lymphoblastoid cell line is not available for distribution. The human DNA sample provided is that of J. Craig Venter whose DNA from white blood cells and sperm was sequenced using Sanger chemistry (ABI Capillary Electrophoresis Platforms 3700 and 3730xl), assembled using the Celera Assembler and was published in PLoS Biology . J. Craig Venter, born on 14 October 1946, is a Caucasian male of self-reported European-American ancestry. The data available on this sample, whose genome assembly is referred to as HuRef, includes: * Whole Genome Shotgun Sequencing data * Sequence trace set deposited by JCVI in the NCBI trace archive * Human Genome Browser displaying sequence assembly, DNA variants and gene annotations Additional data sets from this study include: * Full set of Sanger reads used for genome assembly * SNP and insertion/deletion variant on the human genome sequence coordinates (NCBI version 36) * Affymetrix 500K GeneChip data * Illumina HumanHap650Y Genotyping BeadChip data Given the amount of data publicly available the genomic content of this sample, HuRef will be useful as a reference for many genetic studies.
Proper citation: Human Reference Genetic Material Repository (RRID:SCR_004693) Copy
Part of zebrafish genome project. ZGC project to produce cDNA libraries, clones and sequences to provide complete set of full-length (open reading frame) sequences and cDNA clones of expressed genes for zebrafish. All ZGC sequences are deposited in GenBank and clones can be purchased from distributors of IMAGE consortium. With conclusion of ZGC project in September 2008, GenBank records of ZGC sequences will be frozen, without further updates. Since definition of what constitutes full-length coding region for some of genes and transcripts for which we have ZGC clones will likely change in future, users planning to order ZGC clones will need to monitor for these changes. Users can make use of genome browsers and gene-specific databases, such as UCSC Genome browser, NCBI's Map Viewer, and Entrez Gene, to view relevant regions of genome (browsers) or gene-related information (Entrez Gene).
Proper citation: Zebrafish Gene Collection (RRID:SCR_007054) Copy
http://www.sanger.ac.uk/Projects/D_rerio/zmp/
Create knockout alleles in protein coding genes in the zebrafish genome, using a combination of whole exome enrichment and Illumina next generation sequencing, with the aim to cover them all. Each allele created is analyzed for morphological differences and published on the ZMP site. Transcript counting is performed on alleles with a morphological phenotype. Alleles generated are archived and can be requested from this site through the Zebrafish International Resource Center (ZIRC). You may register to receive updates on genes of interest, or browse a complete list, or search by Ensembl ID, gene name or human and mouse orthologue.
Proper citation: ZMP (RRID:SCR_006161) Copy
http://www.mousephenotype.org/
Center that produces knockout mice and carries out high-throughput phenotyping of each line in order to determine function of every gene in mouse genome. These mice will be preserved in repositories and made available to scientific community representing valuable resource for basic scientific research as well as generating new models for human diseases.
Proper citation: International Mouse Phenotyping Consortium (IMPC) (RRID:SCR_006158) Copy
http://proline.bic.nus.edu.sg/dedb/
Database on Drosophila melanogaster exons presented in a splicing graph form. Data is based on release 3.2 of the Drosophila melanogaster genome annotations available at FlyBase. The gene structure information extracted from the annotations were checked, clustered and transformed into splicing graph. The splicing graph form of the gene constructs were then used for classification of the various types of alternative splicing events. In addition, Pfam domains were mapped onto the gene structure. Users can query the database using the query page using BLAST, FlyBase Gene Name, FlyBase Gene Symbol, Pfam Accession Number and Pfam Identifier. This allows users to determine the Drosophila melanogaster homology of their gene using a BLAST search and to visualize the alternative splicing variants if any. Users can also determine genes containing a particular domain using the Pfam Accession Numbers and Identifiers.
Proper citation: Drosophila melanogaster Exon Database (RRID:SCR_013441) Copy
A genome and functional genomic database for the protozoan parasite Toxoplasma gondii. It incorporates the sequence and annotation of the T. gondii ME49 strain, as well as genome sequences for the GT1, VEG and RH (Chr Ia, Chr Ib) strains. Sequence information is integrated with various other genomic-scale data, including community annotation, ESTs, gene expression and proteomics data. Organisms * Toxoplasma gondii (ME49, RH, GT1, Veg strains) * Neospora caninum * environmental isolate sequences from numerous species Tools * BLAST: Identify Sequence Similarities * Sequence Retrieval: Retrieve Specific Sequences using IDs and coordinates * PubMed and Entrez: View the Latest Toxoplasma, Neospora Pubmed and Entrez Results * Genome Browser: View Sequences and Features in the genome browser * Ancillary Genome Browse: Access Additional info like Probeset data and Toxoplasma Array info
Proper citation: ApiDB ToxoDB (RRID:SCR_013453) Copy
Database that provide a genomic information and comparative genomics platform on sea urchins and related echinoderms. It provide collection of information to directly support experimental work on these useful research models in cell and developmental biology.
Proper citation: EchinoBase (RRID:SCR_013732) Copy
This site has been developed by Kazusa DNA Research Institute for the purpose of offering the science community the analyzed sequence data produced by a multi-national Arabidopsis genome sequencing project coordinated by the Arabidopsis Genome Initiatives (AGI). The aim of this service is to enable users to browse the annotated sequence data produced by all the sequencing teams of AGI through an user-friendly graphic display system and search engines. Gene structures proposed on the annotated sequences as well as those predicted by computer programs are presented and each graphic item has a hyperlink to detailed information of the corresponding area. The nucleotide sequence data deposited in GenBank by AGI was downloaded, re-computer-analyzed at Kazusa and parsed results are displayed graphically.
Proper citation: Kazusa Arabidopsis data opening site (RRID:SCR_013511) Copy
http://www.informatics.jax.org/genes.shtml
Searchable database of mouse genes, DNA segments, cytogenetic markers and QTLs. MGI provides access to integrated data on mouse genes and genome features, from sequences and genomic maps to gene expression and disease models.
Proper citation: Genes, Genome Features and Maps (RRID:SCR_017524) Copy
http://www.mirtoolsgallery.org/miRToolsGallery/node/1055
Comprehensive resource of microRNA target predictions and expression profiles. Used for whole genome prediction of miRNA target genes. For each miRNA, target genes are selected on basis of sequence complementarity using position weighted local alignment algorithm, free energies of RNA-RNA duplexes, and conservation of target sites in related genomes. Provides information about set of genes potentially regulated by particular microRNA, co-occurrence of predicted target sites for multiple microRNAs in mRNA and microRNA expression profiles in tissues. Users are allowed to customize algorithm, numerical parameters, and position-specific rules., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: miRanda (RRID:SCR_017496) Copy
Data collection of large scale genome wide DNA methylation analysis of 1,000 mother-child pairs at serial time points across life course (ARIES).
Proper citation: mqtldb (RRID:SCR_018002) Copy
http://www.clcbio.com/products/clc-main-workbench/
A suite of software for DNA, RNA and protein sequence data analysis. The software allows for the analysis and visualization of Sanger sequencing data as well as gene expression analysis, molecular cloning, primer design, phylogenetic analyses, and sequence data management.
Proper citation: CLC Main Workbench (RRID:SCR_000354) Copy
https://github.com/MicrosoftGenomics/FaST-LMM
FaST-LMM (Factored Spectrally Transformed Linear Mixed Models) is a set of tools for efficiently performing genome-wide association studies (GWAS), prediction, and heritability estimation on large data sets.
Proper citation: FaST LMM (RRID:SCR_015506) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.