Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://compgen.bscb.cornell.edu/phast/
A freely available software package for comparative and evolutionary genomics that consists of about half a dozen major programs, plus more than a dozen utilities for manipulating sequence alignments, phylogenetic trees, and genomic annotations. For the most part, PHAST focuses on two kinds of applications: the identification of novel functional elements, including protein-coding exons and evolutionarily conserved sequences; and statistical phylogenetic modeling, including estimation of model parameters, detection of signatures of selection, and reconstruction of ancestral sequences. It consists of over 60,000 lines of C code.
Proper citation: PHAST (RRID:SCR_003204) Copy
http://gene3d.biochem.ucl.ac.uk/Gene3D/
A large database of CATH protein domain assignments for ENSEMBL genomes and Uniprot sequences. Gene3D is a resource of form studying proteins and the component domains. Gene3D takes CATH domains from Protein Databank (PDB) structures and assigns them to the millions of protein sequences with no PDB structures using Hidden Markov models. Assigning a CATH superfamily to a region of a protein sequence gives information on the gross 3D structure of that region of the protein. CATH superfamilies have a limited set of functions and so the domain assignment provides some functional insights. Furthermore most proteins have several different domains in a specific order, so looking for proteins with a similar domain organization provides further functional insights. Strict confidence cut-offs are used to ensure the reliability of the domain assignments. Gene3D imports functional information from sources such as UNIPROT, and KEGG. They also import experimental datasets on request to help researchers integrate there data with the corpus of the literature. The website allows users to view descriptions for both single proteins and genes and large protein sets, such as superfamilies or genomes. Subsets can then be selected for detailed investigation or associated functions and interactions can be used to expand explorations to new proteins. The Gene3D web services provide programmatic access to the CATH-Gene3D annotation resources and in-house software tools. These services include Gene3DScan for identifying structural domains within protein sequences, access to pre-calculated annotations for the major sequence databases, and linked functional annotation from UniProt, GO and KEGG., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Gene3D (RRID:SCR_007672) Copy
http://hapmap.ncbi.nlm.nih.gov/
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 22, 2016. A multi-country collaboration among scientists and funding agencies to develop a public resource where genetic similarities and differences in human beings are identified and catalogued. Using this information, researchers will be able to find genes that affect health, disease, and individual responses to medications and environmental factors. All of the information generated by the Project will be released into the public domain. Their goal is to compare the genetic sequences of different individuals to identify chromosomal regions where genetic variants are shared. Public and private organizations in six countries are participating in the International HapMap Project. Data generated by the Project can be downloaded with minimal constraints. HapMap project related data, software, and documentation include: bulk data on genotypes, frequencies, LD data, phasing data, allocated SNPs, recombination rates and hotspots, SNP assays, Perlegen amplicons, raw data, inferred genotypes, and mitochondrial and chrY haplogroups; Generic Genome Browser software; protocols and information on assay design, genotyping and other protocols used in the project; and documentation of samples/individuals and the XML format used in the project.
Proper citation: International HapMap Project (RRID:SCR_002846) Copy
http://www.broadinstitute.org/gsea/
Software package for interpreting gene expression data. Used for interpretation of a large-scale experiment by identifying pathways and processes.
Proper citation: Gene Set Enrichment Analysis (RRID:SCR_003199) Copy
The European resource for the collection, organization and dissemination of data on biological macromolecular structures. In collaboration with the other worldwide Protein Data Bank (wwPDB) partners - the Research Collaboratory for Structural Bioinformatics (RCSB) and BioMagResBank (BMRB) in the USA and the Protein Data Bank of Japan (PDBj) - they work to collate, maintain and provide access to the global repository of macromolecular structure data. The main objectives of the work at PDBe are: * to provide an integrated resource of high-quality macromolecular structures and related data and make it available to the biomedical community via intuitive user interfaces. * to maintain in-house expertise in all the major structure-determination techniques (X-ray, NMR and EM) in order to stay abreast of technical and methodological developments in these fields, and to work with the community on issues of mutual interest (such as data representation, harvesting, formats and standards, or validation of structural data). * to provide high-quality deposition and annotation facilities for structural data as one of the wwPDB deposition sites. Several sophisticated tools are also available for the structural analysis of macromolecules.
Proper citation: PDBe - Protein Data Bank in Europe (RRID:SCR_004312) Copy
http://www.metabolomicsworkbench.org
Repository for metabolomics data and metadata which provides analysis tools and access to various resources. NIH grantees may upload data and general users can search metabolomics database. Provides protocols for sample preparation and analysis, information about NIH Metabolomics Program, data sharing guidelines, funding opportunities, services offered by its Regional Comprehensive Metabolomics Resource Cores (RCMRC)s, and training workshops.
Proper citation: Metabolomics Workbench (RRID:SCR_013794) Copy
Open source Java based image processing software program designed for scientific multidimensional images. ImageJ has been transformed to ImageJ2 application to improve data engine to be sufficient to analyze modern datasets.
Proper citation: ImageJ (RRID:SCR_003070) Copy
https://panoramaweb.org/project/home/begin.view?
Repository software for targeted mass spectrometry assays from Skyline. Targeted proteomics knowledge base. Public repository for quantitative data sets processed in Skyline. Facilitates viewing, sharing, and disseminating results contained in Skyline documents.
Proper citation: PanoramaWeb (RRID:SCR_017136) Copy
Integrated genomic and functional genomic database for Entamoeba and Acanthamoeba parasites. Contains genomes of three Entamoeba species and microarray expression data for E. histolytica. Integrates whole genome sequence and annotation and includes experimental data and environmental isolate sequences provided by community researchers.
Proper citation: AmoebaDB (RRID:SCR_017592) Copy
https://github.com/galaxyproteomics/mvpapplication-git.git
Software tool as plugin to enable viewing of results produced from workflows integrating genomic sequencing data and mass spectrometry proteomics data. Plugin to Galaxy bioinformatics workbench which enables visualization of mass spectrometry-based proteomics data integrated with genomic and/or transcriptomic sequencing data. Useful for verifying quality of results and characterizing novel peptide sequences identified using multi-omic proteogenomic approach.
Proper citation: Multi-omics Visualization Platform (RRID:SCR_018077) Copy
http://bioinformatics.hungry.com/clearcut/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023.Software as a stand-alone reference implementation for the Relaxed Neighbor Joining (RNJ) algorithm. Used in distance-based phylogenetic tree reconstruction method to process large sequence datasets., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Clearcut (RRID:SCR_016059) Copy
https://niaid.github.io/spice/
Software application for data mining and visualization. Used for analyzes of large FLOWJO data sets from polychromatic flow cytometry and organizing the normalized data graphically.
Proper citation: SPICE (RRID:SCR_016603) Copy
Software tool for genome and metagenome distance estimation using MinHash. Reduces large sequences and sequence sets to small, representative sketches, from which global mutation distances can be rapidly estimated.
Proper citation: Mash (RRID:SCR_019135) Copy
http://www.neuroepigenomics.org/methylomedb/
A database containing genome-wide brain DNA methylation profiles for human and mouse brains. The DNA methylation profiles were generated by Methylation Mapping Analysis by Paired-end Sequencing (Methyl-MAPS) method and analyzed by Methyl-Analyzer software package. The methylation profiles cover over 80% CpG dinucleotides in human and mouse brains in single-CpG resolution. The integrated genome browser (modified from UCSC Genome Browser allows users to browse DNA methylation profiles in specific genomic loci, to search specific methylation patterns, and to compare methylation patterns between individual samples. Two species were included in the Brain Methylome Database: human and mouse. Human postmortem brain samples were obtained from three distinct cortical regions, i.e., dorsal lateral prefrontal cortex (dlPFC), ventral prefrontal cortex (vPFC), and auditory cortex (AC). Human samples were selected from our postmortem brain collection with extensive neuropathological and psychopathological data, as well as brain toxicology reports. The Department of Psychiatry of Columbia University and the New York State Psychiatric Institute have assembled this brain collection, where a validated psychological autopsy method is used to generate Axis I and II DSM IV diagnoses and data are obtained on developmental history, history of psychiatric illness and treatment, and family history for each subject. The mouse sample (strain 129S6/SvEv) DNA was collected from the entire left cerebral hemisphere. The three human brain regions were selected because they have been implicated in the neuropathology of depression and schizophrenia. Within each cortical region, both disease and non-psychiatric samples have been profiled (matching subjects by age and sex in each group). Such careful matching of subjects allows one to perform a wide range of queries with the ability to characterize methylation features in non-psychiatric controls, as well as detect differentially methylated domains or features between disease and non-psychiatric samples. A total of 14 non-psychiatric, 9 schizophrenic, and 6 depression methylation profiles are included in the database.
Proper citation: MethylomeDB (RRID:SCR_005583) Copy
http://llama.mshri.on.ca/funcassociate/
A web-based tool that accepts as input a list of genes, and returns a list of GO attributes that are over- (or under-) represented among the genes in the input list. Only those over- (or under-) representations that are statistically significant, after correcting for multiple hypotheses testing, are reported. Currently 37 organisms are supported. In addition to the input list of genes, users may specify a) whether this list should be regarded as ordered or unordered; b) the universe of genes to be considered by FuncAssociate; c) whether to report over-, or under-represented attributes, or both; and d) the p-value cutoff. A new version of FuncAssociate supports a wider range of naming schemes for input genes, and uses more frequently updated GO associations. However, some features of the original version, such as sorting by LOD or the option to see the gene-attribute table, are not yet implemented. Platform: Online tool
Proper citation: FuncAssociate: The Gene Set Functionator (RRID:SCR_005768) Copy
Web-based microarray data analysis and visualization system powered by CRC, or Chinese Restaurant cluster, a Dirichlet process model-based clustering algorithm recently developed by Dr. Steve Qin. It also incorporates several gene expression analysis programs from Bioconductor, including GOStats, genefilter, and Heatplus. CRCView also installs from the Bioconductor system 78 annotation libraries of microarray chips for human (31), mouse (24), rat (14), zebrafish (1), chicken (1), Drosophila (3), Arabidopsis (2), Caenorhabditis elegans (1), and Xenopus Laevis (1). CRCView allows flexible input data format, automated model-based CRC clustering analysis, rich graphical illustration, and integrated Gene Ontology (GO)-based gene enrichment for efficient annotation and interpretation of clustering results. CRC has the following features comparing to other clustering tools: 1) able to infer number of clusters, 2) able to cluster genes displaying time-shifted and/or inverted correlations, 3) able to tolerate missing genotype data and 4) provide confidence measure for clusters generated. You need to register for an account in the system to store your data and analyses. The data and results can be visited again anytime you log in.
Proper citation: CRCView (RRID:SCR_007092) Copy
https://run.biosimulations.org
Web tool for executing broad range of modeling studies and visualizing their results. Provides web interface for reusing any model. Models, simulations, and visualizations are available under licenses specified for each resource.
Proper citation: runBioSimulations (RRID:SCR_019110) Copy
http://ftp://ftp.ncbi.nlm.nih.gov/pub/mhc/rbc/Final Archive
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 23, 2019.BGMUT was database that provided publicly accessible platform for DNA sequences and curated set of blood mutation information. Data Archive are available at ftp://ftp.ncbi.nlm.nih.gov/pub/mhc/rbc/Final Archive.
Proper citation: Blood Group Antigen Gene Mutation Database (RRID:SCR_002297) Copy
http://www.ncbi.nlm.nih.gov/biosystems/
Database that provides access to biological systems and their component genes, proteins, and small molecules, as well as literature describing those biosystems and other related data throughout Entrez. A biosystem, or biological system, is a group of molecules that interact directly or indirectly, where the grouping is relevant to the characterization of living matter. BioSystem records list and categorize components, such as the genes, proteins, and small molecules involved in a biological system. The companion FLink tool, in turn, allows you to input a list of proteins, genes, or small molecules and retrieve a ranked list of biosystems. A number of databases provide diagrams showing the components and products of biological pathways along with corresponding annotations and links to literature. This database was developed as a complementary project to (1) serve as a centralized repository of data; (2) connect the biosystem records with associated literature, molecular, and chemical data throughout the Entrez system; and (3) facilitate computation on biosystems data. The NCBI BioSystems Database currently contains records from several source databases: KEGG, BioCyc (including its Tier 1 EcoCyc and MetaCyc databases, and its Tier 2 databases), Reactome, the National Cancer Institute's Pathway Interaction Database, WikiPathways, and Gene Ontology (GO). It includes several types of records such as pathways, structural complexes, and functional sets, and is desiged to accomodate other record types, such as diseases, as data become available. Through these collaborations, the BioSystems database facilitates access to, and provides the ability to compute on, a wide range of biosystems data. If you are interested in depositing data into the BioSystems database, please contact them.
Proper citation: NCBI BioSystems Database (RRID:SCR_004690) Copy
Open source software package for comparative sequence analysis using stochastic evolutionary models. Used for analysis of genetic sequence data in particular the inference of natural selection using techniques in phylogenetics, molecular evolution, and machine learning.
Proper citation: HyPhy (RRID:SCR_016162) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.