Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://agbase.msstate.edu/cgi-bin/tools/goprofiler_select.pl
Service that provides a summary of GO annotations available for each species. The user provides a taxon id and GOProfiler displays the number of GO associations and the number of annotated proteins for that species. The results are listed by evidence code and a separate list of unannotated proteins is also provided.
Proper citation: GOProfiler (RRID:SCR_005683) Copy
http://www.pandora.cs.huji.ac.il/
With PANDORA, you can search for any non-uniform sets of proteins and detect subsets of proteins that share unique biological properties and the intersections of such sets. PANDORA supports GO annotations as well as additional keywords (from UniProt Knowledgebase, InterPro, ENZYME, SCOP etc). It is also integrated into the ProtoNet system, thus allowing testing of thousands of automatically generated protein families. Note that PANDORA replaces the ProtoGO browser developed by the same group. Platform: Online tool
Proper citation: Pandora - Protein ANnotation Diagram ORiented Analysis (RRID:SCR_005686) Copy
http://mcbc.usm.edu/gofetcher/
THIS RESOURCE IS NO LONGER IN SERVICE, documented on June 29, 2012. We developed a web application, GOfetcher, with a very comprehensive search facility for the GO project and a variety of output formats for the results. GOfetcher has three different levels for searching the GO: Quick Search, Advanced Search, and Upload Files for searching. The application includes a unique search option which generates gene information given a nucleotide or protein accession number which can then be used in generating gene ontology information. The output data in GOfetcher can be saved into several different formats; including spreadsheet, comma-separated values, and the Extensible Markup Language (XML) format. Platform: Online tool
Proper citation: GOfetcher (RRID:SCR_005681) Copy
http://www.compbio.dundee.ac.uk/downloads/oxbench/
A suite of programs aimed at developers of alignment methods rather than end-users to assess the accuracy of multiple sequence alignment methods. It includes a reference database of protein multiple sequence alignments that were generated by consideration of protein three-dimensional structure.
Proper citation: OXBench (RRID:SCR_005591) Copy
http://www.garban.org/garban/home.php
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 12, 2012. GARBAN is a tool for analysis and rapid functional annotation of data arising from cDNA microarrays and proteomics techniques. GARBAN has been implemented with bioinformatic tools to rapidly compare, classify, and graphically represent multiple sets of data (genes/ESTs, or proteins), with the specific aim of facilitating the identification of molecular markers in pathological and pharmacological studies. GARBAN has links to the major genomic and proteomic databases (Ensembl, GeneBank, UniProt Knowledgebase, InterPro, etc.), and follows the criteria of the Gene Ontology Consortium (GO) for ontological classifications. Source may be shared: e-mail garban (at) ceit.es. Platform: Online tool
Proper citation: GARBAN (RRID:SCR_005778) Copy
Bioinformatics Resource Center for invertebrate vectors. Provides web-based resources to scientific community conducting basic and applied research on organisms considered potential agents of biowarfare or bioterrorism or causing emerging or re-emerging diseases.
Proper citation: VectorBase (RRID:SCR_005917) Copy
http://www.ebi.ac.uk/Tools/pfa/iprscan/
Software package for functional analysis of sequences by classifying them into families and predicting presence of domains and sites. Scans sequences against InterPro's signatures. Characterizes nucleotide or protein function by matching it with models from several different databases. Used in large scale analysis of whole proteomes, genomes and metagenomes. Available as Web based version and standalone Perl version and SOAP Web Service.
Proper citation: InterProScan (RRID:SCR_005829) Copy
http://www.ebi.ac.uk/webservices/whatizit/info.jsf
A text processing system that allows you to do textmining tasks on text. It is great at identifying molecular biology terms and linking them to publicly available databases. Whatizit is also a Medline abstracts retrieval/search engine. Instead of providing the text by Copy&Paste, you can launch a Medline search. The abstracts that match your search criteria are retrieved and processed by a pipeline of your choice. Whatizit is also available as 1) a webservice and as 2) a streamed servlet. The webservice allows you to enrich content within your website in a similar way as in the wikipedia. The streamed servlet allows you to process large amounts of text.
Proper citation: Whatizit (RRID:SCR_005824) Copy
http://crdd.osdd.net/raghava/ccpdb/
ccPDB (Compilation and Creation of datasets from PDB) is designed to provide service to scientific community working in the field of function or structure annoation of proteins. This database of datasets is based on Protein Data Bank (PDB), where all datasets were derived from PDB. ccPDB have four modules; i) compilation of datasets, ii) creation of datasets, iii) web services and iv) Important links. * Compilation of Datasets: Datasets at ccPDB can be classified in two categories, i) datasets collected from literature and ii) datasets compiled from PDB. We are in process of collecting PDB datasetsfrom literature and maintaining at ccPDB. We are also requesting community to suggest datasets. In addition, we generate datasets from PDB, these datasets were generated using commonly used standard protocols like non-redundant chains, structures solved at high resolution. * Creation of datasets: This module developed for creating customized datasets where user can create a dataset using his/her conditions from PDB. This module will be useful for those users who wish to create a new dataset as per ones requirement. This module have six steps, which are described in help page. * Web Services: We integrated following web services in ccPDB; i) Analyze of PDB ID service allows user to submit their PDB on around 40 servers from single point, ii) BLAST search allows user to perform BLAST search of their protein against PDB, iii) Structural information service is designed for annotating a protein structure from PDB ID, iv) Search in PDB facilitate user in searching structures in PDB, v)Generate patterns service facility to generate different types of patterns required for machine learning techniques and vi) Download useful information allows user to download various types of information for a given set of proteins (PDB IDs). * Important Links: One of major objectives of this web site is to provide links to web servers related to functional annotation of proteins. In first phase we have collected and compiled these links in different categories. In future attempt will be made to collect as many links as possible.
Proper citation: ccPDB - Compilation and Creation of datasets from PDB (RRID:SCR_005870) Copy
It facilitates the search for and dissemination of mass spectra from biologically active metabolites quantified using Gas chromatography (GC) coupled to mass spectrometry (MS). Use the Search Page to search for a compound of your interest, using the name, mass, formula, InChI etc. as query input. Additionally, a Library Search service enables the search of user submitted mass spectra within the GMD. In parallel to the library search, a prediction of chemical sub-groups is performed. This approach has reached beta level and a publication is currently under review. Using several sub-group specific Decision Trees (DTs), mass spectra are classified with respect to the presence of the chemical moieties within the linked (unknown) compound. Prediction of functional groups (ms analysis) facilitates the search of metabolites within the GMD by means of user submitted GC-MS spectra consisting of retention index (n-alkanes, if vailable) and mass intensities ratios. In addition, a functional group prediction will help to characterize those metabolites without available reference mass spectra included in the GMD so far. Instead, the unknown metabolite is characterized by predicted presence or absence of functional groups. For power users this functionality presented here is exposed as soap based web services. Functional group prediction of compounds by means of GC-EI-MS spectra using Microsoft analysis service decision trees All currently available trained decision trees and sub-structure predictions provided by the GMD interface. Table describes the functional group, optional use of an RI system, record date of the trained decision tree, number of MSTs with proportion of MSTs linked to metabolites with the functional group present for each tree. Average and standard deviation of the 50-fold CV error, namely the ratio false over correctly sorted MSTs in the trained DT, are listed. The GMD website offers a range of mass spectral reference libraries to academic users which can be downloaded free of charge in various electronic formats. The libraries are constituted by base peak normalized consensus spectra of single analytes and contain masses in the range 70 to 600 amu, while the ubiquitous mass fragments typically generated from compounds carrying a trimethylsilyl-moiety, namely the fragments at m/z 73, 74, 75, 147, 148, and 149, were excluded.
Proper citation: GMD (RRID:SCR_006625) Copy
The Global Proteome Machine Organization was set up so that scientists involved in proteomics using tandem mass spectrometry could use that data to analyze proteomes. The projects supported by the GPMO have been selected to improve the quality of analysis, make the results portable and to provide a common platform for testing and validating proteomics results. The Global Proteome Machine Database was constructed to utilize the information obtained by GPM servers to aid in the difficult process of validating peptide MS/MS spectra as well as protein coverage patterns. This database has been integrated into GPM server pages, allowing users to quickly compare their experimental results with the best results that have been previously observed by other scientists.
Proper citation: Global Proteome Machine Database (GPM DB) (RRID:SCR_006617) Copy
http://ligand-expo.rutgers.edu/
An integrated data resource for finding chemical and structural information about small molecules bound to proteins and nucleic acids within the structure entries of the Protein Data Bank. Tools are provided to search the PDB dictionary for chemical components, to identify structure entries containing particular small molecules, and to download the 3D structures of the small molecule components in the PDB entry. A sketch tool is also provided for building new chemical definitions from reported PDB chemical components.
Proper citation: Ligand Expo (RRID:SCR_006636) Copy
DPVweb provides a central source of information about viruses, viroids and satellites of plants, fungi and protozoa. Comprehensive taxonomic information, including brief descriptions of each family and genus, and classified lists of virus sequences are provided. The database also holds detailed, curated, information for all sequences of viruses, viroids and satellites of plants, fungi and protozoa that are complete or that contain at least one complete gene. For comparative purposes, it also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA genome. The start and end positions of each feature (gene, non-translated region and the like) have been recorded and checked for accuracy. As far as possible, nomenclature for genes and proteins are standardized within genera and families. Sequences of features (either as DNA or amino acid sequences) can be directly downloaded from the website in FASTA format. The sequence information can also be accessed via client software for PC computers (freely downloadable from the website) that enable users to make an easy selection of sequences and features of a chosen virus for further analyses. The public sequence databases contain vast amounts of data on virus genomes but accessing and comparing the data, except for relatively small sets of related viruses can be very time consuming. The procedure is made difficult because some of the sequences on these databases are incorrectly named, poorly annotated or redundant. The NCBI Reference Sequence project (1) provides a comprehensive, integrated, non-redundant set of sequences, including genomic DNA, transcript (RNA) and protein products, for major research organisms. This now includes curated information for a single sequence of each fully sequenced virus species. While this is a welcome development, it can only deal with complete sequences. An important feature of DPV is the opportunity to access genes (and other features) of multiple sequences quickly and accurately. Thus, for example, it is easy to obtain the nucleotide or amino acid sequences of all the available accessions of the coat protein gene of a given virus species or for a group of viruses. To increase its usefulness further, DPVweb also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA (ssDNA) genome. Sponsors: This site is supported by the Association of Applied Biologists and the Zhejiang Academy of Agricultural Sciences, Hangzhou, People''s Republic of China.
Proper citation: Descriptions of Plant Viruses (RRID:SCR_006656) Copy
Publicly available database of the genes, proteins, experimentally-verified interactions and signaling pathways involved in the innate immune response of humans, mice and bovines to microbial infection. The database captures coverage of the innate immunity interactome by integrating known interactions and pathways from major public databases together with manually-curated data into a centralized resource. The database can be mined as a knowledgebase or used with the integrated bioinformatics and visualization tools for the systems level analysis of the innate immune response. Although InnateDB curation focuses on innate immunity-relevant interactions and pathways, it also incorporates detailed annotation on the entire human, mouse and bovine interactomes by integrating data (178,000+ interactions & 3,900+ pathways) from several of the major public interaction and pathway databases. InnateDB also has integrated human, mouse and bovine orthology predictions generated using Ortholgue software. Ortholgue uses a phylogenetic distance-based method to identify possible paralogs in high-throughput orthology predictions. Integrated human and mouse conserved gene order and synteny information has also been determined to provide further support for orthology predictions. InnateDB Capabilities: * View statistics for manually-curated innate immunity relevant molecular interactions. New manually curated interactions are submitted weekly. * Search for genes and proteins of interest. * Search for experimentally-verified molecular interactions by gene/protein name, interaction type, cell type, etc. * Search genes/interactions belonging to 3,900 pathways. * Visualize interactions using an intuitive subcellular localization-based layout in Cerebral. * Upload your own list of genes along with associated gene expression data (from up to 10 experimental conditions) to interactively analyze this data in a molecular interaction network context. Once you have uploaded your data, you will be able to interactively visualize interaction networks with expression data overlaid; carry out Pathway, Gene Ontology and Transcription Factor Binding Site over-representation analyses; construct orthologous interaction networks in other species; and much more. * Access curated interaction data via a dedicated PSICQUIC webservice.
Proper citation: InnateDB (RRID:SCR_006714) Copy
http://inparanoid.sbc.su.se/cgi-bin/index.cgi
Collection of pairwise comparisons between 100 whole genomes generated by a fully automatic method for finding orthologs and in-paralogs between TWO species. Ortholog clusters in the InParanoid are seeded with a two-way best pairwise match, after which an algorithm for adding in-paralogs is applied. The method bypasses multiple alignments and phylogenetic trees, which can be slow and error-prone steps in classical ortholog detection. Still, it robustly detects complex orthologous relationships and assigns confidence values for in-paralogs. The original data sets can be downloaded.
Proper citation: InParanoid: Eukaryotic Ortholog Groups (RRID:SCR_006801) Copy
Multi-organism, publicly accessible compendium of peptides identified in a large set of tandem mass spectrometry proteomics experiments. Mass spectrometer output files are collected for human, mouse, yeast, and several other organisms, and searched using the latest search engines and protein sequences. All results of sequence and spectral library searching are subsequently processed through the Trans Proteomic Pipeline to derive a probability of correct identification for all results in a uniform manner to insure a high quality database, along with false discovery rates at the whole atlas level. The raw data, search results, and full builds can be downloaded for other uses. All results of sequence searching are processed through PeptideProphet to derive a probability of correct identification for all results in a uniform manner ensuring a high quality database. All peptides are mapped to Ensembl and can be viewed as custom tracks on the Ensembl genome browser. The long term goal of the project is full annotation of eukaryotic genomes through a thorough validation of expressed proteins. The PeptideAtlas provides a method and a framework to accommodate proteome information coming from high-throughput proteomics technologies. The online database administers experimental data in the public domain. You are encouraged to contribute to the database.
Proper citation: PeptideAtlas (RRID:SCR_006783) Copy
canSAR is an integrated database that brings together biological, chemical, pharmacological (and eventually clinical) data. Its goal is to integrate this data and make it accessible to cancer research scientists from multiple disciplines, in order to help with hypothesis generation in cancer research and support translational research. This cancer research and drug discovery resource was developed to utilize the growing publicly available biological annotation, chemical screening, RNA interference screening, expression, amplification and 3D structural data. Scientists can, in a single place, rapidly identify biological annotation of a target, its structural characterization, expression levels and protein interaction data, as well as suitable cell lines for experiments, potential tool compounds and similarity to known drug targets. canSAR has, from the outset, been completely use-case driven which has dramatically influenced the design of the back-end and the functionality provided through the interfaces. The Web interface provides flexible, multipoint entry into canSAR. This allows easy access to the multidisciplinary data within, including target and compound synopses, bioactivity views and expert tools for chemogenomic, expression and protein interaction network data.
Proper citation: canSAR (RRID:SCR_006794) Copy
Portal to the PSORT family of computer programs for the prediction of protein localization sites in cells, as well as other datasets and resources relevant to localization prediction. The standalone versions are available for download for larger analyses.
Proper citation: Psort (RRID:SCR_007038) Copy
Comprehensive set of protein domain families automatically generated from UniProt Knowledge Database. Automated clustering of homologous domains generated from global comparison of all available protein sequences., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: ProDom (RRID:SCR_006969) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented on August 27, 2019.
Database for those interested in the consequences of Factor VIII genetic variation at the DNA and protein level, it provides access to data on the molecular pathology of haemophilia A. The database presents a review of the structure and function of factor VIII and the molecular genetics of haemophilia A, a real time update of the biostatistics of each parameter in the database, a molecular model of the A1, A2 and A3 domains of the factor VIII protein (based on the crystal structure of caeruloplasmin) and a bulletin board for discussion of issues in the molecular biology of factor VIII. The database is completely updated with easy submission of point mutations, deletions and insertions via e-mail of custom-designed forms. A methods section devoted to mutation detection is available, highlighting issues such as choice of technique and PCR primer sequences. The FVIII structure section now includes a download of a FVIII A domain homology model in Protein Data Bank format and a multiple alignment of the FVIII amino-acid sequences from four species (human, murine, porcine and canine) in addition to the virtual reality simulations, secondary structural data and FVIII animation already available. Finally, to aid navigation across this site, a clickable roadmap of the main features provides easy access to the page desired. Their intention is that continued development and updating of the site shall provide workers in the fields of molecular and structural biology with a one-stop resource site to facilitate FVIII research and education. To submit your mutants to the Haemophilia A Mutation Database email the details. (Refer to Submission Guidelines)
Proper citation: HAMSTeRS - The Haemophilia A Mutation Structure Test and Resource Site (RRID:SCR_006883) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.