Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Alignment software for large-scale protein contact or protein-protein interaction prediction optimized for speed through shorter runtimes. FreeContact provides the opportunity to compute contact predictions in any environment (desktop or cloud).
Proper citation: FreeContact (RRID:SCR_016113) Copy
http://iclab.life.nctu.edu.tw/iclab_webtools/sodock/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on July 31,2025. An optimization algorithm based on particle swarm optimization (PSO) for solving flexible protein-ligand docking problems.
Proper citation: SODOCK (RRID:SCR_000193) Copy
Web server for flexible protein structure comparison. Structure alignment is formulated as the aligned fragment pairs chaining process allowing at most t twists, and the flexible structure alignment is transformed into a rigid structure alignment when t is forced to be 0., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: FATCAT (RRID:SCR_014631) Copy
A computer algorithm to predict aggregation nucleating regions in proteins as well the effect of mutations and environmental conditions on the aggregation propensity of these regions.
Proper citation: TANGO (RRID:SCR_001770) Copy
https://bitbucket.org/nsegata/phylophlan/wiki/Home
Software pipeline for reconstructing highly accurate and resolved phylogenetic trees based on whole-genome sequence information. Pipeline is scalable to thousands of genomes and uses the most conserved 400 proteins for extracting the phylogenetic signal. PhyloPhlAn also implements taxonomic curation, estimation, and insertion operations., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: PhyloPhlAn (RRID:SCR_013082) Copy
http://www.cbs.dtu.dk/services/ProP/
Web application which predicts arginine and lysine propeptide cleavage sites in eukaryotic protein sequences using an ensemble of neural networks. Furin-specific prediction is the default. It is also possible to perform a general proprotein convertase prediction.
Proper citation: ProP Server (RRID:SCR_014936) Copy
Collection of genome databases for vertebrates and other eukaryotic species with DNA and protein sequence search capabilities. Used to automatically annotate genome, integrate this annotation with other available biological data and make data publicly available via web. Ensembl tools include BLAST, BLAT, BioMart and the Variant Effect Predictor (VEP) for all supported species.
Proper citation: Ensembl (RRID:SCR_002344) Copy
A web-based software package for comparative genomics.
Proper citation: Sybil (RRID:SCR_005593) Copy
Curated, open-source, integrated data resource for comparative functional genomics in crops and model plant species to facilitate the study of cross-species comparisons using information generated from projects supported by public funds. It currently hosts annotated whole genomes in over two dozen plant species and partial assemblies for almost a dozen wild rice species in the Ensembl browser, genetic and physical maps with genes, ESTs and QTLs locations, genetic diversity data sets, structure-function analysis of proteins, plant pathways databases (BioCyc and Plant Reactome platforms), and descriptions of phenotypic traits and mutations. The web-based displays for phenotypes include the Genes and Quantitative Trait Loci (QTL) modules. Sequence based relationships are displayed in the Genomes module using the genome browser adapted from Ensembl, in the Maps module using the comparative map viewer (CMap) from GMOD, and in the Proteins module displays. BLAST is used to search for similar sequences. Literature supporting all the above data is organized in the Literature database. In addition, Gramene now hosts a variety of web services including a Distributed Annotation Server (DAS), BLAST and a public MySQL database. Twice a year, Gramene releases a major build of the database and makes interim releases to correct errors or to make important updates to software and/or data. Additionally you can access Gramene through an FTP site.
Proper citation: Gramene (RRID:SCR_002829) Copy
Model organism database that provides organization of and access to scientific data for the fission yeast Schizosaccharomyces pombe. PomBase supports genomic sequence and features, genome-wide datasets and manual literature curation. PomBase also provides a community hub for researchers, providing genome statistics, a community curation interface, news, events, documentation, mailing lists, and welcomes data submissions.
Proper citation: PomBase (RRID:SCR_006586) Copy
Database that provides a collection of transmembrane, monotopic and peripheral proteins from the Protein Data Bank whose spatial arrangements in the lipid bilayer have been calculated theoretically and compared with experimental data. The database allows analysis, sorting and searching of membrane proteins based on their structural classification, species, destination membrane, numbers of transmembrane segments and subunits, numbers of secondary structures and the calculated hydrophobic thickness or tilt angle with respect to the bilayer normal.
Proper citation: Orientations of Proteins in Membranes database (RRID:SCR_011961) Copy
http://www.ihop-net.org/UniPub/iHOP/
Information system that provides a network of concurring genes and proteins extends through the scientific literature touching on phenotypes, pathologies and gene function. It provides this network as a natural way of accessing millions of PubMed abstracts. By using genes and proteins as hyperlinks between sentences and abstracts, the information in PubMed can be converted into one navigable resource, bringing all advantages of the internet to scientific literature research. Moreover, this literature network can be superimposed on experimental interaction data (e.g., yeast-two hybrid data from Drosophila melanogaster and Caenorhabditis elegans) to make possible a simultaneous analysis of new and existing knowledge. The network contains half a million sentences and 30,000 different genes from humans, mice, D. melanogaster, C. elegans, zebrafish, Arabidopsis thaliana, yeast and Escherichia coli.
Proper citation: Information Hyperlinked Over Proteins (RRID:SCR_004829) Copy
Database of apo and holo structure pairs of proteins before and after binding. Various protein functions have been shown directly associated with conformational transitions triggered by binding other molecules. Tertiary structures determined in the unbound and bound state are usually named apo and holo structures, respectively. AH-DB is the largest database of apo-holo structure pairs and provides a sophisticated interface to search and view the collected data. It contains 746314 apo-holo pairs of 3638 proteins from 702 organisms.
Proper citation: Apo and Holo structures DataBase (RRID:SCR_004800) Copy
http://prorepeat.bioinformatics.nl/
ProRepeat is an integrated curated repository and analysis platform for in-depth research on the biological characteristics of amino acid tandem repeats. ProRepeat collects repeats from all proteins included in the UniProt knowledgebase, together with 85 completely sequenced eukaryotic proteomes contained within the RefSeq collection. It contains non-redundant perfect tandem repeats, approximate tandem repeats and simple, low-complexity sequences, covering the majority of the amino acid tandem repeat patterns found in proteins. The ProRepeat web interface allows querying the repeat database using repeat characteristics like repeat unit and length, number of repetitions of the repeat unit and position of the repeat in the protein. Users can also search for repeats by the characteristics of repeat containing proteins, such as entry ID, protein description, sequence length, gene name and taxon. ProRepeat offers powerful analysis tools for finding biological interesting properties of repeats, such as the strong position bias of leucine repeats in the N-terminus of eukaryotic protein sequences, the differences of repeat abundance among proteomes, the functional classification of repeat containing proteins and GC content constrains of repeats' corresponding codons.
Proper citation: ProRepeat (RRID:SCR_006113) Copy
The database of protein-chemical structural interactions includes all existing 3D structures of complexes of proteins with low molecular weight ligands. When one considers the proteins and chemical vertices of a graph, all these interactions form a network. Biological networks are powerful tools for predicting undocumented relationships between molecules. The underlying principle is that existing interactions between molecules can be used to predict new interactions. For pairs of proteins sharing a common ligand, we use protein and chemical superimpositions combined with fast structural compatibility screens to predict whether additional compounds bound by one protein would bind the other. The current version includes data from the Protein Data Bank as of August 2011. The database is updated monthly.
Proper citation: ProtChemSI (RRID:SCR_006115) Copy
http://prism.ccbb.ku.edu.tr/hotregion/index.php
Hot spots are energetically important residues at protein interfaces and they are not randomly distributed across the interface but rather clustered. These clustered hot spots form hot regions. Hot regions are important for the stability of protein complexes, as well as providing specificity to binding sites. HotRegion provides the hot region information of the interfaces by using predicted hot spot residues, and structural properties of these interface residues such as pair potentials of interface residues, accessible surface area (ASA) and relative ASA values of interface residues of both monomer and complex forms of proteins. Also, the 3D visualization of the interface and interactions among hot spot residues are provided. The number of interfaces in the database is 147909 and still growing.
Proper citation: HotRegion - A Database of Cooperative Hotspots (RRID:SCR_006022) Copy
DOMMINO is a comprehensive structural database on macromolecular interactions. As of June, 2011, it contains more than 407,000 binary interactions. The distinctive features of DOMMINO are: # Automated updates: DOMMINO is fully automated and is designed to update itself on a weekly basis, one day after a PDB weekly update. Thus, the community will be able to study macromolecular interactions almost immediately after they are released by PDB. # Coverage of non-domain mediated interactions: In addition to domain-domain and domain-peptide interactions the database characterizes the interaction between domains and unstructured protein regions that are not parts of a domain, such as inter-domain linkers and N- and C-termini. The interactions that involve the latter unstructured parts of proteins have been included to the database for the first time providing additional ~186,000 interactions (~45% of the total number of interactions, as of June, 2011). # Coverage of new structural domains: DOMMINO employs one of the most accurate structural classifications of proteins, SCOP. In addition to the existing SCOP-annotated domains, we employ a state-of-the-art machine learning approach to classify newer protein structures into existing SCOP families. With the progress of structural genomics, we do not expect a significant growth of the number of structurally novel folds or protein families and therefore our method allows covering almost all new protein structures. In total, using this predictive approach has allowed us to add more than 261,000 new interactions, almost twice as many as existing SCOP-annotated interactions. # The web-interface is designed to give the user a possibility of a flexible search as well as the capability to study macromolecular interactions in a PDB structure at the interaction network level and at the individual interface level. The web interface of the DOMMINO database includes a comprehensive list of help topics linked to the specific actions. In addition, we have designed a step-by-step tutorial that covers all aspects of working with the data from DOMMINO using the web interface.
Proper citation: DOMMINO - Database Of MacroMolecular INteractiOns (RRID:SCR_005958) Copy
http://www.jcvi.org/charprotdb/index.cgi/home
The Characterized Protein Database, CharProtDB, is designed and being developed as a resource of expertly curated, experimentally characterized proteins described in published literature. For each protein record in CharProtDB, storage of several data types is supported. It includes functional annotation (several instances of protein names and gene symbols) taxonomic classification, literature links, specific Gene Ontology (GO) terms and GO evidence codes, EC (Enzyme Commisssion) and TC (Transport Classification) numbers and protein sequence. Additionally, each protein record is associated with cross links to all public accessions in major protein databases as ��synonymous accessions��. Each of the above data types can be linked to as many literature references as possible. Every CharProtDB entry requires minimum data types to be furnished. They are protein name, GO terms and supporting reference(s) associated to GO evidence codes. Annotating using the GO system is of importance for several reasons; the GO system captures defined concepts (the GO terms) with unique ids, which can be attached to specific genes and the three controlled vocabularies of the GO allow for the capture of much more annotation information than is traditionally captured in protein common names, including, for example, not just the function of the protein, but its location as well. GO evidence codes implemented in CharProtDB directly correlate with the GO consortium definitions of experimental codes. CharProtDB tools link characterization data from multiple input streams through synonymous accessions or direct sequence identity. CharProtDB can represent multiple characterizations of the same protein, with proper attribution and links to database sources. Users can use a variety of search terms including protein name, gene symbol, EC number, organism name, accessions or any text to search the database. Following the search, a display page lists all the proteins that match the search term. Click on the protein name to view more detailed annotated information for each protein. Additionally, each protein record can be annotated.
Proper citation: CharProtDB: Characterized Protein Database (RRID:SCR_005872) Copy
http://pbildb1.univ-lyon1.fr/virhostnet/
Public knowledge base specialized in the management and analysis of integrated virus-virus, virus-host and host-host interaction networks coupled to their functional annotations. It contains high quality and up-to-date information gathered and curated from public databases (VirusMint, Intact, HIV-1 database). It allows users to search by host gene, host/viral protein, gene ontology function, KEGG pathway, Interpro domain, and publication information. It also allows users to browse viral taxonomy.
Proper citation: VirHostNet: Virus-Host Network (RRID:SCR_005978) Copy
Scansite searches for motifs within proteins that are likely to be phosphorylated by specific protein kinases or bind to domains such as SH2 domains, 14-3-3 domains or PDZ domains. The Motifscanner program utilizes an entropy approach that assesses the probability of a site matching the motif using the selectivity values and sums the logs of the probability values for each amino acid in the candidate sequence. The program then indicates the percentile ranking of the candidate motif in respect to all potential motifs in proteins of a protein database. When available, percentile scores of some confirmed phosphorylation sites for the kinase of interests or confirmed binding sites of the domain of interest are provided for comparison with the scores of the candidate motifs.
Proper citation: Scansite (RRID:SCR_007026) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.