Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://www.ihop-net.org/UniPub/iHOP/
Information system that provides a network of concurring genes and proteins extends through the scientific literature touching on phenotypes, pathologies and gene function. It provides this network as a natural way of accessing millions of PubMed abstracts. By using genes and proteins as hyperlinks between sentences and abstracts, the information in PubMed can be converted into one navigable resource, bringing all advantages of the internet to scientific literature research. Moreover, this literature network can be superimposed on experimental interaction data (e.g., yeast-two hybrid data from Drosophila melanogaster and Caenorhabditis elegans) to make possible a simultaneous analysis of new and existing knowledge. The network contains half a million sentences and 30,000 different genes from humans, mice, D. melanogaster, C. elegans, zebrafish, Arabidopsis thaliana, yeast and Escherichia coli.
Proper citation: Information Hyperlinked Over Proteins (RRID:SCR_004829) Copy
Database that collects all arabidopsis transcription factors (totally 1922 Loci; 2290 Gene Models) and classifies them into 64 families. It uses not only locus (gene), but also gene model (transcript, protein) and the detail information is for each gene model not for locus. It adds multiple alignment of the DNA-binding domain of each family, Neighbor-Joining phylogenetic tree of each family, the GO annotation, homolog with the Database of Rice Transcription Factors (DRTF). It also keeps old information items such as the unique cloned and sequenced information of about 1200 transcription factors, protein domains, 3D structure information with BLAST hits against PDB, predicted Nuclear Location Signals, UniGene information, as well as links to literature reference.
Proper citation: Database of Arabidopsis Transcription Factors (RRID:SCR_007101) Copy
Cross-species microarray expression database focusing on high-throughput expression data relevant for germline development, meiosis and gametogenesis as well as the mitotic cell cycle. The database contains a unique combination of information: 1) High-throughput expression data obtained with whole-genome high-density oligonucleotide microarrays (GeneChips). 2) Sample annotation (mouse over the sample name and click on it) using the Multiomics Information Management and Annotation System (MIMAS 3.0). 3) In vivo protein-DNA binding data and protein-protein interaction data (available for selected species). 4) Genome annotation information from Ensembl version 50. 5) Orthologs are identified using data from Ensembl and OMA and linked to each other via a section in the report pages. The portal provides access to the Saccharomyces Genomics Viewer (SGV) which facilitates online interpretation of complex data from experiments with high-density oligonucleotide tiling microarrays that cover the entire yeast genome. The database displays only expression data obtained with high-density oligonucleotide microarrays (GeneChips)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 15,2026.
Proper citation: GermOnline (RRID:SCR_002807) Copy
A database of high-quality protein-protein interactions in different organisms.
Proper citation: HINT (RRID:SCR_002762) Copy
Web-based tool for the ontological analysis of large lists of genes. It can be used to determine biological annotations or combinations of annotations that are significantly associated to a list of genes under study with respect to a reference list. As well as single annotations, this tool allows users to simultaneously evaluate annotations from different sources, for example Biological Process and Cellular Component categories of Gene Ontology., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: GeneCodis (RRID:SCR_006943) Copy
http://biodev.extra.cea.fr/interoporc/
Automatic prediction tool to infer protein-protein interaction networks, it is applicable for lots of species using orthology and known interactions. The interoPORC method is based on the interolog concept and combines source interaction datasets from public databases as well as clusters of orthologous proteins (PORC) available on Integr8. Users can use this page to ask InteroPorc for all species present in Integr8. Some results are already computed and users can run InteroPorc to investigate any other species. Currently, the following databases are processed and merged (with datetime of the last available public release for each database used): IntAct, MINT, DIP, and Integr8.
Proper citation: InteroPorc (RRID:SCR_002067) Copy
A web-based tool that provides composite interpretations for microarray data comparing two sample groups as well as lists of genes from diverse sources of biological information. It provides multiple gene set analysis methods for microarray inputs as well as enrichment analyses for lists of genes. It screens redundant composite annotations when generating and prioritizing them. It also incorporates union and subtracted sets as well as intersection sets. Users can upload their gene sets (e.g. predicted miRNA targets) to generate and analyze new composite sets.
Proper citation: ADGO (RRID:SCR_006343) Copy
Curated protein-protein and genetic interaction repository of raw protein and genetic interactions from major model organism species, with data compiled through comprehensive curation efforts.
Proper citation: Biological General Repository for Interaction Datasets (BioGRID) (RRID:SCR_007393) Copy
http://zope.bioinfo.cnio.es/plan2l/plan2l.html
A web-based online search system that integrates text mining and information extraction techniques to access systematically information useful for analyzing genetic, cellular and molecular aspects of the plant model organism Arabidopsis thaliana. The system facilitates a more efficient retrieval of information relevant to heterogeneous biological topics, from implications in biological relationships at the level of protein interactions and gene regulation, to sub-cellular locations of gene products and associations to cellular and developmental processes, i.e. cell cycle, flowering, root, leaf and seed development. Beyond single entities, also predefined pairs of entities can be provided as queries for which literature-derived relations together with textual evidences are returned.
Proper citation: PLAN2L (RRID:SCR_013346) Copy
DNAtraffic database is dedicated to be an unique comprehensive and richly annotated database of genome dynamics during the cell life. DNAtraffic contains extensive data on the nomenclature, ontology, structure and function of proteins related to control of the DNA integrity mechanisms such as chromatin remodeling, DNA repair and damage response pathways from eight model organisms commonly used in the DNA-related study: Homo sapiens, Mus musculus, Drosophila melanogaster, Caenorhabditis elegans, Saccharomyces cerevisiae, Schizosaccharomyces pombe, Escherichia coli and Arabidopsis thaliana. DNAtraffic contains comprehensive information on diseases related to the assembled human proteins. Database is richly annotated in the systemic information on the nomenclature, chemistry and structure of the DNA damage and drugs targeting nucleic acids and/or proteins involved in the maintenance of genome stability. One of the DNAtraffic database aim is to create the first platform of the combinatorial complexity of DNA metabolism pathway analysis. Database includes illustrations of pathway, damage, protein and drug. Since DNAtraffic is designed to cover a broad spectrum of scientific disciplines it has to be extensively linked to numerous external data sources. Database represents the result of the manual annotation work aimed at making the DNAtraffic database much more useful for a wide range of systems biology applications. DNAtraffic database is freely available and can be queried by the name of DNA network process, DNA damage, protein, disease, and drug.
Proper citation: DNAtraffic (RRID:SCR_008886) Copy
http://genetrail.bioinf.uni-sb.de/
A web-based application that analyzes gene sets for statistically significant accumulations of genes that belong to some functional category. Considered category types are: KEGG Pathways, TRANSPATH Pathways, TRANSFAC Transcription Factor, GeneOntology Categories, Genomic Localization, Protein-Protein Interactions, Coiled-coil domains, Granzyme-B clevage sites, and ELR/RGD motifs. The web server provides two statistical approaches, "Over-Representation Analysis" (ORA) comparing a reference set of genes to a test set, and "Gene Set Enrichment Analysis" (GSEA) scoring sorted lists of genes., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: GeneTrail (RRID:SCR_006250) Copy
http://clipserve.clip.ubc.ca/topfind
An integrated knowledgebase focused on protein termini, their formation by proteases and functional implications. It contains information about the processing and the processing state of proteins and functional implications thereof derived from research literature, contributions by the scientific community and biological databases. It lists more than 120,000 N- and C-termini and almost 10,000 cleavages. TopFIND is a resource for comprehensive coverage of protein N- and C-termini discovered by all available in silico, in vitro as well as in vivo methodologies. It makes use of existing knowledge by seamless integration of data from UniProt and MEROPS and provides access to new data from community submission and manual literature curating. It renders modifications of protein termini, such as acetylation and citrulination, easily accessible and searchable and provides the means to identify and analyse extend and distribution of terminal modifications across a protein. The data is presented to the user with a strong emphasis on the relation to curated background information and underlying evidence that led to the observation of a terminus, its modification or proteolytic cleavage. In brief the protein information, its domain structure, protein termini, terminus modifications and proteolytic processing of and by other proteins is listed. All information is accompanied by metadata like its original source, method of identification, confidence measurement or related publication. A positional cross correlation evaluation matches termini and cleavage sites with protein features (such as amino acid variants) and domains to highlight potential effects and dependencies in a unique way. Also, a network view of all proteins showing their functional dependency as protease, substrate or protease inhibitor tied in with protein interactions is provided for the easy evaluation of network wide effects. A powerful yet user friendly filtering mechanism allows the presented data to be filtered based on parameters like methodology used, in vivo relevance, confidence or data source (e.g. limited to a single laboratory or publication). This provides means to assess physiological relevant data and to deduce functional information and hypotheses relevant to the bench scientist. TopFIND PROVIDES: * Integration of protein termini with proteolytic processing and protein features * Displays proteases and substrates within their protease web including detailed evidence information * Fully supports the Human Proteome Project through search by chromosome location CONTRIBUTE * Submit your N- or C-termini datasets * Contribute information on protein cleavages * Provide detailed experimental description, sample information and raw data
Proper citation: TopFIND (RRID:SCR_008918) Copy
http://plantgrn.noble.org/LegumeIP/
LegumeIP is an integrative database and bioinformatics platform for comparative genomics and transcriptomics to facilitate the study of gene function and genome evolution in legumes, and ultimately to generate molecular based breeding tools to improve quality of crop legumes. LegumeIP currently hosts large-scale genomics and transcriptomics data, including: * Genomic sequences of three model legumes, i.e. Medicago truncatula, Glycine max (soybean) and Lotus japonicus, including two reference plant species, Arabidopsis thaliana and Poplar trichocarpa, with the annotation based on UniProt TrEMBL, InterProScan, Gene Ontology and KEGG databases. LegumeIP covers a total 222,217 protein-coding gene sequences. * Large-scale gene expression data compiled from 104 array hybridizations from L. japonicas, 156 array hybridizations from M. truncatula gene atlas database, and 14 RNA-Seq-based gene expression profiles from G. max on different tissues including four common tissues: Nodule, Flower, Root and Leaf. * Systematic synteny analysis among M. truncatula, G. max, L. japonicus and A. thaliana. * Reconstruction of gene family and gene family-wide phylogenetic analysis across the five hosted species. LegumeIP features comprehensive search and visualization tools to enable the flexible query on gene annotation, gene family, synteny, relative abundance of gene expression.
Proper citation: LegumeIP (RRID:SCR_008906) Copy
A database designed for plant comparative and functional genomics based on complete genomes. It comprises complete proteome sequences from the major phylum of plant evolution. The clustering of these proteomes was performed to define a consistent and extensive set of homeomorphic plant families. Based on this, lists of gene families such as plant or species specific families and several tools are provided to facilitate comparative genomics within plant genomes. The analyses follow two main steps: gene family clustering and phylogenomic analysis of the generated families. Once a group of sequences (cluster) is validated, phylogenetic analyses are performed to predict homolog relationships such as orthologs and ultraparalogs.
Proper citation: GreenPhylDB (RRID:SCR_002834) Copy
Collection of pathways and pathway annotations. The core unit of the Reactome data model is the reaction. Entities (nucleic acids, proteins, complexes and small molecules) participating in reactions form a network of biological interactions and are grouped into pathways (signaling, innate and acquired immune function, transcriptional regulation, translation, apoptosis and classical intermediary metabolism) . Provides website to navigate pathway knowledge and a suite of data analysis tools to support the pathway-based analysis of complex experimental and computational data sets.
Proper citation: Reactome (RRID:SCR_003485) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.