Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Integrated database resource consisting of 16 main databases, broadly categorized into systems information, genomic information, and chemical information. In particular, gene catalogs in completely sequenced genomes are linked to higher-level systemic functions of cell, organism, and ecosystem. Analysis tools are also available. KEGG may be used as reference knowledge base for biological interpretation of large-scale datasets generated by sequencing and other high-throughput experimental technologies.
Proper citation: KEGG (RRID:SCR_012773) Copy
http://www.kegg.jp/kegg/download/kegtools.html
Freely available desktop applications including KegHier: Java application for browsing BRITE hierarchy files, KegArray: Java application for microarray data analysis, and KegDraw: Java application for drawing compound and glycan structures. They run on the Mac OS X, Windows, and Linux platforms.
Proper citation: KegTools (RRID:SCR_006432) Copy
http://www.bioextract.org/GuestLogin
An open, web-based system designed to aid researchers in the analysis of genomic data by providing a platform for the creation of bioinformatic workflows. Scientific workflows are created within the system by recording tasks performed by the user. These tasks may include querying multiple, distributed data sources, saving query results as searchable data extracts, and executing local and web-accessible analytic tools. The series of recorded tasks can then be saved as a reproducible, sharable workflow available for subsequent execution with the original or modified inputs and parameter settings. Integrated data resources include interfaces to the National Center for Biotechnology Information (NCBI) nucleotide and protein databases, the European Molecular Biology Laboratory (EMBL-Bank) non-redundant nucleotide database, the Universal Protein Resource (UniProt), and the UniProt Reference Clusters (UniRef) database. The system offers access to numerous preinstalled, curated analytic tools and also provides researchers with the option of selecting computational tools from a large list of web services including the European Molecular Biology Open Software Suite (EMBOSS), BioMoby, and the Kyoto Encyclopedia of Genes and Genomes (KEGG). The system further allows users to integrate local command line tools residing on their own computers through a client-side Java applet.
Proper citation: BioExtract (RRID:SCR_005397) Copy
http://www.ici.upmc.fr/cluego/
A Cytoscape plug-in that visualizes the non-redundant biological terms for large clusters of genes in a functionally grouped network. It can be used in combination with GOlorize. The identifiers can be uploaded from a text file or interactively from a network of Cytoscape. The type of identifiers supported can be easily extended by the user. ClueGO performs single cluster analysis and comparison of clusters. From the ontology sources used, the terms are selected by different filter criteria. The related terms which share similar associated genes can be combined to reduce redundancy. The ClueGO network is created with kappa statistics and reflects the relationships between the terms based on the similarity of their associated genes. On the network, the node colour can be switched between functional groups and clusters distribution. ClueGO charts are underlying the specificity and the common aspects of the biological role. The significance of the terms and groups is automatically calculated. ClueGO is easy updatable with the newest files from Gene Ontology and KEGG. Platform: Windows compatible, Mac OS X compatible, Linux compatible, Unix compatible, THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: ClueGO (RRID:SCR_005748) Copy
National university in Kyoto, Japan. It is the second oldest Japanese university, one of Asia highest ranked universities and one of Japan National Seven Universities.
Proper citation: Kyoto University; Kyoto; Japan (RRID:SCR_004108) Copy
http://kt.ijs.si/software/SEGS/
A web tool for descriptive analysis of microarray data. The analysis is performed by looking for descriptions of gene sets that are statistically significantly over- or under-expressed between different scenarios within the context of a genome-scale experiments (DNA microarray). Descriptions are defined by using the terms from the Gene Ontology (GO), the Kyoto Encyclopedia of Genes and Genomes (KEGG) pathways and gene-gene interactions found in the ENTREZ database. Gene annotations by GO and KEGG terms can also be found in the ENTREZ database. The tool provides three procedures for testing the enrichment of the gene sets (over- or under-expressed): Fisher's exact test, GSEA and PAGE, and option for combining the results of the tests. Because of the multiple-hypothesis testing nature of the problem, all the p-values are computed using the permutation testing method.
Proper citation: SEGS (RRID:SCR_003554) Copy
A web-based tool to support meta-analysis of multiple gene-expression data sets, as well as to enable integration of data sets from gene expression and metabolomics experiments. INMEX contains three functional modules. The data preparation module supports flexible data processing, annotation and visualization of individual data sets. The statistical analysis module allows researchers to combine multiple data sets based on P-values, effect sizes, rank orders and other features. The significant genes can be examined in functional analysis module for enriched Gene Ontology terms or Kyoto Encyclopedia of Genes and Genomes (KEGG) pathways, or expression profile visualization. INMEX has built-in support for common gene/metabolite identifiers (IDs), as well as 45 popular microarray platforms for human, mouse and rat. Complex operations are performed through a user-friendly web interface in a step-by-step manner.
Proper citation: INMEX (RRID:SCR_004173) Copy
http://www.ncbi.nlm.nih.gov/biosystems/
Database that provides access to biological systems and their component genes, proteins, and small molecules, as well as literature describing those biosystems and other related data throughout Entrez. A biosystem, or biological system, is a group of molecules that interact directly or indirectly, where the grouping is relevant to the characterization of living matter. BioSystem records list and categorize components, such as the genes, proteins, and small molecules involved in a biological system. The companion FLink tool, in turn, allows you to input a list of proteins, genes, or small molecules and retrieve a ranked list of biosystems. A number of databases provide diagrams showing the components and products of biological pathways along with corresponding annotations and links to literature. This database was developed as a complementary project to (1) serve as a centralized repository of data; (2) connect the biosystem records with associated literature, molecular, and chemical data throughout the Entrez system; and (3) facilitate computation on biosystems data. The NCBI BioSystems Database currently contains records from several source databases: KEGG, BioCyc (including its Tier 1 EcoCyc and MetaCyc databases, and its Tier 2 databases), Reactome, the National Cancer Institute's Pathway Interaction Database, WikiPathways, and Gene Ontology (GO). It includes several types of records such as pathways, structural complexes, and functional sets, and is desiged to accomodate other record types, such as diseases, as data become available. Through these collaborations, the BioSystems database facilitates access to, and provides the ability to compute on, a wide range of biosystems data. If you are interested in depositing data into the BioSystems database, please contact them.
Proper citation: NCBI BioSystems Database (RRID:SCR_004690) Copy
http://equilibrator.weizmann.ac.il/
Web interface designed for thermodynamic analysis of biochemical systems. eQuilibrator enables free-text search for biochemical compounds and reactions and provides thermodynamic estimates for both in a variety of conditions. It can provide estimates for compounds in the KEGG database, and individual compounds and enzymes can be searched for by their common names (water, glucosamine, hexokinase). Reactions can be entered in a free-text format that eQuilibrator parses automatically. eQuilibrator also allows manipulation of the conditions of a reaction - pH, ionic strength, and reactant and product concentrations.
Proper citation: eQuilibrator (RRID:SCR_006011) Copy
Web application that filters and links enriched output data identifying sets of associated genes and terms, producing metagroups of coherent biological significance. The method uses fuzzy reciprocal linkage between genes and terms to unravel their functional convergence and associations. It can also be accessed through its web service.
Proper citation: GeneTerm Linker (RRID:SCR_006385) Copy
http://genetrail.bioinf.uni-sb.de/
A web-based application that analyzes gene sets for statistically significant accumulations of genes that belong to some functional category. Considered category types are: KEGG Pathways, TRANSPATH Pathways, TRANSFAC Transcription Factor, GeneOntology Categories, Genomic Localization, Protein-Protein Interactions, Coiled-coil domains, Granzyme-B clevage sites, and ELR/RGD motifs. The web server provides two statistical approaches, "Over-Representation Analysis" (ORA) comparing a reference set of genes to a test set, and "Gene Set Enrichment Analysis" (GSEA) scoring sorted lists of genes., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: GeneTrail (RRID:SCR_006250) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 14,2026. Integrated database of genomic, expression and protein data for Drosophila, Anopheles, C. elegans and other organisms. You can run flexible queries, export results and analyze lists of data. FlyMine presents data in categories, with each providing information on a particular type of data (for example Gene Expression or Protein Interactions). Template queries, as well as the QueryBuilder itself, allow you to perform searches that span data from more than one category. Advanced users can use a flexible query interface to construct their own data mining queries across the multiple integrated data sources, to modify existing template queries or to create your own template queries. Access our FlyMine data via our Application Programming Interface (API). We provide client libraries in the following languages: Perl, Python, Ruby and & Java API
Proper citation: FlyMine (RRID:SCR_002694) Copy
Web server to identify statistically enriched pathways, diseases, and GO terms for a set of genes or proteins, using pathway, disease, and GO knowledge from multiple famous databases. It allows for both ID mapping and cross-species sequence similarity mapping. It then performs statistical tests to identify statistically significantly enriched pathways and diseases. KOBAS 2.0 incorporates knowledge across 1327 species from 5 pathway databases (KEGG PATHWAY, PID, BioCyc, Reactome and Panther) and 5 human disease databases (OMIM, KEGG DISEASE, FunDO, GAD and NHGRI GWAS Catalog). A standalone command line version is also available, THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: KOBAS (RRID:SCR_006350) Copy
http://www.genome.jp/kegg/expression/
Database for mapping gene expression profiles to pathways and genomes. Repository of microarray gene expression profile data for Synechocystis PCC6803 (syn), Bacillus subtilis (bsu), Escherichia coli W3110 (ecj), Anabaena PCC7120 (ana), and other species contributed by the Japanese research community.
Proper citation: Kyoto Encyclopedia of Genes and Genomes Expression Database (RRID:SCR_001120) Copy
http://www.ebi.ac.uk/thornton-srv/databases/FunTree/
FunTree provides a range of data resources to detect the evolution of enzyme function within distant structurally related clusters within domain super families as determined by CATH. To access the resource enter a specific CATH superfamily code or search for a structure / sequence / function (either via a EC code or KEGG ligand / reaction ID, PDB ID or UniProtKB ID). Or browse the resource via superfamily / function / structure / metabolites & reactions via the menu on the left panel. FunTree is a new resource that brings together sequence, structure, phylogenetic, chemical and mechanistic information for structurally defined enzyme superfamilies. Gathering together this range of data into a single resource allows the investigation of how novel enzyme functions have evolved within a structurally defined superfamily as well as providing a means to analyse trends across many superfamilies. This is done not only within the context of an enzyme''''s sequence and structure but also the relationships of their reactions. Developed in tandem with the CATH database, it currently comprises 276 superfamilies covering 1800 (70%) of sequence assigned enzyme reactions. Central to the resource are phylogenetic trees generated from structurally informed multiple sequence alignments using both domain structural alignments supplemented with domain sequences and whole sequence alignments based on commonality of multi-domain architectures. These trees are decorated with functional annotations such as metabolite similarity as well as annotations from manually curated resources such the catalytic site atlas and MACiE for enzyme mechanisms.
Proper citation: FunTree (RRID:SCR_006014) Copy
Database about gene regulation and gene expression in prokaryotes. It includes a manually curated and unique collection of transcription factor binding sites. A variety of bioinformatics tools for the prediction, analysis and visualization of regulons and gene reglulatory networks is included. The integrated approach provides information about molecular networks in prokaryotes with focus on pathogenic organisms. In detail this concerns: * transcriptional regulation (transcription factors and their DNA binding sites * signal transduction (two-component systems, phosphylation cascades) * protein interactions (complex formation, oligomerization) * biochemical pathways (chemical reactions) * other regulation events (e.g. codon usage, etc. ...) It aims to be a resource to model protein-host interactions and to be a suitable platform to analyze high-throughput data from proteomis and transcriptomics experiments (systems biology). Currently it mainly contains detailed information about operon and promoter structures including huge collections of transcription factor binding sites. If an appropriate number of regulatory binding sites is available, a position weight matrix (PWM) and a sequence logo is provided, which can be used to predict new binding sites. This data is collected manually by screening the original scientific literature. PRODORIC also handles protein-protein interactions and signal-transduction cascades that commonly occur in form of two-component systems in prokaryotes. Furthermore it contains metabolic network data imported from the KEGG database., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: PRODORIC (RRID:SCR_007074) Copy
A public repository of metabolite information as well as tandem mass spectrometry data is provided to facilitate metabolomics experiments. It contains structures and represents a data management system designed to assist in a broad array of metabolite research and metabolite identification. An annotated list of known metabolites and their mass, chemical formula, and structure are available. Each metabolite is linked to outside resources for further reference and inquiry. MS/MS data is also available on many of the metabolites.
Proper citation: METLIN (RRID:SCR_010500) Copy
http://www.arabidopsisreactome.org
Curated database of core pathways and reactions in plant biology that covers biological pathways ranging from the basic processes of metabolism to high-level processes such as cell cycle regulation. While it is targeted at Arabidopsis pathways, it also includes many biological events from other plant species. This makes the database relevant to the large number of researchers who work on other plants. Arabidopsis Reactome currently contains both in-house curated pathways as well as imported pathways from AraCyc and KEGG databases. All the curated information is backed up by its provenance: either a literature citation or an electronic inference based on sequence similarity. Their ontology ensures that the various events are linked in an appropriate spatial and temporal context.
Proper citation: Arabidopsis Reactome (RRID:SCR_002063) Copy
An integrative interaction database that integrates different types of functional interactions from heterogeneous interaction data resources. Physical protein interactions, metabolic and signaling reactions and gene regulatory interactions are integrated in a seamless functional association network that simultaneously describes multiple functional aspects of genes, proteins, complexes, metabolites, etc. With human, yeast and mouse complex functional interactions, it currently constitutes the most comprehensive publicly available interaction repository for these species. Different ways of utilizing these integrated interaction data, in particular with tools for visualization, analysis and interpretation of high-throughput expression data in the light of functional interactions and biological pathways is offered.
Proper citation: ConsensusPathDB (RRID:SCR_002231) Copy
http://agem.cnb.csic.es/VisualOmics/aGEM/
Database platform of an integrated view of eight databases (mouse gene expression resources: EMAGE, GXD, GENSAT, BioGPS, ABA, EUREXPRESS; human gene expression databases: HUDSEN, BioGPS and Human Protein Atlas) that allows the experimentalist to retrieve relevant statistical information relating gene expression, anatomical structure (space) and developmental stage (time). Moreover, general biological information from databases such as KEGG, OMIM and MTB is integrated too. It can be queried using gene and anatomical structure. Output information is presented in a friendly format, allowing the user to display expression maps and correlation matrices for a gene or structure during development. An in-depth study of a specific developmental stage is also possible using heatmaps that relate gene expression with anatomical components. This is a powerful tool in the gene expression field that makes easy the access to information related to the anatomical pattern of gene expression in human and mouse, so that it can complement many functional genomics studies. The platform allows the integration of gene expression data with spatial-temporal anatomic data by means of an intuitive and user friendly display., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: aGEM (RRID:SCR_013349) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.