Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Database of information about restriction enzymes and related proteins containing published and unpublished references, recognition and cleavage sites, isoschizomers, commercial availability, methylation sensitivity, crystal, genome, and sequence data. DNA methyltransferases, homing endonucleases, nicking enzymes, specificity subunits and control proteins are also included. Several tools are available including REBsites, BLAST against REBASE, NEBcutter and REBpredictor. Putative DNA methyltransferases and restriction enzymes, as predicted from analysis of genomic sequences, are also listed. REBASE is updated daily and is constantly expanding. Users may submit new enzyme and/or sequence information, recommend references, or send them corrections to existing data. The contents of REBASE may be browsed from the web and selected compilations can be downloaded by ftp (ftp.neb.com). Additionally, monthly updates can be requested via email.,
Proper citation: REBASE (RRID:SCR_007886) Copy
A portal to the Mouse Atlas of Gene Expression Project and Dissecting Gene Expression Networks in Mammalian Organogenesis Project. This Atlas will define the normal state for many tissues by determining, in a comprehensive and quantitative fashion, the number and identity of genes expressed throughout development. The resource will be comprehensive, quantitative, and publicly accessible, containing data on essentially all genes expressed throughout select stages of mouse development. Serial Analysis of Gene Expression (SAGE) is the gene expression methodology of choice for this work. Unlike expressed sequence tags (ESTs) and gene chip data, SAGE data are independent of prior gene discovery and are quantitative. Furthermore, SAGE data are digital, easily exchanged between laboratories for comparison and can be added to by scientists for years to come. Thus, this Atlas will include a data structure and data curation strategy that will facilitate the ongoing collection of gene expression data, even after the completion of this project. The Mouse Atlas project compromises 202 SAGE Libraries from 198 tissues. The list of libraries is available in a number of different groupings, including groups of libraries taken from specific tissue locations and libraries taken from specific developmental stages. Furthermore, this atlas will assemble gene expression profiles for a few focused experiments that will test hypotheses related to the techniques employed, tumor models and models of abnormal development. This will test the resource and provide quality control, validation and demonstrate applicability. Additionally, The Mammalian Organogenesis - Regulation by Gene Expression Networks (MORGEN) project will provide a complete, permanent, and accurate picture of mouse gene expression in the heart (atrioventricular canal and outflow tract), pancreas, and liver; new techniques to understand the interplay of proteins governing the expression of genes key to the development of these organ systems; and the identification of the master regulatory switches that control development of the tissues.
Proper citation: Mouse Gene Expression at the BC Cancer Agency (RRID:SCR_008091) Copy
MitoRes, is a comprehensive and reliable resource for massive extraction of sequences and sub-sequences of nuclear genes and encoded products targeting mitochondria in metazoa. It has been developed for supporting high-throughput in-silico analyses aimed to studies of functional genomics related to mitochondrial biogenesis, metabolism and to their pathological dysfunctions. It integrates information from the most accredited world-wide databases to bring together gene, transcript and encoded protein sequences associated to annotations on species name and taxonomic classification, gene name, functional product, organelle localization, protein tissue specificity, Enzyme Classification (EC), Gene Ontology (GO) classification and links to other related public databases. The section Cluster, has been dedicated to the collection of data on protein clustering of the entire catalogue of MitoRes protein sequences based on all versus all global pair-wise alignments for assessing putative intra- and inter-species functional relationships. The current version of MitoRes is based on the UniProt release 4 and contains 64 different metazoan species. The incredible explosion of knowledge production in Biology in the past two decades has created a critical need for bioinformatic instruments able to manage data and facilitate their retrieval and analysis. Hundreds of biological databases have been produced and the integration of biological data from these different resources is very important when we want to focus our efforts towards the study of a particular layer of biological knowledge. MitoRes is a completely rebuilt edition of MitoNuc database, which has been extensively modified to deal successfully with the challenges of the post genomic era. Its goal is to represent a comprehensive and reliable resource supporting high-quality in-silico analyses aimed to the functional characterization of gene, transcript and amino acid sequences, encoded by the nuclear genome and involved in mitochondrial biogenesis, metabolism and pathological dysfunctions in metazoa. The central features of MitoRes are: # an integrated catalogue of protein, transcript and gene sequences and sub-sequences # a Web-based application composed of a wide spectrum of search/retrieval facilities # a sequence export manager allowing massive extraction of bio-sequences (genes, introns, exons, gene flanking regions, transcripts, UTRs, CDS, proteins and signal peptides) in FASTA, EMBL and GenBank formats. It is an interconnected knowledge management system based on a MySQL relational database, which ensures data consistency and integrity, and on a Web Graphical User Interface (GUI), built in Seagull PHP Framework, offering a wide range of search and sequence extraction facilities. The database is compiled extracting and integrating information from public resources and data generated by the MitoRes team. The MitoRes database consists of comprehensive sequence entries whose core data are protein, transcript and gene sequences and taxonomic information describing the biological source of the protein. Additional information include: bio-sequences structure and location, biological function of protein product and dynamic links to both, external public databases used as data resources and public databases reporting complementary information. The core entity of the MitoRes database is represented by the protein so that each MitoRes entry is generated for each protein reported in the UniProt database as a nuclear encoded protein involved in mitochondrial biogenesis and function. Sponsors: MitoRes has been supported by Ministero Universit e Ricerca Scientifica, Italy (PRIN, Programma Biotecnologie legge 95/95-MURST 5, Proiect MURST Cluster C03/2000, CEGBA). Currently it is supported by operating grants from the Ministero dellIstruzione, dellUniversit e della Ricerca (MIUR), Italy (PNR 2001-2003 (FIRB art.8) D.M. 199, Strategic Program: Post-genome, grant 31-063933 and Project n.2, Cluster C03 L. 488/929).
Proper citation: MitoRes (RRID:SCR_008208) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 29, 2016. An algorithm that finds articles most relevant to a genetic sequence. In the genomic era, researchers often want to know more information about a biological sequence by retrieving its related articles. However, there is no available tool yet to achieve conveniently this goal. Here, a new literature-mining tool MedBlast is developed, which uses natural language processing techniques, to retrieve the related articles of a given sequence. An online server of this program is also provided. The genome sequencing projects generate such a large amount of data every day that many molecular biologists often encounter some sequences that they know nothing about. Literature is usually the principal resource of such information. It is relatively easy to mine the articles cited by the sequence annotation; however, it is a difficult task to retrieve those relevant articles without direct citation relationship. The related articles are those described in the given sequence (gene/protein), or its redundant sequences, or the close homologs in various species. They can be divided into two classes: direct references, which include those either cited by the sequence annotation or citing the sequence in its text; indirect references, those which contain gene symbols of the given sequence. A few additional issues make the task even more complicated: (1) symbols may have aliases; and (2) one sequence may have a couple of relatives that we want to take into account too, which include redundant (e.g. protein and gene sequences) and close homologs. Here the issues are addressed by the development of the software MedBlast, which can retrieve the related articles of the given sequence automatically. MedBlast uses BLAST to extend homology relationships, precompiled species-specific thesauruses, a useful semantics technique in natural language processing (NLP), to extend alias relationship, and EUtilities toolset to search and retrieve corresponding articles of each sequence from PubMed. MedBlast take a sequence in FASTA format as input. The program first uses BLAST to search the GenBank nucleic acid and protein non-redundant (nr) databases, to extend to those homologous and corresponding nucleic acid and protein sequences. Users can input the BLAST results directly, but it is recommended to input the result of both protein and nucleic acid nr databases. The hits with low e-values are chosen as the relatives because the low similarity hits often do not contain specific information. Very long sequences, e.g. 100k, which are usually genomic sequences, are discarded too, for they do not contain specific direct references. User can adjust these parameters to meet their own needs.
Proper citation: MedBlast (RRID:SCR_008202) Copy
http://www.koki.hu/main.php?folderID=903
The aim of this laboratory is to understand how information is encoded in specific spatiotemporal activity patterns and structural configurations at the circuit, cellular, and molecular levels in the hippocampus, thereby enabling the process of memory. A major task is to find the neuronal codes of internal representations of memory items and the mapping rules between the levels of gene expression/proteins synthesis and the level of cognitive processing. Novel combinations of approaches, including multiple single-cell recording technology, patch-clamp electrophysiology, neuroanatomy/neurochemistry at the cellular and subcellular levels, and computational models are employed to test specific hypotheses about that mapping process (such as local circuit anatomy and activity-dependent short-term and long-term synaptic plasticity). Collaborations within the Institute allows the group to also incorporate gene targeting methods and behavioral learning/memory tests in their methodological repertoire. The laboratory has been focusing on the normal and pathological (epileptic, ischemic) activity of cortical networks, with particular attention to the generation of behaviour-dependent population discharge patterns (theta and gamma oscillations, hippocampal sharp waves). Anatomical, in vitro and in vivo electrophysiological, pharmacological and molecular techniques and modeling are combined to elucidate the functional roles of inhibitory cell types in the control of population synchrony and synaptic plasticity in the hippocampus, their local and subcortical modulation via selective afferent pathways (GABAergic and cholinergic septal, as well as serotonergic raphe input) and pre- or postsynaptic receptors. An expanding new direction of research is related to the role of endocannabinoid signaling in the activity-dependent modulation of GABAergic and glutamatergic transmission, and its involvement in anxiety-like behavior.
Proper citation: Institute of Experimental Medicine of the Hungarian Academy of Sciences: Laboratory of Cerebral Cortext Reserach (RRID:SCR_008041) Copy
https://wiki.med.harvard.edu/SysBio/Megason/GoFigure
GoFigure is a software platform for quantitating complex 4d in vivo microscopy based data in high-throughput at the level of the cell. A prime goal of GoFigure is the automatic segmentation of nuclei and cell membranes and in temporally tracking them across cell migration and division to create cell lineages. GoFigure v2.0 is a major new release of our software package for quantitative analysis of image data. The research focuses on analyzing cells in intact, whole zebrafish embryos using 4d (xyzt) imaging which tends to make automatic segmentation more difficult than with 2d or 2d+time imaging of cells in culture. This resource has developed an automatic segmentation pipeline that includes ICA based channel unmixing, membrane nuclear channel subtraction, Gaussian correlation, shape models, and level set based variational active contours. GoFigure was designed to meet the challenging requirements of in toto imaging. In toto imaging is a technology that we are developing in which we seek to track all the cell movements and divisions that form structures during embryonic development of zebrafish and to quantitate protein expression and localization on top of this digital lineage. For in toto imaging, GoFigure uses zebrafish embryos in which the nuclei and cell membranes have been marked with 2 different color fluorescent proteins to allow cells to be segmented and tracked. A transgenic line in a third color can be used to mark protein expression and localization using a genetic approach that this resource developed called FlipTraps or using traditional transgenic approaches. Embryos are imaged using confocal or 2-photon microscopy to capture high-resolution xyzt image sets used for cell tracking. The GoFigure GUI will provide many tools for visualization and analysis of bioimages. Since fully automatic segmentation of cells is never perfect, GoFigure will provide easy to use tools for semi-automatically and manually adding, deleting, and editing traces in 2d (figures-xy, xz, or yz), 3d (meshes- xyz), 4d (tracks- xyzt) and 4d+cell division (lineages). GoFigure will also provide a number of views into complex image data sets including 3d XYZ and XYT image views, tabular list views of traces, histograms, and scattergrams. Importantly, all these views will be linked together to allow the user to explore their data from multiple angles. Data will be easily sorted and color-coded in many ways to explore correlations in higher dimensional data. The GoFigure architecture is designed to allow additional segmentation, visualization, and analysis filters to be plugged in. Sponsors: GoFigure is developed by Harvard University., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Harvard Medical School, Department of Systems Biology: The Megason Lab -GoFigure Software (RRID:SCR_008037) Copy
http://www.ebi.ac.uk/parasites/parasite-genome.html
This website contains information about the genomic sequence of parasites. It also contains multiple search engines to search six frame translations of parasite nucleotide databases for motifs, parasite protein databases for motifs, and parasite protein databases for keywords and text terms. * Guide to Internet Access to Parasite Genome Information * Guide to web-based analysis tools * Parasite Genome BLAST Server: Search a range of parasite specific nucleotide sequence databases with your own sequence. * Parasite Proteome Keyword Search Facility: Search parasite protein databases for keywords and text terms * Parasite Proteome Motif Search Facility: Search parasite protein databases for motifs * Parasite Six Frame Translation Motif Search Facility: Search six frame translations of parasite nucleotide databases for motifs * Genome computing resources: A list of ftp and gopher sites where genome computing applications and other resources can be found.
Proper citation: Parasite genome databases and genome research resources (RRID:SCR_008150) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 23, 2016. PDBfun is a web server for structural and functional analysis of proteins at the residue level. pdbFun gives fast access to the whole Protein Data Bank (PDB) organized as a database of annotated residues. The available data (features) range from solvent exposure to ligand binding ability, location in a protein cavity, secondary structure, residue type, sequence functional pattern, protein domain and catalytic activity. PDBfun is an integrated web tool for querying the PDB at the residue level and for local structural comparison. It integrates knowledge on single residues in protein structures coming from other databases or calculated with available or in-house developed instruments for structural analysis. Each set of different annotations represents a feature. Features are listed in PDBfun main page in orange. Features can be used for building residues selections.
Proper citation: Protein Databank Fun (RRID:SCR_008226) Copy
http://www.poissonboltzmann.org/apbs/
APBS is a software package for modeling biomolecular solvation through solution of the Poisson-Boltzmann equation (PBE), one of the most popular continuum models for describing electrostatic interactions between molecular solutes in salty, aqueous media. APBS was designed to efficiently evaluate electrostatic properties for such simulations for a wide range of length scales to enable the investigation of molecules with tens to millions of atoms. It also provides implicit solvent models of nonpolar solvation which accurately account for both repulsive and attractive solute-solvent interactions. APBS uses FEtk (the Finite Element ToolKit) to solve the Poisson-Boltzmann equation numerically. FEtk is a portable collection of finite element modeling class libraries written in an object-oriented version of C. It is designed to solve general coupled systems of nonlinear partial differential equations using adaptive finite element methods, inexact Newton methods, and algebraic multilevel methods.
Proper citation: Adaptive Poisson-Boltzmann Solver (RRID:SCR_008387) Copy
http://ophid.utoronto.ca/navigator/
A software package for visualizing and analyzing protein-protein interaction networks. NAViGaTOR can query OPHID / I2D - online databases of interaction data - and display networks in 2D or 3D. To improve scalability and performance, NAViGaTOR combines Java with OpenGL to provide a 2D/3D visualization system on multiple hardware platforms. NAViGaTOR also provides analytical capabilities and supports standard import and export formats such as GO and the Proteomics Standards Initiative (PSI). NAViGaTOR can be installed and run on Microsoft Windows, Linux / UNIX, and Mac OS systems. NAViGaTOR is written in Java and uses JOGL (Java bindings for OpenGL) to support scalability, highlighting or suppressing of information, and other advanced graphic approaches.
Proper citation: Network Analysis, Visualization and Graphing TORonto (RRID:SCR_008373) Copy
The JCSG is a multi-institutional consortium that aims to explore the expanding protein universe to find new challenges and opportunities to significantly contribute to new biology, chemistry and medicine through development of HT approaches to structural genomics. The mission of JCSG is to to operate a robust HT protein structure determination pipeline as a large-scale production center for PSI-2. A major goal is to ensure that innovative high-throughput approaches are developed that advance not only structural genomics, but also structural biology in general, via investigation of large numbers of high-value structures that populate protein fold and family space and by increasing the efficiency of structure determination at substantially reduced cost. The JCSG centralizes each core activity into single dedicated sites, each handling distinct, but interconnected objectives. This unique approach allows each specialized group to focus on its own area of expertise and provides well-defined interfaces among the groups. In addition, this approach addresses the requirements for the scalability needed to process large numbers of targets at a greatly reduced cost per target. JCSG production groups are: - Administrative Core - Bioinformatics Core - Crystallomics Core - Structure Determination Core - NMR Core JCSG is deeply committed to the development of new technologies that facilitate high throughput structural genomics. The areas of development include hardware, software, new experimental methods, and adaptation of existing technologies to advance genome research. In the hardware arena, their commitment is to the development of technologies that accelerate structure solution by increasing throughput rates at every stage of the production pipeline. Therefore, one major area of hardware development has been the implementation of robotics. In the software arena, they have developed enterprise resource software that track success, failures, and sample histories from target selection to PDB deposition, annotation and target management tools, and helper applications aimed at facilitating and automating multiple steps in the pipeline. Sponsors: The Joint Center for Structural Genomics is funded by the National Institute of General Medical Sciences (NIGMS), as part of the second phase of the Protein Structure Initiative (PSI) of the National Institutes of Health (U54 GM074898).
Proper citation: Joint Center for Structural Genomics (RRID:SCR_008251) Copy
http://degradome.uniovi.es/diseases.html
This resource has cataloged a total of 80 human hereditary diseases caused by mutations in protease-coding genes, which implies that more than 10% of the human protease genes are involved in human pathologies. They are classified in three groups: loss of function, gain of function, and an heterogeneous group including non-protease homologs (np), putative proteases, and hedgehog proteins with only autoprocessing activity. Type of inheritance is indicated by R (recessive) or D (dominant).
Proper citation: Human Hereditary Diseases of Proteolysis (RRID:SCR_008344) Copy
https://services.healthtech.dtu.dk/services/NetNGlyc-1.0/
Server that predicts N-Glycosylation sites in human proteins using artificial neural networks that examine the sequence context of Asn-Xaa-Ser/Thr sequons. NetNGlyc 1.0 is also available as a stand-alone software package, with the same functionality as the service above. Ready-to-ship packages exist for the most common UNIX platforms.
Proper citation: NetNGlyc (RRID:SCR_001570) Copy
https://services.healthtech.dtu.dk/services/YinOYang-1.2/
Server that produces neural network predictions for O-beta-GlcNAc attachment sites in eukaryotic protein sequences. This server can also use NetPhos, to mark possible phosphorylated sites and hence identify Yin-Yang sites. YinOYang 1.2 is available as a stand-alone software package, with the same functionality. Ready-to-ship packages exist for the most common UNIX platforms.
Proper citation: YinOYang (RRID:SCR_001605) Copy
https://www.ebi.ac.uk/jdispatcher/msa/clustalo?stype=protein
Software package as multiple sequence alignment tool that uses seeded guide trees and HMM profile-profile techniques to generate alignments between three or more sequences. Accepts nucleic acid or protein sequences in multiple sequence formats NBRF/PIR, EMBL/UniProt, Pearson (FASTA), GDE, ALN/Clustal, GCG/MSF, RSF.
Proper citation: Clustal Omega (RRID:SCR_001591) Copy
http://matrixdb.univ-lyon1.fr/
Freely available database focused on interactions established by extracellular proteins and polysaccharides, taking into account the multimeric nature of the extracellular proteins (e.g. collagens, laminins and thrombospondins are multimers). MatrixDB is an active member of the International Molecular Exchange (IMEx) consortium and has adopted the PSI-MI standards for annotating and exchanging interaction data. It includes interaction data extracted from the literature by manual curation, and offers access to relevant data involving extracellular proteins provided by the IMEx partner databases through the PSICQUIC webservice, as well as data from the Human Protein Reference Database. The database reports mammalian protein-protein and protein-carbohydrate interactions involving extracellular molecules. Interactions with lipids and cations are also reported. MatrixDB is focused on mammalian interactions, but aims to integrate interaction datasets of model organisms when available. MatrixDB provides direct links to databases recapitulating mutations in genes encoding extracellular proteins, to UniGene and to the Human Protein Atlas that shows expression and localization of proteins in a large variety of normal human tissues and cells. MatrixDB allows researchers to perform customized queries and to build tissue- and disease-specific interaction networks that can be visualized and analyzed with Cytoscape or Medusa. Statistics (2013): 2283 extracellular matrix interactions including 2095 protein-protein and 169 protein-glycosaminoglycan interactions.
Proper citation: MatrixDB (RRID:SCR_001727) Copy
http://datahub.io/dataset/kupkb
A collection of omics datasets (mRNA, proteins and miRNA) that have been extracted from PubMed and other related renal databases, all related to kidney physiology and pathology giving KUP biologists the means to ask queries across many resources in order to aggregate knowledge that is necessary for answering biological questions. Some microarray raw datasets have also been downloaded from the Gene Expression Omnibus and analyzed by the open-source software GeneArmada. The Semantic Web technologies, together with the background knowledge from the domain's ontologies, allows both rapid conversion and integration of this knowledge base. SPARQL endpoint http://sparql.kupkb.org/sparql The KUPKB Network Explorer will help you visualize the relationships among molecules stored in the KUPKB. A simple spreadsheet template is available for users to submit data to the KUPKB. It aims to capture a minimal amount of information about the experiment and the observations made.
Proper citation: Kidney and Urinary Pathway Knowledge Base (RRID:SCR_001746) Copy
A computer algorithm to predict aggregation nucleating regions in proteins as well the effect of mutations and environmental conditions on the aggregation propensity of these regions.
Proper citation: TANGO (RRID:SCR_001770) Copy
Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
Proper citation: BLASTX (RRID:SCR_001653) Copy
The Physiome Project is a worldwide public domain effort to provide a computational framework for understanding human and other eukaryotic physiology. It aims to develop integrative models at all levels of biological organization, from genes to the whole organism via gene regulatory networks, protein pathways, integrative cell function, and tissue and whole organ structure/function relations. Additionally, an important goal of the project is to develop applications for teaching physiology. Current projects include the development of: - ontologies to organize biological knowledge and access to databases - markup languages to encode models of biological structure and function in a standard format for sharing between different application programs and for re-use as components of more comprehensive models - databases of structure at the cell, tissue and organ levels - software to render computational models of cell function such as ion channel electrophysiology, cell signaling and metabolic pathways, transport, motility, the cell cycle, etc. in 2 & 3D graphical form - software for displaying and interacting with the organ models which will allow the user to move across all spatial scales Sponsors: This project is supported by the International Union of Physiological Sciences (IUPS), the IEEE Engineering. in Medicine and Biology (EMBS), and the International Federation for Medical and Biological Engineering (IFMBE)
Proper citation: International Union of Physiological Sciences: Physiome Project (RRID:SCR_001760) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.