Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A database of new exon boundaries induced by pathogenic mutations in human disease genes.
Proper citation: DBASS (RRID:SCR_002107) Copy
A tool for annotating, exploring, and analyzing gene sets that may be associated with cancer.
Proper citation: Mutation Annotation and Genomic Interpretation (RRID:SCR_002800) Copy
http://www.tanpaku.org/autophagy/
Database that provides basic, up-to-date information on relevant literature, and a list of autophagy-related proteins and their homologs in eukaryotes.
Proper citation: Autophagy Database (RRID:SCR_002671) Copy
http://greengenes.secondgenome.com/downloads
Database that provides access to the current and comprehensive 16S rRNA gene sequence alignment for browsing, blasting, probing, and downloading. The data and tools can assist the researcher in choosing phylogenetically specific probes, interpreting microarray results, and aligning/annotating novel sequences. The 16S rRNA gene database provides chimera screening, standard alignment, and taxonomic classification using multiple published taxonomies. ARB users can use Greengenes to update local databases., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Greengenes (RRID:SCR_002830) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 13,2026. Database of known and predicted protein domain (domain-domain) interactions containing interactions inferred from PDB entries, and those that are predicted by 8 different computational approaches using Pfam domain definitions. DOMINE contains a total of 26,219 domain-domain interactions (among 5,410 domains) out of which 6,634 are inferred from PDB entries, and 21,620 are predicted by at least one computational approach. Of the 21,620 computational predictions, 2,989 interactions are high-confidence predictions (HCPs), 2,537 interactions are medium-confidence predictions (MCPs), and the remaining 16,094 are low-confidence predictions (LCPs). (May 2014)
Proper citation: DOMINE: Database of Protein Interactions (RRID:SCR_002399) Copy
Database of experimentally validated gene regulatory relations and the corresponding transcription factor binding sites upstream of Bacillus subtilis genes. The database allows the comparison of systematic experiments with individual experimental results in order to facilitate the elucidation of the complete B. subtilis gene regulatory network. The current version is constructed by surveying 947 references and contains the information of 120 binding factors and 1475 gene regulatory relations. For each promoter, all of its known cis-elements are listed according to their positions, while these cis-elements are aligned to illustrate the consensus sequence for each transcription factor. All probable transcription factors coded in the genome were classified using Pfam motifs. The DBTBS database was reorganized to show operons instead of individual genes as the building blocks of gene regulatory networks. It now contains 463 experimentally known operons, as well as their terminator sequences if identifiable. In addition, 517 transcriptional terminators were identified computationally. (De Hoon, M.J.L. et al., PLoS Comput. Biol. 1, e25 (2005)). A new section was added under "Motif conservation", which presents hexameric motifs found to be conserved to different extents between upstream intergenic regions of genus-specific subgroups of homologous proteins.
Proper citation: DBTBS (RRID:SCR_002345) Copy
A database of orthologous groups of genes. The orthologous groups are annotated with functional description lines (derived by identifying a common denominator for the genes based on their various annotations), with functional categories (i.e derived from the original COG/KOG categories). eggNOG's database currently counts 1.7 million orthologous groups in 3686 species, covering over 7.7 million proteins (built from 9.6 million proteins). (Jan 30, 2014)
Proper citation: eggNOG (RRID:SCR_002456) Copy
http://bioinf-apache.charite.de/supertarget_v2/
Database for analyzing drug-target interactions, it integrates drug-related information associated with medical indications, adverse drug effects, drug metabolism, pathways and Gene Ontology (GO) terms for target proteins. At present (May 2013), the updated database contains >6000 target proteins, which are annotated with >330 000 relations to 196 000 compounds (including approved drugs); the vast majority of interactions include binding affinities and pointers to the respective literature sources. The user interface provides tools for drug screening and target similarity inclusion. A query interface enables the user to pose complex queries, for example, to find drugs that target a certain pathway, interacting drugs that are metabolized by the same cytochrome P450 or drugs that target proteins within a certain affinity range.
Proper citation: SuperTarget (RRID:SCR_002696) Copy
Database of protein-ligand crystal structures that is a subset of the Protein Data Bank (PDB), containing every high-quality example of ligand-protein binding. The resolved protein crystal structures with clearly identified biologically relevant ligands are annotated with experimentally determined binding data extracted from literature. A viewer is provided to examine the protein-ligand structures. Ligands have additional chemical data, allowing for cheminformatics mining. The binding-affinity data ranges 13 orders of magnitude. The issue of redundancy in the data has also been addressed. To create a nonredundant dataset, one protein from each of the 1780 protein families was chosen as a representative. Representatives were chosen by tightest binding, best resolution, etc. For the 1780 best complexes that comprise the nonredundant version of Binding MOAD, 475 (27%) have binding data. This collection of protein-ligand complexes will be useful in elucidating the biophysical patterns of molecular recognition and enzymatic regulation. The complexes with binding-affinity data will help in the development of improved scoring functions and structure-based drug discovery techniques.
Proper citation: Binding MOAD (RRID:SCR_002294) Copy
http://edas2.bioinf.fbb.msu.ru/
Databases of alternatively spliced genes with data on the alignment of proteins, mRNAs, and EST. It contains information on all exons and introns observed, as well as elementary alternatives formed from them. The database makes it possible to filter the output data by changing the cut-off threshold by the significance level. It contains splicing information on human, mouse, dog (not yet functional) and rat (not yet functional). For each database, users can search by keyword or by overall gene expression. They can also view genes based on chromosomal arrangement or other position in genome (exon, intron, acceptor site, donor site), functionality, position, conservation, and EST coverage. Also offered is an online Fisher test.
Proper citation: EDAS - EST-Derived Alternative Splicing Database (RRID:SCR_002449) Copy
http://bioinf.gen.tcd.ie/casbah/
Database which contains information pertaining to all currently known caspase substrates.
Proper citation: CASBAH (RRID:SCR_002728) Copy
http://www.liu.se/hu/mdl/main/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 23,2022. An on-line database and publically accessible depository that is dedicated to the omics of small biomolecules.
Proper citation: NMR metabolomics database of Linkoping (RRID:SCR_002758) Copy
http://www.genscript.com/psort/wolf_psort.html
Data analysis service for protein subcellular localization prediction.
Proper citation: WoLF PSORT (RRID:SCR_002472) Copy
http://www.ccmb.med.umich.edu/ccdu/SNPAAMapper
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 19,2025. A downstream variant annotation program that can effectively classify variants by region (e.g. exon, intron, etc), predict amino acid change type (e.g. synonymous, non-synonymous mutation, etc), and prioritize mutation effects (e.g. CDS versus 5?UTR, etc). Major features: * The pipeline accepts the VCF (Variant Call Format) input file in tab-delimited format and processes the vcf input file containing all cases (G5, lowFreq, and novel) * The variant mapping step has the option of letting users select whether they want to report the bp distance between each identified intron variant and its nearby exon * The pipeline can deal with VCF files called by different SAMTools versions (0.1.18 and older ones) and also offers flexibility in dealing with vcf input files generated using SAMTools with two or three samples * The spreadsheet result file contains full protein sequences for both ref and alt alleles, which makes it easier for downstream protein structure/function analysis tools to take
Proper citation: SNPAAMapper (RRID:SCR_002012) Copy
http://mutdb.org/mutpredsplice/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 6,2023. Tool for identifying coding region variants which disrupt pre-mRNA splicing and the underlying mechanism.
Proper citation: MutPred Splice (RRID:SCR_000594) Copy
http://sourceforge.net/apps/mediawiki/mummergpu/index.php?title=MUMmerGPU
Software tool as high throughput DNA sequence alignment program that runs on nVidia G80-class GPUs. Aligns sequences in parallel on video card to accelerate widely used serial CPU program MUMmer.
Proper citation: MUMmerGPU (RRID:SCR_001200) Copy
http://sourceforge.net/projects/skewer/
Software program for adapter trimming that is specially designed for processing Illumina paired-end sequences.
Proper citation: skewer (RRID:SCR_001151) Copy
https://github.com/princelab/mspire-simulator
A free, open-source shotgun proteomic simulator that goes beyond previous simulation attempts by generating LC-MS features with realistic m/z and intensity variance along with other noise components.
Proper citation: Mspire-Simulator (RRID:SCR_001431) Copy
A collection of high quality multiple sequence alignments for objective, comparative studies of alignment algorithms. The alignments are constructed based on 3D structure superposition and manually refined to ensure alignment of important functional residues. A number of subsets are defined covering many of the most important problems encountered when aligning real sets of proteins. It is specifically designed to serve as an evaluation resource to address all the problems encountered when aligning complete sequences. The first release provided sets of reference alignments dealing with the problems of high variability, unequal repartition and large N/C-terminal extensions and internal insertions. Version 2.0 of the database incorporates three new reference sets of alignments containing structural repeats, trans-membrane sequences and circular permutations to evaluate the accuracy of detection/prediction and alignment of these complex sequences.
Within the resource, users can look at a list of all the alignments, download the whole database by ftp, get the "c" program to compare a test alignment with the BAliBASE reference (The source code for the program is freely available), or look at the results of a comparison study of several multiple alignment programs, using BAliBASE reference sets.
Proper citation: BAliBASE (RRID:SCR_001940) Copy
Society that develop standards for biological research data quality, annotation and exchange. They facilitate the creation and use of software tools that build on these standards and allow researchers to annotate and share their data easily. They promote scientific discovery that is driven by genome wide and other biological research data integration and meta-analysis. Historically, FGED began with a focus on microarrays and gene expression data. However, the scope of FGED now includes data generated using any technology when applied to genome-scale studies of gene expression, binding, modification and other related applications.
Proper citation: FGED (RRID:SCR_001897) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.