Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://sourceforge.net/projects/bless-ec/
Software tool for Bloom-filter-based error correction for next-generation sequencing (NGS) reads. The algorithm produces accurate correction results with much less memory.
Proper citation: BLESS (RRID:SCR_005963) Copy
http://jjwanglab.org:8080/gwasdb/
Combines collections of genetic variants (GVs) from GWAS and their comprehensive functional annotations, as well as disease classifications. Used to maximize utilility of GWAS data to gain biological insights through integrative, multi-dimensional functional annotation portal. In addition to all GVs annotated in NHGRI GWAS Catalog, we manually curate GVs that are marginally significant (P value < 10-3) by looking into supplementary materials of each original publication and provide extensive functional annotations for these GVs. GVs are manually classified by diseases according to Disease Ontology Lite and HPO (Human Phenotype Ontology) for easy access. Database can also conduct gene based pathway enrichment and PPI network association analysis for those diseases with sufficient variants. SOAP services are available. You may Download GWASdb SNP. (This file contains all of the significant SNP in GWASdb. In the pvalue column, 0 means this P-value is not reported in the study but it is significant SNP. In the source column, GWAS:A represents the original data in GWAS catalog, while GWAS:B is our curation data which P-value < 10-3)
Proper citation: GWASdb (RRID:SCR_006015) Copy
http://equilibrator.weizmann.ac.il/
Web interface designed for thermodynamic analysis of biochemical systems. eQuilibrator enables free-text search for biochemical compounds and reactions and provides thermodynamic estimates for both in a variety of conditions. It can provide estimates for compounds in the KEGG database, and individual compounds and enzymes can be searched for by their common names (water, glucosamine, hexokinase). Reactions can be entered in a free-text format that eQuilibrator parses automatically. eQuilibrator also allows manipulation of the conditions of a reaction - pH, ionic strength, and reactant and product concentrations.
Proper citation: eQuilibrator (RRID:SCR_006011) Copy
http://aias.biol.uoa.gr/OMPdb/
A database of Beta-barrel outer membrane proteins from Gram-negative bacteria. The web interface of OMPdb offers the user the ability not only to view the available data, but also to submit advanced queries for text search within the database''s protein entries or run BLAST searches against the database. The most up-to-date version of the database (as well as all past versions) can be downloaded in various formats (flat text, XML format or raw FASTA sequences). For constructing OMPdb, multiple freely accessible resources were combined and a detailed literature search was performed. The classification of OMPdb''s protein entries into families is based mainly on structural and functional criteria. Information included in the database consists of sequence data, as well as annotation for structural characteristics (such as the transmembrane segments), literature references and links to other public databases, features that are unique worldwide. Along with the database, a collection of profile Hidden Markov Models that were shown to be characteristic for Beta-barrel outer membrane proteins was also compiled. This set, when used in combination with our previously developed algorithms (PRED-TMBB, MCMBB and ConBBPRED) will serve as a powerful tool in matters of discrimination and classification of novel Beta-barrel proteins and whole-genome analyses., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: OMPdb (RRID:SCR_006221) Copy
http://www.bioconductor.org/packages/devel/bioc/html/deepSNV.html
Software package that provides quantitative variant callers for detecting subclonal mutations in ultra-deep (>=100x coverage) sequencing experiments. The algorithm is used for a comparative setup with a control experiment of the same loci and uses a beta-binomial model and a likelihood ratio test to discriminate sequencing errors and subclonal SNVs (single nucleotide variants).
Proper citation: deepSNV (RRID:SCR_006214) Copy
http://evolution.genetics.washington.edu/phylip.html
A free package of software programs for inferring phylogenies (evolutionary trees). The source code is distributed (in C), and executables are also distributed. In particular, already-compiled executables are available for Windows (95/98/NT/2000/me/xp/Vista), Mac OS X, and Linux systems. Older executables are also available for Mac OS 8 or 9 systems.
Proper citation: PHYLIP (RRID:SCR_006244) Copy
http://stormo.wustl.edu/ScerTF
Catalog of over 1,200 position weight matrices (PWMs) for 196 different yeast transcription factors (TFs). They've curated 11 literature sources, benchmarked the published position-specific scoring matrices against in-vivo TF occupancy data and TF deletion experiments, and combined the most accurate models to produce a single collection of the best performing weight matrices for Saccharomyces cerevisiae. ScerTF is useful for a wide range of problems, such as linking regulatory sites with transcription factors, identifying a transcription factor based on a user-input matrix, finding the genes bound/regulated by a particular TF, and finding regulatory interactions between transcription factors. Enter a TF name to find the recommended matrix for a particular TF, or enter a nucleotide sequence to identify all TFs that could bind a particular region.
Proper citation: ScerTF (RRID:SCR_006121) Copy
http://www-bionet.sscc.ru/sitex/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 19,2019. Analyzing protein structure projection on exon-intron structure of corresponding gene through years led to several fundamental conclusions about structural and functional organization of the protein. According to these results we decided to map the protein functional sites. So we created the database SitEx that keep the information about this mapping and included the BLAST search and 3D similar structure search using PDB3DScan for the polypeptide encoded by one exon, participating in organizing the functional site. This will help: # to study the positions of the functional sites in exon structure; # to make the complex analysis of the protein function; # to exposure the exons that took part in exon shuffling and came from bacterial genomes; # to study the peculiarities of coding the polypeptide structures. Currently, SitEx contains information about 9994 functional sites presented in 2021 proteins described in proteomes of 17 organisms.
Proper citation: SitEx (RRID:SCR_006122) Copy
https://compbio.dfci.harvard.edu/predictivenetworks//
A flexible, open-source, web-based application and data services framework that enables the integration, navigation, visualization and analysis of gene interaction networks. The primary goal of PN is to allow biomedical researchers to evaluate experimentally derived gene lists in the context of large-scale gene interaction networks. The PN analytical pipeline involves two key steps. The first is the collection of a comprehensive set of known gene interactions derived from a variety of publicly available sources. The second is to use these ''known'' interactions together with gene expression data to infer robust gene networks. The regression-based network inference algorithm creates a graph of gene interactions in which cycles may be present (but no self-loops). Based on information-theoretic techniques, a causal gene interaction network is inferred from both prior knowledge (interactions extracted from biomedical literature and structured biological databases) and gene expression data. A prediction model is fitted for each gene, given its parents, enabling assessment of the predictive ability of the network model.
Proper citation: Predictive Networks (RRID:SCR_006110) Copy
http://202.38.126.151:8080/SDisease/
Curated database of experimentally supported data of RNA Splicing mutation and disease. The RNA Splicing mutations include cis-acting mutations that disrupt splicing and trans-acting mutations that affecting RNA-dependent functions that cause disease. Information such as EntrezGeneID, gene genomic sequence, mutation (nucleotide substitutions, deletions and insertions), mutation location within the gene, organism, detailed description of the splicing mutation and references are also given. Users are able to submit new entries to the database. This database integrating RNA splicing and disease associations would be helpful for understanding not only the RNA splicing but also its contribution to disease. In SpliceDisease database, they manually curated 2337 splicing mutation disease entries involving 303 genes and 370 diseases, which have been supported experimentally in 898 publications. The SpliceDisease database provides information including the change of the nucleotide in the sequence, the location of the mutation on the gene, the reference PubMed ID and detailed description for the relationship among gene mutations, splicing defects and diseases. They standardized the names of the diseases and genes and provided links for these genes to NCBI and UCSC genome browser for further annotation and genomic sequences. For the location of the mutation, they give direct links of the entry to the respective position/region in the genome browser.
Proper citation: SpliceDisease (RRID:SCR_006130) Copy
http://weizhong-lab.ucsd.edu/cd-hit/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: CD-HIT (RRID:SCR_007105) Copy
This database provides a platform to query and compare gene expression data during the development of the major model animals (zebrafish, drosophila, medaka, mouse). The name 4DXpress stands for expression database in 4D. The 4D (four dimensions) of 4DXpress can be interpreted either as: 3 spatial dimensions plus time, or as 1. species 2. gene 3. developmental stage 4. anatomical structure. The major focus of this database lies in cross species comparison. The high resolution expression data was acquired through whole mount in situ hybridsation-, antibody- or transgenic experiments. Data was integrated from several species specific expression pattern databases, such as ZFIN, BDGP, GXD, MEPD as well as directly submitted by researchers of the participating groups at EMBL. The 4DXpress database is a project within the Centre for Computational Biology at EMBL. It is developed by Yannick Haudry, Thorsten Henrich and Ivica Letunic and coordinated by Thorsten Henrich. Hugo Berube is developing the 4D ArrayExpress Data Warehouse at EBI for integrating in situ data with microarray data.
Proper citation: Expression Database in 4D (RRID:SCR_007066) Copy
Database containing the DNA sequence and annotation of the entire human chromosome 7, encompassing nearly 158 million nucleotides of DNA and 1917 gene structures, are presented; the most up to date collation of sequence, gene, and other annotations from all databases (eg. Celera published, NCBI, Ensembl, RIKEN, UCSC) as well as unpublished data. To generate a higher order description, additional structural features such as imprinted genes, fragile sites, and segmental duplications were integrated at the level of the DNA sequence with medical genetic data, including 440 chromosome rearrangement breakpoints associated with disease. The objective of this project is to generate a comprehensive description of human chromosome 7 to facilitate biological discovery, disease gene research and medical genetic applications. There are over 360 disease-associated genes or loci on chromosome 7. A major challenge ahead will be to represent chromosome alterations, variants, and polymorphisms and their related phenotypes (or lack thereof), in an accessible way. In addition to being a primary data source, this site serves as a weighing station for testing community ideas and information to produce highly curated data to be submitted to other databases such as NCBI, Ensembl, and UCSC. Therefore, any useful data submitted will be curated and shown in this database. All Chromosome 7 genomic clones (cosmids, BACs, YACs) listed in GBrowser and in other data tables are freely distributed.
Proper citation: Chromosome 7 Annotation Project (RRID:SCR_007134) Copy
Curated protein-protein and genetic interaction repository of raw protein and genetic interactions from major model organism species, with data compiled through comprehensive curation efforts.
Proper citation: Biological General Repository for Interaction Datasets (BioGRID) (RRID:SCR_007393) Copy
http://bioinfo3d.cs.tau.ac.il/FlexProt/
FlexProt detects the optimal flexible structural alignment of a pair of protein structures. The first structure is assumed to be rigid, while in the second structure potential flexible regions are automatically detected.
Proper citation: FlexProt: flexible protein alignment (RRID:SCR_007306) Copy
http://sourceforge.net/projects/taipan/
A fast hybrid short-read assembly tool.
Proper citation: Taipan (RRID:SCR_007330) Copy
http://genotan.sourceforge.net/
A free software tool to identify length variation of microsatellites from short sequence reads.
Proper citation: GenoTan (RRID:SCR_007935) Copy
Resource for experimentally validated human and mouse noncoding fragments with gene enhancer activity as assessed in transgenic mice. Most of these noncoding elements were selected for testing based on their extreme conservation in other vertebrates or epigenomic evidence (ChIP-Seq) of putative enhancer marks. Central public database of experimentally validated human and mouse noncoding fragments with gene enhancer activity as assessed in transgenic mice. Users can retrieve elements near single genes of interest, search for enhancers that target reporter gene expression to particular tissue, or download entire collections of enhancers with defined tissue specificity or conservation depth.
Proper citation: VISTA Enhancer Browser (RRID:SCR_007973) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 26,2019. In October 2016, T1DBase has merged with its sister site ImmunoBase (https://immunobase.org). Documented on March 2020, ImmunoBase ownership has been transferred to Open Targets (https://www.opentargets.org). Results for all studies can be explored using Open Targets Genetics (https://genetics.opentargets.org). Database focused on genetics and genomics of type 1 diabetes susceptibility providing a curated and integrated set of datasets and tools, across multiple species, to support and promote research in this area. The current data scope includes annotated genomic sequences for suspected T1D susceptibility regions; genetic data; microarray data; and global datasets, generally from the literature, that are useful for genetics and systems biology studies. The site also includes software tools for analyzing the data.
Proper citation: T1DBase (RRID:SCR_007959) Copy
http://www.ebi.ac.uk/huber-srv/hilbert/
Software tool that allows to display very long data vectors in a space-efficient manner, allowing the user to visually judge the large scale structure and distribution of features simultaneously with the rough shape and intensity of individual features.
Proper citation: HilbertVis (RRID:SCR_007862) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.