Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://sourceforge.net/projects/dnaclust/
Software program for clustering large number of short similar DNA sequences. It was originally designed for clustering targeted 16S rRNA pyrosequencing reads.
Proper citation: DNACLUST (RRID:SCR_001771) Copy
An open source data warehouse system built for the integration and analysis of complex biological data that enables the creation of biological databases accessed by sophisticated web query tools. Parsers are provided for integrating data from many common biological data sources and formats, and there is a framework for adding data. InterMine includes a user-friendly web interface that works "out of the box" and can be easily customized for specific needs, as well as a powerful, scriptable web-service API to allow programmatic access to data.
Proper citation: InterMine (RRID:SCR_001772) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 23,2022. Friend is a bioinformatics application designed for simultaneous analysis and visualization of multiple structures and sequences of proteins and/or DNA/RNA. The application provides basic functionalities such as: structure visualization with different rendering and coloring, sequence alignment, and simple phylogeny analysis, along with a number of extended features to perform more complex analyses of sequence structure relationships, including: structural alignment of proteins, investigation of specific interaction motifs, studies of protein-protein and protein-DNA interactions, and protein super-families. Friend is also useful for the functional annotation of proteins, protein modeling, and protein folding studies. Friend provides three levels of usage; 1) an extensive GUI for a scientist with no programming experience, 2) a command line interface for scripting for a scientist with some programming experience, and 3) the ability to extend Friend with user written libraries for an experienced programmer. The application is linked and communicates with local and remote sequence and structure databases.
Proper citation: An Integrated Multiple Structure Visualization and Multiple Sequence Alignment Application (RRID:SCR_001646) Copy
http://protein.bio.unipd.it/pasta2/
Online interface that utilizes an algorithm to predict the most aggregation-prone portions and the corresponding beta-strand inter-molecular pairing for a given input sequence. Users can paste the sequence into the interface and output the appropriate sequence.
Proper citation: Prediction of Amyloid Structure Aggregation (RRID:SCR_001768) Copy
Issue
http://www.nitrc.org/projects/plink
Open source whole genome association analysis toolset, designed to perform range of basic, large scale analyses in computationally efficient manner. Used for analysis of genotype/phenotype data. Through integration with gPLINK and Haploview, there is some support for subsequent visualization, annotation and storage of results. PLINK 1.9 is improved and second generation of the software.
Proper citation: PLINK (RRID:SCR_001757) Copy
http://www.bioconductor.org/packages/release/bioc/html/unifiedWMWqPCR.html
Software package that implements the unified Wilcoxon-Mann-Whitney Test for qPCR data. This modified test allows for testing differential expression in qPCR data.
Proper citation: unifiedWMWqPCR (RRID:SCR_001706) Copy
http://www.bioconductor.org/packages/2.13/bioc/html/cqn.html
A normalization tool for RNA-Seq data, implementing the conditional quantile normalization method.
Proper citation: CQN (RRID:SCR_001786) Copy
Suite of motif-based sequence analysis tools to discover motifs using MEME, DREME (DNA only) or GLAM2 on groups of related DNA or protein sequences; search sequence databases with motifs using MAST, FIMO, MCAST or GLAM2SCAN; compare a motif to all motifs in a database of motifs; associate motifs with Gene Ontology terms via their putative target genes, and analyze motif enrichment using SpaMo or CentriMo. Source code, binaries and a web server are freely available for noncommercial use.
Proper citation: MEME Suite - Motif-based sequence analysis tools (RRID:SCR_001783) Copy
http://cmb.gis.a-star.edu.sg/ChIPSeq/paperCCAT.htm
THIS RESOURCE IS OUT OF SERVICE, documented on April 5, 2017, A software package for the analysis of ChIP-seq data with negative control., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: CCAT (RRID:SCR_001843) Copy
http://www.genabel.org/packages/GenABEL
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. R software library for genome-wide association analysis for quantitative, binary and time-till-event traits.
Proper citation: GenABEL (RRID:SCR_001842) Copy
http://surfer.nmr.mgh.harvard.edu/
Open source software suite for processing and analyzing human brain MRI images. Used for reconstruction of brain cortical surface from structural MRI data, and overlay of functional MRI data onto reconstructed surface. Contains automatic structural imaging stream for processing cross sectional and longitudinal data. Provides anatomical analysis tools, including: representation of cortical surface between white and gray matter, representation of the pial surface, segmentation of white matter from rest of brain, skull stripping, B1 bias field correction, nonlinear registration of cortical surface of individual with stereotaxic atlas, labeling of regions of cortical surface, statistical analysis of group morphometry differences, and labeling of subcortical brain structures.Operating System: Linux, macOS.
Proper citation: FreeSurfer (RRID:SCR_001847) Copy
https://urgi.versailles.inra.fr/Tools/S-Mart
Software toolbox that manages your RNA-Seq and ChIP-Seq data and also produces many different plots to visualize your data. It performs several tasks that are usually required during the analysis of mapped RNA-Seq and ChIP-Seq reads, including data selection and data visualization. It includes the selection (or the exclusion) of the data that overlaps with a reference set, clustering and comparative analysis. It also provides many ways to visualize data: size of the reads, density on the genome, distance with respect to a reference set, and the correlation of two data sets (with cloud plots). A computer science background is not required to run it through a graphical interface and it can be run on any personal computer, yielding results within an hour for most queries.
Proper citation: S-MART (RRID:SCR_001908) Copy
http://www.bioconductor.org/packages/release/bioc/html/SamSPECTRAL.html
Software that identifies cell population in flow cytometry data. It demonstrates significant advantages in proper identification of populations with non-elliptical shapes, low density populations close to dense ones, minor subpopulations of a major population and rare populations. It samples large data such that spectral clustering is possible while preserving density information in edge weights. More specifically, given a matrix of coordinates as input, SamSPECTRAL first builds the communities to sample the data points. Then, it builds a graph and after weighting the edges by conductance computation, the graph is passed to a classic spectral clustering algorithm to find the spectral clusters. The last stage of SamSPECTRAL is to combine the spectral clusters. The resulting connected components estimate biological cell populations in the data sample.
Proper citation: SamSPECTRAL (RRID:SCR_001858) Copy
http://bioconductor.org/packages/2.9/bioc/html/RamiGO.html
Software package with an R interface sending requests to AmiGO visualize, retrieving DAG GO trees, parsing GraphViz DOT format files and exporting GML files for Cytoscape. Also uses RCytoscape to interactively display AmiGO trees in Cytoscape.
Proper citation: RamiGO (RRID:SCR_006922) Copy
http://www.ebi.ac.uk/thornton-srv/databases/WSsas/
SAS is a tool for applying structural information to a given protein sequence. It uses FASTA to scan a given protein sequence against all the proteins of known 3D structure in the Protein Data Bank and provides functional residue annotation based on data from the Catalytic Site Atlas and PDBsum. The web service is aimed to facilitate the use of the SAS tool when having a huge number of queries. Currently, the web service provides annotation for binding sites (to ligand, metal or nucleic acid), catalytic residues and amino acids related to protein-protein interactions.
Proper citation: WSsas - Web Service for the SAS tool (RRID:SCR_007051) Copy
http://autismkb.cbi.pku.edu.cn/
Genetic factors contribute significantly to ASD. AutismKB is an evidence-based knowledgebase of Autism spectrum disorder (ASD) genetics. The current version contains 2193 genes (99 syndromic autism related genes and 2135 non-syndromic autism related genes), 4617 Copy Number Variations (CNVs) and 158 linkage regions associated with ASD by one or more of the following six experimental methods: # Genome-Wide Association Studies (GWAS); # Genome-wide CNV studies; # Linkage analysis; # Low-scale genetic association studies; # Expression profiling; # Other low-scale gene studies. Based on a scoring and ranking system, 99 syndromic autism related genes and 383 non-syndromic autism related genes (434 genes in total) were designated as having high confidence. Autism spectrum disorder (ASD) is a heterogeneous neurodevelopmental disorder with a prevalence of 1.0-2.6%. The three core symptoms of ASD are: # impairments in reciprocal social interaction; # communication impairments; # presence of restricted, repetitive and stereotyped patterns of behavior, interests and activities.
Proper citation: AutismKB (RRID:SCR_006937) Copy
http://bowtie-bio.sourceforge.net/myrna/index.shtml
A cloud computing tool for calculating differential gene expression in large RNA-seq datasets. It uses Bowtie for short read alignment and R/Bioconductor for interval calculations, normalization, and statistical testing. These tools are combined in an automatic, parallel pipeline that runs in the cloud (Elastic MapReduce in this case) on a local Hadoop cluster, or on a single computer, exploiting multiple computers and CPUs wherever possible.
Proper citation: Myrna (RRID:SCR_006951) Copy
https://github.com/jstjohn/SimSeq
An illumina paired-end and mate-pair short read simulator. This project attempts to model as many of the quirks that exist in Illumina data as possible. Some of these quirks include the potential for chimeric reads, and non-biotinylated fragment pull down in mate-pair libraries .
Proper citation: SimSeq (RRID:SCR_006947) Copy
http://weizhong-lab.ucsd.edu/cd-hit/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: CD-HIT (RRID:SCR_007105) Copy
This database provides a platform to query and compare gene expression data during the development of the major model animals (zebrafish, drosophila, medaka, mouse). The name 4DXpress stands for expression database in 4D. The 4D (four dimensions) of 4DXpress can be interpreted either as: 3 spatial dimensions plus time, or as 1. species 2. gene 3. developmental stage 4. anatomical structure. The major focus of this database lies in cross species comparison. The high resolution expression data was acquired through whole mount in situ hybridsation-, antibody- or transgenic experiments. Data was integrated from several species specific expression pattern databases, such as ZFIN, BDGP, GXD, MEPD as well as directly submitted by researchers of the participating groups at EMBL. The 4DXpress database is a project within the Centre for Computational Biology at EMBL. It is developed by Yannick Haudry, Thorsten Henrich and Ivica Letunic and coordinated by Thorsten Henrich. Hugo Berube is developing the 4D ArrayExpress Data Warehouse at EBI for integrating in situ data with microarray data.
Proper citation: Expression Database in 4D (RRID:SCR_007066) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.