Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Multi-organism, publicly accessible compendium of peptides identified in a large set of tandem mass spectrometry proteomics experiments. Mass spectrometer output files are collected for human, mouse, yeast, and several other organisms, and searched using the latest search engines and protein sequences. All results of sequence and spectral library searching are subsequently processed through the Trans Proteomic Pipeline to derive a probability of correct identification for all results in a uniform manner to insure a high quality database, along with false discovery rates at the whole atlas level. The raw data, search results, and full builds can be downloaded for other uses. All results of sequence searching are processed through PeptideProphet to derive a probability of correct identification for all results in a uniform manner ensuring a high quality database. All peptides are mapped to Ensembl and can be viewed as custom tracks on the Ensembl genome browser. The long term goal of the project is full annotation of eukaryotic genomes through a thorough validation of expressed proteins. The PeptideAtlas provides a method and a framework to accommodate proteome information coming from high-throughput proteomics technologies. The online database administers experimental data in the public domain. You are encouraged to contribute to the database.
Proper citation: PeptideAtlas (RRID:SCR_006783) Copy
http://www.bioconductor.org/packages/2.11/bioc/html/ShortRead.html
Software package for input, quality assessment and exploration of high-throughput sequence data. Used for input, quality assurance, and basic manipulation of `short read'' DNA sequences such as those produced by Solexa, 454, and related technologies, including exible import of common short read data formats.
Proper citation: ShortRead (RRID:SCR_006813) Copy
http://www.ensemblgenomes.org/
Database portal offering integrated access to genome-scale data from non-vertebrate species of scientific interest, developed using the Ensembl genome annotation and visualization platform. Ensembl Genomes consists of five sub-portals (for bacteria, protists, fungi, plants and invertebrate metazoa) designed to complement the availability of vertebrate genomes in Ensembl. Many of the databases supporting the portal have been built in close collaboration with the scientific community - essential for maintaining the accuracy and usefulness of the resource. A common set of user interfaces (which include a graphical genome browser, FTP, BLAST search, a query optimized data warehouse, programmatic access, and a Perl API) is provided for all domains. Data types incorporated include annotation of (protein and non-protein coding) genes, cross references to external resources, and high throughput experimental data (e.g. data from large scale studies of gene expression and polymorphism visualized in their genomic context). Additionally, extensive comparative analysis has been performed, both within defined clades and across the wider taxonomy, and sequence alignments and gene trees resulting from this can be accessed through the site.
Proper citation: Ensembl Genomes (RRID:SCR_006773) Copy
canSAR is an integrated database that brings together biological, chemical, pharmacological (and eventually clinical) data. Its goal is to integrate this data and make it accessible to cancer research scientists from multiple disciplines, in order to help with hypothesis generation in cancer research and support translational research. This cancer research and drug discovery resource was developed to utilize the growing publicly available biological annotation, chemical screening, RNA interference screening, expression, amplification and 3D structural data. Scientists can, in a single place, rapidly identify biological annotation of a target, its structural characterization, expression levels and protein interaction data, as well as suitable cell lines for experiments, potential tool compounds and similarity to known drug targets. canSAR has, from the outset, been completely use-case driven which has dramatically influenced the design of the back-end and the functionality provided through the interfaces. The Web interface provides flexible, multipoint entry into canSAR. This allows easy access to the multidisciplinary data within, including target and compound synopses, bioactivity views and expert tools for chemogenomic, expression and protein interaction network data.
Proper citation: canSAR (RRID:SCR_006794) Copy
Curated collection of known Drosophila transcriptional cis-regulatory modules (CRMs) and transcription factor binding sites (TFBSs). Includes experimentally verified fly regulatory elements along with their DNA sequence, associated genes, and expression patterns they direct. Submission of experimentally verified cis-regulatory elements that are not included in REDfly database are welcome.
Proper citation: REDfly Regulatory Element Database for Drosophilia (RRID:SCR_006790) Copy
https://github.com/friend1ws/EBCall
A software package for somatic mutation detection (including InDels). EBCall uses not only paired tumor/normal sequence data of a target sample, but also multiple non-paired normal reference samples for evaluating distribution of sequencing errors, which leads to an accurate mutaiton detection even in case of low sequencing depths and low allele frequencies.
Proper citation: EBCall (RRID:SCR_006791) Copy
Centralized, standards compliant, public data repository for proteomics data, including protein and peptide identifications, post-translational modifications and supporting spectral evidence. Originally it was developed to provide a common data exchange format and repository to support proteomics literature publications. This remit has grown with PRIDE, with the hope that PRIDE will provide a reference set of tissue-based identifications for use by the community. The future development of PRIDE has become closely linked to HUPO PSI. PRIDE encourages and welcomes direct user submissions of protein and peptide identification data to be published in peer-reviewed publications. Users may Browse public datasets, use PRIDE BioMart for custom queries, or download the data directly from the FTP site. PRIDE has been developed through a collaboration of the EMBL-EBI, Ghent University in Belgium, and the University of Manchester.
Proper citation: Proteomics Identifications (PRIDE) (RRID:SCR_003411) Copy
https://github.com/ggloor/ALDEx2
Software tool to examine compositional high-throughput sequence data with Welch's t-test. A differential relative count abundance analysis for the comparison of two conditions. For example, single-organism and meta-rna-seq high-throughput sequencing assays, or of selected and unselected values from in-vitro sequence selections. Uses a Dirichlet-multinomial model to infer abundance from counts, that has been optimized for three or more experimental replicates. Infers sampling variation and calculates the expected Benjamini-Hochberg false discovery rate given the biological and sampling variation using several parametric and non-parametric tests. Can to glm and Kruskal-Wallace tests on one-way ANOVA style designs.
Proper citation: ALDEx2 (RRID:SCR_003364) Copy
http://www.bioconductor.org/packages/release/bioc/html/ggbio.html
An R package for extending the grammar of graphics for genomic data. The graphics are designed to answer common scientific questions, in particular those often asked of high throughput genomics data. All core Bioconductor data structures are supported, where appropriate. The package supports detailed views of particular genomic regions, as well as genome-wide overviews. Supported overviews include ideograms and grand linear views. High-level plots include sequence fragment length, edge-linked interval to data view, mismatch pileup, and several splicing summaries.
Proper citation: ggbio (RRID:SCR_003313) Copy
https://bitbucket.org/dranew/defuse
Software package for gene fusion discovery using RNA-Seq data. It uses clusters of discordant paired end alignments to inform a split read alignment analysis for finding fusion boundaries.
Proper citation: deFuse (RRID:SCR_003279) Copy
http://primerseq.sourceforge.net/
Software that designs RT-PCR primers that evaluate alternative splicing events by incorporating RNA-Seq data. It is particularly advantageous for designing a large number of primers for validating alternative splicing events found in RNA-Seq data. It incorporates RNA-Seq data in the design process to weight exons by their read counts. Essentially, the RNA-Seq data allows primers to be placed using actually expressed transcripts. This could be for a particular cell line or experimental condition, rather than using annotations that incorporate transcripts that are not expressed for the data. Alternatively, you can design primers that are always on constitutive exons. PrimerSeq does not limit the use of gene annotations and can be used for a wide array of species.
Proper citation: PrimerSeq (RRID:SCR_003295) Copy
http://shendurelab.github.io/MIPGEN/
Software for a fast, simple way to generate designs for MIP assays targeting hundreds or thousands of genomic loci in parallel. Packaged with MIPgen are scripts that aid in visualization of MIP designs and processing of MIP sequence reads to SAM files that can then be passed through any standard variant calling pipeline.
Proper citation: MIPgen (RRID:SCR_003325) Copy
https://bitbucket.org/johanneskoester/snakemake/wiki/
A Python based language and execution environment for make-like workflows. The system supports the use of automatically inferred multiple named wildcards (or variables) in input and output filenames.
Proper citation: Snakemake (RRID:SCR_003475) Copy
http://knowledgemap.mc.vanderbilt.edu/research/content/phewas-r-package
Software package contains methods for performing Phenome-Wide Association Study.
Proper citation: PheWAS R Package (RRID:SCR_003512) Copy
http://www.cellimagelibrary.org/
Freely accessible, public repository of vetted and annotated microscopic images, videos, and animations of cells from a variety of organisms, showcasing cell architecture, intracellular functionalities, and both normal and abnormal processes. Explore by Cell Process, Cell Component, Cell Type or Organism. The Cell includes images acquired from historical and modern collections, publications, and by recruitment.
Proper citation: Cell Image Library (CIL) (RRID:SCR_003510) Copy
https://code.google.com/p/bpipe/
Software tool for running and managing bioinformatics pipelines. It specializes in enabling users to turn existing pipelines based on shell scripts or command line tools into highly flexible, adaptable and maintainable workflows with a minimum of effort. Bpipe ensures that pipelines execute in a controlled and repeatable fashion and keeps audit trails and logs to ensure that experimental results are reproducible. Requiring only Java as a dependency, it is fully self-contained and cross-platform, making it very easy to adopt and deploy into existing environments.
Proper citation: Bpipe (RRID:SCR_003471) Copy
http://www.lgm.upmc.fr/parseq/
Statistical software for transcription landscape reconstruction at a basepair resolution from RNA Seq read counts. It is based on a state-space model which describes, in terms of abrupt shifts and more progressive drifts, the transcription level dynamics along the genome. Alongside variations of transcription level, it incorporates a component of short-range variation to pull apart local artifacts causing correlated dispersion. Reconstruction of the transcription level relies on a conditional sequential Monte Carlo approach that is combined with parameter estimation in a Markov chain Monte Carlo algorithm known as particle Gibbs. The method allows to estimate the local transcription level, to call transcribed regions, and to identify the transcript borders.
Proper citation: Parseq (RRID:SCR_003464) Copy
http://cran.r-project.org/web/packages/MultiPhen/
Software package that performs genetic association tests between SNPs (one-at-a-time) and multiple phenotypes (separately or in joint model).
Proper citation: MultiPhen (RRID:SCR_003498) Copy
http://www.biostat.wisc.edu/~kendzior/EBSEQ/
Software R package for RNA-Seq Differential Expression Analysis.
Proper citation: EBSeq (RRID:SCR_003526) Copy
Collection of pathways and pathway annotations. The core unit of the Reactome data model is the reaction. Entities (nucleic acids, proteins, complexes and small molecules) participating in reactions form a network of biological interactions and are grouped into pathways (signaling, innate and acquired immune function, transcriptional regulation, translation, apoptosis and classical intermediary metabolism) . Provides website to navigate pathway knowledge and a suite of data analysis tools to support the pathway-based analysis of complex experimental and computational data sets.
Proper citation: Reactome (RRID:SCR_003485) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.