Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A comprehensive encyclopedia of genomic functional elements in the model organisms C. elegans and D. melanogaster. modENCODE is run as a Research Network and the consortium is formed by 11 primary projects, divided between worm and fly, spanning the domains of gene structure, mRNA and ncRNA expression profiling, transcription factor binding sites, histone modifications and replacement, chromatin structure, DNA replication initiation and timing, and copy number variation. The raw and interpreted data from this project is vetted by a data coordinating center (DCC) to ensure consistency and completeness. The entire modENCODE data corpus is now available on the Amazon Web Services EC2 cloud. What this means is that virtual machines and virtual compute clusters that you run within the EC2 cloud can mount the modENCODE data set in whole or in part. Your software can run analyses against the data files directly without experiencing the long waits and logistics associated with copying the datasets over to your local hardware. You may also view the data using GBrowse, Dataset Search, or download the data via FTP, as well as download pre-release datasets.
Proper citation: modENCODE (RRID:SCR_006206) Copy
A clustering and visualization tool that enables the interactive exploration of genome-wide data, with a specialization in epigenomics data. Spark is also available as a service within the Epigenome toolset of the Genboree Workbench. The approach utilizes data clusters as a high-level visual guide and supports interactive inspection of individual regions within each cluster. The cluster view links to gene ontology analysis tools and the detailed region view connects to existing genome browser displays taking advantage of their wealth of annotation and functionality.
Proper citation: Spark (RRID:SCR_006207) Copy
An open-membership International community to promote mechanisms that standardize the description of genomes and the exchange and integration of genomic data. Community-driven standards have the best chance of success if developed within the auspices of international working groups. Participants in the GSC include biologists, computer scientists, those building genomic databases and conducting large-scale comparative genomic analyses, and those with experience of building community-based standards. The mission of the GSC is to work with the wider community towards: * the implementation of new genomic standards * methods of capturing and exchanging metadata * harmonization of metadata collection and analysis efforts across the wider genomics community
Proper citation: Genomic Standards Consortium (RRID:SCR_006273) Copy
http://www.uni-koeln.de/med-fak/cgars/
Software package to dissect random from non-random patterns in copy number data and thereby to assess significantly enriched somatic copy number aberrations (SCNA) across a set of tumor specimens or cell lines.
Proper citation: CGARS (RRID:SCR_006404) Copy
http://www.clipz.unibas.ch/downloads/TSSer/index.php
A computational pipeline to analyze differential RNA sequencing (dRNA-seq) data to determine transcription start sites genome-wide.
Proper citation: TSSer (RRID:SCR_006419) Copy
Model organism database that provides organization of and access to scientific data for the fission yeast Schizosaccharomyces pombe. PomBase supports genomic sequence and features, genome-wide datasets and manual literature curation. PomBase also provides a community hub for researchers, providing genome statistics, a community curation interface, news, events, documentation, mailing lists, and welcomes data submissions.
Proper citation: PomBase (RRID:SCR_006586) Copy
http://bioinformatics.ubc.ca/ermineJ/
Data analysis software for gene sets in expression microarray data or other genome-wide data that results in rankings of genes. A typical goal is to determine whether particular biological pathways are doing something interesting in the data. The software is designed to be used by biologists with little or no informatics background. A command-line interface is available for users who wish to script the use of ermineJ. Major features include: * Implementation of multiple methods for gene set analysis: ** Over-representation analysis ** A resampling-based method that uses gene scores ** A rank-based method that uses gene scores ** A resampling-based method that uses correlation between gene expression profiles (a type of cluster-enrichment analysis). * Gene sets receive statistical scores (p-values), and multiple test correction is supported. * Support of the Gene Ontology terminology; users can choose which aspects to analyze. * User files use simple text formats. * Users can modify gene sets or create new ones. * The results can be visualized within the software. * It is simple to compare multiple analyses of the same data set with different settings. * User-definable hyperlinks are provided to external sites to allow more efficient browsing of the results. * For programmers, there is a command line interface as well as a simple application programming interface that can be used to plug ermineJ functionality into your own code Platform: Online tool, Windows compatible, Mac OS X compatible, Linux compatible, Unix compatible
Proper citation: ErmineJ (RRID:SCR_006450) Copy
Public archive providing a comprehensive record of the world''''s nucleotide sequencing information, covering raw sequencing data, sequence assembly information and functional annotation. All submitted data, once public, will be exchanged with the NCBI and DDBJ as part of the INSDC data exchange agreement. The European Nucleotide Archive (ENA) captures and presents information relating to experimental workflows that are based around nucleotide sequencing. A typical workflow includes the isolation and preparation of material for sequencing, a run of a sequencing machine in which sequencing data are produced and a subsequent bioinformatic analysis pipeline. ENA records this information in a data model that covers input information (sample, experimental setup, machine configuration), output machine data (sequence traces, reads and quality scores) and interpreted information (assembly, mapping, functional annotation). Data arrive at ENA from a variety of sources including submissions of raw data, assembled sequences and annotation from small-scale sequencing efforts, data provision from the major European sequencing centers and routine and comprehensive exchange with their partners in the International Nucleotide Sequence Database Collaboration (INSDC). Provision of nucleotide sequence data to ENA or its INSDC partners has become a central and mandatory step in the dissemination of research findings to the scientific community. ENA works with publishers of scientific literature and funding bodies to ensure compliance with these principles and to provide optimal submission systems and data access tools that work seamlessly with the published literature. ENA is made up of a number of distinct databases that includes the EMBL Nucleotide Sequence Database (Embl-Bank), the newly established Sequence Read Archive (SRA) and the Trace Archive. The main tool for downloading ENA data is the ENA Browser, which is available through REST URLs for easy programmatic use. All ENA data are available through the ENA Browser. Note: EMBL Nucleotide Sequence Database (EMBL-Bank) is entirely included within this resource.
Proper citation: European Nucleotide Archive (ENA) (RRID:SCR_006515) Copy
Set of measures intended for use in large-scale genomic studies. Facilitate replication and validation across studies. Includes links to standards and resources in effort to facilitate data harmonization to legacy data. Measurement protocols that address wide range of research domains. Information about each protocol to ensure consistent data collection.Collections of protocols that add depth to Toolkit in specific areas.Tools to help investigators implement measurement protocols.
Proper citation: Phenotypes and eXposures Toolkit (RRID:SCR_006532) Copy
Database for genetic, genomic, phenotype, and disease data generated from rat research. Centralized database that collects, manages, and distributes data generated from rat genetic and genomic research and makes these data available to scientific community. Curation of mapped positions for quantitative trait loci, known mutations and other phenotypic data is provided. Facilitates investigators research efforts by providing tools to search, mine, and analyze this data. Strain reports include description of strain origin, disease, phenotype, genetics, immunology, behavior with links to related genes, QTLs, sub-strains, and strain sources.
Proper citation: Rat Genome Database (RGD) (RRID:SCR_006444) Copy
http://www.gigasciencejournal.com/
An online open-access open-data journal, publishing ''big-data'' studies from the entire spectrum of life and biomedical sciences whose publication format links standard manuscript publication with its affiliated database, GigaDB, that hosts all associated data, provides data analysis tools, cloud-computing resources, and a DOI assignment to every dataset. GigaScience covers not just ''omic'' type data and the fields of high-throughput biology currently serviced by large public repositories, but also the growing range of more difficult-to-access data, such as imaging, neuroscience, ecology, cohort data, systems biology and other new types of large-scale sharable data. Supporting the open-data movement, they require that all supporting data and source code be publicly available in a suitable public repository and/or under a public domain CC0 license in the BGI GigaScience database. Using the BGI cloud as a test environment, they also consider open-source software tools / methods for the analysis or handling of large-scale data. When submitting a manuscript, please contact them if you have datasets or cloud applications you would like them to host. To maximize data usability submitters are encouraged to follow best practice for metadata reporting and are given the opportunity to submit in ISA-Tab format.
Proper citation: GigaScience (RRID:SCR_006565) Copy
Database of Drosophila genetic and genomic information with information about stock collections and fly genetic tools. Gene Ontology (GO) terms are used to describe three attributes of wild-type gene products: their molecular function, the biological processes in which they play a role, and their subcellular location. Additionally, FlyBase accepts data submissions. FlyBase can be searched for genes, alleles, aberrations and other genetic objects, phenotypes, sequences, stocks, images and movies, controlled terms, and Drosophila researchers using the tools available from the "Tools" drop-down menu in the Navigation bar.
Proper citation: FlyBase (RRID:SCR_006549) Copy
Non profit research organization for genome sequences to advance understanding of biology of humans and pathogens in order to improve human health globally. Provides data which can be translated for diagnostics, treatments or therapies including over 100 finished genomes, which can be downloaded. Data are publicly available on limited basis, and provided more extensively upon request.
Proper citation: Wellcome Trust Sanger Institute; Hinxton; United Kingdom (RRID:SCR_011784) Copy
http://www.g-language.org/GenomeProjector/
A searchable database browser with zoomable user interface using Google Map API. Genome Projector currently contains 4 views: Genome map, Plasmid map, Pathway map, and DNA walk.
Proper citation: Genome Projector (RRID:SCR_011790) Copy
http://anntools.sourceforge.net/
Software tool for annotating single nucleotide substitutions (SNP/SNV), small insertions/deletions (indels), and copy number variations (CNV) calls generated from sequencing and microarray data. Only human genome build 37/hg19 can be annotated at this time.
Proper citation: AnnTools (RRID:SCR_005170) Copy
http://cbrc.kaust.edu.sa/readscan/
A highly scalable parallel software program to identify non-host sequences (of potential pathogen origin) and estimate their genome relative abundance in high-throughput sequence datasets.
Proper citation: READSCAN (RRID:SCR_005204) Copy
http://odin.mdacc.tmc.edu/~xsu1/VirusSeq.html
An algorithmic software tool for detecting known viruses and their integration sites using next-generation sequencing of human cancer tissue. VirusSeq takes FASTQ files (paired-end reads) as input.
Proper citation: VirusSeq (RRID:SCR_005206) Copy
http://snpeff.sourceforge.net/
Genetic variant annotation and effect prediction software toolbox that annotates and predicts effects of variants on genes (such as amino acid changes). By using standards, such as VCF, SnpEff makes it easy to integrate with other programs.
Proper citation: SnpEff (RRID:SCR_005191) Copy
A web server designed to rapidly and accurately identify, annotate and graphically display prophage sequences within bacterial genomes or plasmids. It accepts either raw DNA sequence data or partially annotated GenBank formatted data and rapidly performs a number of database comparisons as well as phage cornerstone feature identification steps to locate, annotate and display prophage sequences and prophage features. Relative to other prophage identification tools, PHAST is up to 40 times faster and up to 15% more sensitive. It is also able to process and annotate both raw DNA sequence data and Genbank files, provide richly annotated tables on prophage features and prophage quality and distinguish between intact and incomplete prophage. PHAST also generates downloadable, high quality, interactive graphics that display all identified prophage components in both circular and linear genomic views. Databases available for download include Virus DB, Prophage and virus DB, Bacteria DB, and PHAST result DB. Pre-calculated genomes for viewing are also available.
Proper citation: PHAge Search Tool (RRID:SCR_005184) Copy
http://code.google.com/p/snpdat/
A simple and easy to use high through-put analysis tool which can provide comprehensive annotation of both novel and known single nucleotide polymorphisms (SNPs) for any organism with a draft sequence and annotation. SNPdat makes possible analyses involving non-model organisms that are not supported by the vast majority of SNP annotation tools currently available. It is especially intended for use by researchers with limited bioinformatic experience.
Proper citation: SNPdat (RRID:SCR_005187) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.