Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
The Distributed Annotation System (DAS) defines a communication protocol used to exchange annotations on genomic or protein sequences. It is motivated by the idea that such annotations should not be provided by single centralized databases, but should instead be spread over multiple sites. Data distribution, performed by DAS servers, is separated from visualization, which is done by DAS clients. The advantages of this system are that control over the data is retained by data providers, data is freed from the constraints of specific organisations and the normal issues of release cycles, API updates and data duplication are avoided. DAS is a client-server system in which a single client integrates information from multiple servers. It allows a single machine to gather up sequence annotation information from multiple distant web sites, collate the information, and display it to the user in a single view. Little coordination is needed among the various information providers. DAS is heavily used in the genome bioinformatics community. Over the last years we have also seen growing acceptance in the protein sequence and structure communities. A DAS-enabled website or application can aggregate complex and high-volume data from external providers in an efficient manner. For the biologist, this means the ability to plug in the latest data, possibly including a user''s own data. For the application developer, this means protection from data format changes and the ability to add new data with minimal development cost. Here are some examples of DAS-enabled applications or websites for end users: :- Dalliance Experimental Web/Javascript based Genome Viewer :- IGV Integrative Genome Viewer java based browser for many genomes :- Ensembl uses DAS to pull in genomic, gene and protein annotations. It also provides data via DAS. :- Gbrowse is a generic genome browser, and is both a consumer and provider of DAS. :- IGB is a desktop application for viewing genomic data. :- SPICE is an application for projecting protein annotations onto 3D structures. :- Dasty2 is a web-based viewer for protein annotations :- Jalview is a multiple alignment editor. :- PeppeR is a graphical viewer for 3D electron microscopy data. :- DASMI is an integration portal for protein interaction data. :- DASher is a Java-based viewer for protein annotations. :- EpiC presents structure-function summaries for antibody design. :- STRAP is a STRucture-based sequence Alignment Program. Hundreds of DAS servers are currently running worldwide, including those provided by the European Bioinformatics Institute, Ensembl, the Sanger Institute, UCSC, WormBase, FlyBase, TIGR, and UniProt. For a listing of all available DAS sources please visit the DasRegistry. Sponsors: The initial ideas for DAS were developed in conversations with LaDeana Hillier of the Washington University Genome Sequencing Center.
Proper citation: Distributed Annotation System (RRID:SCR_008427) Copy
http://bioinf.uni-greifswald.de/augustus/
Software for gene prediction in eukaryotic genomic sequences. Serves as a basis for further steps in the analysis of sequenced and assembled eukaryotic genomes.
Proper citation: Augustus (RRID:SCR_008417) Copy
Web application for simulating SNP genotypes for case-control and affected-child trio studies by resampling from Phase I/II HapMap SNP data. The user provides a list of SNPs to be genotyped, along with a disease model file that describes causal SNPs and their effect sizes. The simulation tool is appropriate for candidate regions or whole-genome scans. (entry from Genetic Analysis Software)
Proper citation: HAP-SAMPLE (RRID:SCR_009234) Copy
http://pages.stat.wisc.edu/~yandell/qtl/software/qtlbim/
Software library for QTL Bayesian Interval Mapping that provides a Bayesian model selection approach to map multiple interacting QTL. It works on experimentally inbred lines and performs a genome-wide search to locate multiple potential QTL. The package can handle continuous, binary and ordinal traits. (entry from Genetic Analysis Software)
Proper citation: R/QTLBIM (RRID:SCR_009375) Copy
Project focused on cerebral aneurysms and provides integrated decision support system to assess risk of aneurysm rupture in patients and to optimize their treatments. IT infrastructure has been developeded for management and processing of vast amount of heterogeneous data acquired during diagnosis.
Proper citation: aneurIST (RRID:SCR_007427) Copy
http://human.brain-map.org/static/brainexplorer
Multi modal atlas of human brain that integrates anatomic and genomic information, coupled with suite of visualization and mining tools to create open public resource for brain researchers and other scientists. Data include magnetic resonance imaging (MRI), diffusion tensor imaging (DTI), histology and gene expression data derived from both microarray and in situ hybridization (ISH) approaches. Brain Explorer 2 is desktop software application for viewing human brain anatomy and gene expression data in 3D.
Proper citation: Allen Human Brain Atlas (RRID:SCR_007416) Copy
This is a blog about post genomic knowledge. The website''s goal is to make public datasets from the bioinformatics community available in RDF format via standard SPARQL endpoints.
Proper citation: Bio2RDF atlas of post genomic knowledge (RRID:SCR_007991) Copy
https://cran.r-project.org/web/packages/tdthap/index.html
Software package for TDT with extended haplotypes in the R language. R is the public domain dialect of S. It should be possible to port this library to the commercial Splus product. The main problem would be translation of the help files. (entry from Genetic Analysis Software)
Proper citation: R/TDTHAP (RRID:SCR_007625) Copy
http://locus.jouy.inra.fr/cgi-bin/lgbc/mapping/common/intro2.pl?BASE=goat
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 16, 2013. This website contains information about the mapping of the caprine genome. It contains loci list, phenes list, cartography, gene list, and other sequence information about goats. This website contains 731 loci, 271 genes, and 1909 homologue loci on 112 species. It also allows users to summit their own data for Goatmap. ARK-Genomics is not-for-profit and has collaborators from all over the world with an interest in farm animal genomics and genetics. ARK-Genomics was initially set up in 2000 with a grant awarded from the BBSRC IGF (Investigating Gene Function) initiative and from core resources of the Roslin Institute to provide a laboratory for automated analysis of gene expression using state-of-the-art genomic facilities. Since then, ARK-Genomics has expanded considerably, building up considerable expertise and resources.
Proper citation: GoatMap Database (RRID:SCR_008144) Copy
http://www.sanger.ac.uk/Projects/Microbes/
This website includes a list of projects that the Sanger Institute is currently working on or completed. All projects consist of the genomic sequencing of different bacteria. Each description of the bacteria includes its classification, a description, and the types of diseases that the bacteria is likely to cause. The Sanger Institute bacterial sequencing effort is concentrated on pathogens and model organisms. Data is accessible in a number of ways; for each organism there is a BLAST server, allowing users to search the sequences with their own query and retrieve the matching contigs. Sequences can also be downloaded directly by FTP. Data is accessible in a number of ways; for each organism there is a BLAST server, allowing you to search the sequences with your own query and retrieve the matching contigs. Sequences can also be downloaded directly by FTP. The primary sequence viewer and annotation tool, Artemis is available for download. This is a portable Java program which is used extensively within the Microbial Genomes group for the analysis and annotation of sequence data from cosmids to whole genomes. The Artemis Comparison Tool (ACT) is also useful for interactive viewing of the comparisons between large and small sequences.
Proper citation: Bacterial Genomes (RRID:SCR_008141) Copy
http://www.animalgenome.org/pigs/nagrp.html
Database and resources on the pig genome.
Proper citation: U.S. Pig Genome Project (RRID:SCR_008151) Copy
http://genewindow.nci.nih.gov/
Software tool for pre- and post-genetic bioinformatics and analytical work, developed and used at the Core Genotyping Facility (CGF) at the National Cancer Institute. While Genewindow is implemented for the human genome and integrated with the CGF laboratory data, it stands as a useful tool to assist investigators in the selection of variants for study in vitro, or in novel genetic association studies. The Genewindow application and source code is publicly available for use in other genomes, and can be integrated with the analysis, storage, and archiving of data generated in any laboratory setting. This can assist laboratories in the choice and tracking of information related to genetic annotations, including variations and genomic positions. Features of GeneWindow include: -Intuitive representation of genomic variation using advanced web-based graphics (SVG) -Search by HUGO gene symbol, dbSNP ID, internal CGF polymorphism ID, or chromosome coordinates -Gene-centric display (only when a gene of interest is in view) oriented 5 to 3 regardless of the reference strand and adjacent genes -Two views, a Locus Overview, which varies in size depending on the gene or genomic region being viewed and, below it, a Sequence View displaying 2000 base pairs within the overview -Navigate the genome by clicking along the gene in the Locus Overview to change the Sequence View, expand or contract the genomic interval, or shift the view in the 5 or 3 direction (relative to the current gene) -Lists of available genomic features -Search for sequence matches in the Locus Overview -Genomic features are represented by shape, color and opacity with contextual information visible when the user moves over or clicks on a feature -Administrators can insert newly-discovered polymorphisms into the Genewindow database by entering annotations directly through the GUI -Integration with a Laboratory Information Management System (LIMS) or other databases is possible
Proper citation: GeneWindow (RRID:SCR_008183) Copy
It facilitates the search for and dissemination of mass spectra from biologically active metabolites quantified using Gas chromatography (GC) coupled to mass spectrometry (MS). Use the Search Page to search for a compound of your interest, using the name, mass, formula, InChI etc. as query input. Additionally, a Library Search service enables the search of user submitted mass spectra within the GMD. In parallel to the library search, a prediction of chemical sub-groups is performed. This approach has reached beta level and a publication is currently under review. Using several sub-group specific Decision Trees (DTs), mass spectra are classified with respect to the presence of the chemical moieties within the linked (unknown) compound. Prediction of functional groups (ms analysis) facilitates the search of metabolites within the GMD by means of user submitted GC-MS spectra consisting of retention index (n-alkanes, if vailable) and mass intensities ratios. In addition, a functional group prediction will help to characterize those metabolites without available reference mass spectra included in the GMD so far. Instead, the unknown metabolite is characterized by predicted presence or absence of functional groups. For power users this functionality presented here is exposed as soap based web services. Functional group prediction of compounds by means of GC-EI-MS spectra using Microsoft analysis service decision trees All currently available trained decision trees and sub-structure predictions provided by the GMD interface. Table describes the functional group, optional use of an RI system, record date of the trained decision tree, number of MSTs with proportion of MSTs linked to metabolites with the functional group present for each tree. Average and standard deviation of the 50-fold CV error, namely the ratio false over correctly sorted MSTs in the trained DT, are listed. The GMD website offers a range of mass spectral reference libraries to academic users which can be downloaded free of charge in various electronic formats. The libraries are constituted by base peak normalized consensus spectra of single analytes and contain masses in the range 70 to 600 amu, while the ubiquitous mass fragments typically generated from compounds carrying a trimethylsilyl-moiety, namely the fragments at m/z 73, 74, 75, 147, 148, and 149, were excluded.
Proper citation: GMD (RRID:SCR_006625) Copy
https://www.fludb.org/brc/home.spg?decorator=influenza
The Influenza Research Database (IRD) serves as a public repository and analysis platform for flu sequence, experiment, surveillance and related data.
Proper citation: Influenza Research Database (IRD) (RRID:SCR_006641) Copy
DPVweb provides a central source of information about viruses, viroids and satellites of plants, fungi and protozoa. Comprehensive taxonomic information, including brief descriptions of each family and genus, and classified lists of virus sequences are provided. The database also holds detailed, curated, information for all sequences of viruses, viroids and satellites of plants, fungi and protozoa that are complete or that contain at least one complete gene. For comparative purposes, it also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA genome. The start and end positions of each feature (gene, non-translated region and the like) have been recorded and checked for accuracy. As far as possible, nomenclature for genes and proteins are standardized within genera and families. Sequences of features (either as DNA or amino acid sequences) can be directly downloaded from the website in FASTA format. The sequence information can also be accessed via client software for PC computers (freely downloadable from the website) that enable users to make an easy selection of sequences and features of a chosen virus for further analyses. The public sequence databases contain vast amounts of data on virus genomes but accessing and comparing the data, except for relatively small sets of related viruses can be very time consuming. The procedure is made difficult because some of the sequences on these databases are incorrectly named, poorly annotated or redundant. The NCBI Reference Sequence project (1) provides a comprehensive, integrated, non-redundant set of sequences, including genomic DNA, transcript (RNA) and protein products, for major research organisms. This now includes curated information for a single sequence of each fully sequenced virus species. While this is a welcome development, it can only deal with complete sequences. An important feature of DPV is the opportunity to access genes (and other features) of multiple sequences quickly and accurately. Thus, for example, it is easy to obtain the nucleotide or amino acid sequences of all the available accessions of the coat protein gene of a given virus species or for a group of viruses. To increase its usefulness further, DPVweb also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA (ssDNA) genome. Sponsors: This site is supported by the Association of Applied Biologists and the Zhejiang Academy of Agricultural Sciences, Hangzhou, People''s Republic of China.
Proper citation: Descriptions of Plant Viruses (RRID:SCR_006656) Copy
http://inparanoid.sbc.su.se/cgi-bin/index.cgi
Collection of pairwise comparisons between 100 whole genomes generated by a fully automatic method for finding orthologs and in-paralogs between TWO species. Ortholog clusters in the InParanoid are seeded with a two-way best pairwise match, after which an algorithm for adding in-paralogs is applied. The method bypasses multiple alignments and phylogenetic trees, which can be slow and error-prone steps in classical ortholog detection. Still, it robustly detects complex orthologous relationships and assigns confidence values for in-paralogs. The original data sets can be downloaded.
Proper citation: InParanoid: Eukaryotic Ortholog Groups (RRID:SCR_006801) Copy
http://www.nervenet.org/main/dictionary.html
A mouse-related portal of genomic databases and tables of mouse brain data. Most files are intended for you to download and use on your own personal computer. Most files are available in generic text format or as FileMaker Pro databases. The server provides data extracted and compiled from: The 2000-2001 Mouse Chromosome Committee Reports, Release 15 of the MIT microsatellite map (Oct 1997), The recombinant inbred strain database of R.W. Elliott (1997) and R. W. Williams (2001), and the Map Manager and text format chromosome maps (Apr 2001). * LXS genotype (Excel file): Updated, revised positions for 330 markers genotyped using a panel of 77 LXS strain. * MIT SNP DATABASE ONLINE: Search and sort the MIT Single Nucleotide Polymorphism (SNP) database ONLINE. These data from the MIT-Whitehead SNP release of December 1999. * INTEGRATED MIT-ROCHE SNP DATABASE in EXCEL and TEXT FORMATS (1-3 MB): Original MIT SNPs merged with the new Roche SNPs. The Excel file has been formatted to illustrate SNP haplotypes and genetic contrasts. Both files are intended for statistical analyses of SNPs and can be used to test a method outlined in a paper by Andrew Grupe, Gary Peltz, and colleagues (Science 291: 1915-1918, 2001). The Excel file includes many useful equations and formatting that will help in navigating through this large database and in testing the in silico mapping method. * Use of inbred strains for the study of individual differences in pain related phenotypes in the mouse: Elissa J. Chesler''s 2002 dissertation, discussing issues relevant to the integration of genomic and phenomic data from standard inbred strains including genetic interactions with laboratory environmental conditions and the use of various in silico inbred strain haplotype based mapping algorithms for QTL analysis. * SNP QTL MAPPER in EXCEL format (572 KB, updated January 2002 by Elissa Chesler): This Excel workbook implements the Grupe et al. mapping method and outputs correlation plots. The main spreadsheet allows you to enter your own strain data and compares them to haplotypes. Be very cautious and skeptical when using this spreadsheet and the technique. Read all of the caveates. This excel version of the method was developed by Elissa Chesler. This updated version (Jan 2002) handles missing data. * MIT SNP Database (tab-delimited text format): This file is suitable for manipulation in statistics and spreadsheet programs (752 KB, Updated June 27, 2001). Data have been formatted in a way that allows rapid acquisition of the new data from the Roche Bioscience SNP database. * MIT SNP Database (FileMaker 5 Version): This is a reformatted version of the MIT Single Nucleotide Polymorphism (SNP) database in FileMaker 5 format. You will need a copy of this application to open the file (Mac and Windows; 992 KB. Updated July 13, 2001 by RW). * Gene Mapping and Map Manager Data Sets: Genetic maps of mouse chromosomes. Now includes a 10th generation advanced intercross consisting of 500 animals genetoyped at 340 markers. Lots of older files on recombinant inbred strains. * The Portable Dictionary of the Mouse Genome, 21,039 loci, 17,912,832 bytes. Includes all 1997-98 Chromosome Committee Reports and MIT Release 15. * FullDict.FMP.sit: The Portable Dictionary of the Mouse Genome. This large FileMaker Pro 3.0/4.0 database has been compressed with StuffIt. The Dictionary of the Mouse Genome contains data from the 1997-98 chromosome committee reports and MIT Whitehead SSLP databases (Release 15). The Dictionary contains information for 21,039 loci. File size = 4846 KB. Updated March 19, 1998. * MIT Microsatellite Database ONLINE: A database of MIT microsatellite loci in the mouse. Use this FileMaker Pro database with OurPrimersDB. MITDB is a subset of the Portable Dictionary of the Mouse Genome. ONLINE. Updated July 12, 2001. * MIT Microsatellite Database: A database of MIT microsatellite loci in the mouse. Use this FileMaker Pro database with OurPrimersDB. MITDB is a subset of the Portable Dictionary of the Mouse Genome. File size = 3.0 MB. Updated March 19, 1998. * OurPrimersDB: A small database of primers. Download this database if you are using numerous MIT primers to map genes in mice. This database should be used in combination with the MITDB as one part of a relational database. File size = 149 KB. Updated March 19, 1998. * Empty copy (clone) of the Portable Dictionary in FileMaker Pro 3.0 format. Download this file and import individual chromosome text files from the table into the database. File size = 231 KB. Updated March 19, 1998. * Chromosome Text Files from the Dictionary: The table lists data on gene loci for individual chromosomes.
Proper citation: Mouse Genome Databases (RRID:SCR_007147) Copy
http://www.genoscope.cns.fr/externe/tetraodon/
The initial objective of Genoscope was to compare the genomic sequences of this fish to that of humans to help in the annotation of human genes and to estimate their number. This strategy is based on the common genetic heritage of the vertebrates: from one species of vertebrate to another, even for those as far apart as a fish and a mammal, the same genes are present for the most part. In the case of the compact genome of Tetraodon, this common complement of genes is contained in a genome eight times smaller than that of humans. Although the length of the exons is similar in these two species, the size of the introns and the intergenic sequences is greatly reduced in this fish. Furthermore, these regions, in contrast to the exons, have diverged completely since the separation of the lineages leading to humans and Tetraodon. The Exofish method, developed at Genoscope, exploits this contrast such that the conserved regions which can be identified by comparing genomic sequences of the two species, correspond only to coding regions. Using preliminary sequencing results of the genome of Tetraodon in the year 2000, Genoscope evaluated the number of human genes at about 30,000, whereas much higher estimations were current. The progress of the annotation of the human genome has since supported the Genoscope hypothesis, with values as low as 22,000 genes and a consensus of around 25,000 genes. The sequencing of the Tetraodon genome at a depth of about 8X, carried out as a collaboration between Genoscope and the Whitehead Institute Center for Genome Research (now the Broad Institute), was finished in 2002, with the production of an assembly covering 90 of the euchromatic region of the genome of the fish. This has permitted the application of Exofish at a larger scale in comparisons with the genome of humans, but also with those of the two other vertebrates sequenced at the time (Takifugu, a fish closely related to Tetraodon, and the mouse). The conserved regions detected in this way have been integrated into the annotation procedure, along with other resources (cDNA sequences from Tetraodon and ab initio predictions). Of the 28,000 genes annotated, some families were examined in detail: selenoproteins, and Type 1 cytokines and their receptors. The comparison of the proteome of Tetraodon with those of mammals has revealed some interesting differences, such as a major diversification of some hormone systems and of the collagen molecules in the fish. A search for transposable elements in the genomic sequences of Tetraodon has also revealed a high diversity (75 types), which contrasts with their scarcity; the small size of the Tetraodon genome is due to the low abundance of these elements, of which some appear to still be active. Another factor in the compactness of the Tetraodon genome, which has been confirmed by annotation, is the reduction in intron size, which approaches a lower limit of 50-60 bp, and which preferentially affects certain genes. The availability of the sequences from the genomes of humans and mice on one hand, and Takifugu and Tetraodon on the other, provide new opportunities for the study of vertebrate evolution. We have shown that the level of neutral evolution is higher in fish than in mammals. The protein sequences of fish also diverge more quickly than those of mammals. A key mechanism in evolution is gene duplication, which we have studied by taking advantage of the anchoring of the majority of the sequences from the assembly on the chromosomes. The result of this study speaks strongly in favor of a whole genome duplication event, very early in the line of ray-finned fish (Actinopterygians). An even stronger evidence came from synteny studies between the genomes of humans and Tetraodon. Using a high-resolution synteny map, we have reconstituted the genome of the vertebrate which predates this duplication - that is, the last common ancestor to all bony vertebrates (most of the vertebrates apart from cartilaginous fish and agnaths like lamprey). This ancestral karyotype contains 12 chromosomes, and the 21 Tetraodon chromosomes derive from it by the whole genome duplication and a surprisingly small number of interchromosomal rearrangements. On the contrary, exchanges between chromosomes have been much more frequent in the lineage that leads to humans. Sponsors: The project was supported by the Consortium National de Recherche en Genomique and the National Human Genome Research Institute.
Proper citation: Tetraodon Genome Browser (RRID:SCR_007079) Copy
http://www.broadinstitute.org/
Biomedical and genomic research center located in Cambridge, Massachusetts, United States. Nonprofit research organization under the name Broad Institute Inc., and is partners with Massachusetts Institute of Technology, Harvard University, and the five Harvard teaching hospitals. Dedicated to advance understanding of biology and treatment of human disease to improve human health.
Proper citation: Broad Institute (RRID:SCR_007073) Copy
http://sharedresources.fredhutch.org/core-facilities/bioinformatics
THIS RESOURCE IS NO LONGER IN SERVICE.Documented on July 27,2022. Core provides bioinformatics specialists available to assist researchers with processing, exploring, and understanding genomics data.
Proper citation: Fred Hutchinson Cancer Research Center Co-operative Center for Excellence in Hematology Bioinformatics Resource (RRID:SCR_015324) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.