Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
MitoRes, is a comprehensive and reliable resource for massive extraction of sequences and sub-sequences of nuclear genes and encoded products targeting mitochondria in metazoa. It has been developed for supporting high-throughput in-silico analyses aimed to studies of functional genomics related to mitochondrial biogenesis, metabolism and to their pathological dysfunctions. It integrates information from the most accredited world-wide databases to bring together gene, transcript and encoded protein sequences associated to annotations on species name and taxonomic classification, gene name, functional product, organelle localization, protein tissue specificity, Enzyme Classification (EC), Gene Ontology (GO) classification and links to other related public databases. The section Cluster, has been dedicated to the collection of data on protein clustering of the entire catalogue of MitoRes protein sequences based on all versus all global pair-wise alignments for assessing putative intra- and inter-species functional relationships. The current version of MitoRes is based on the UniProt release 4 and contains 64 different metazoan species. The incredible explosion of knowledge production in Biology in the past two decades has created a critical need for bioinformatic instruments able to manage data and facilitate their retrieval and analysis. Hundreds of biological databases have been produced and the integration of biological data from these different resources is very important when we want to focus our efforts towards the study of a particular layer of biological knowledge. MitoRes is a completely rebuilt edition of MitoNuc database, which has been extensively modified to deal successfully with the challenges of the post genomic era. Its goal is to represent a comprehensive and reliable resource supporting high-quality in-silico analyses aimed to the functional characterization of gene, transcript and amino acid sequences, encoded by the nuclear genome and involved in mitochondrial biogenesis, metabolism and pathological dysfunctions in metazoa. The central features of MitoRes are: # an integrated catalogue of protein, transcript and gene sequences and sub-sequences # a Web-based application composed of a wide spectrum of search/retrieval facilities # a sequence export manager allowing massive extraction of bio-sequences (genes, introns, exons, gene flanking regions, transcripts, UTRs, CDS, proteins and signal peptides) in FASTA, EMBL and GenBank formats. It is an interconnected knowledge management system based on a MySQL relational database, which ensures data consistency and integrity, and on a Web Graphical User Interface (GUI), built in Seagull PHP Framework, offering a wide range of search and sequence extraction facilities. The database is compiled extracting and integrating information from public resources and data generated by the MitoRes team. The MitoRes database consists of comprehensive sequence entries whose core data are protein, transcript and gene sequences and taxonomic information describing the biological source of the protein. Additional information include: bio-sequences structure and location, biological function of protein product and dynamic links to both, external public databases used as data resources and public databases reporting complementary information. The core entity of the MitoRes database is represented by the protein so that each MitoRes entry is generated for each protein reported in the UniProt database as a nuclear encoded protein involved in mitochondrial biogenesis and function. Sponsors: MitoRes has been supported by Ministero Universit e Ricerca Scientifica, Italy (PRIN, Programma Biotecnologie legge 95/95-MURST 5, Proiect MURST Cluster C03/2000, CEGBA). Currently it is supported by operating grants from the Ministero dellIstruzione, dellUniversit e della Ricerca (MIUR), Italy (PNR 2001-2003 (FIRB art.8) D.M. 199, Strategic Program: Post-genome, grant 31-063933 and Project n.2, Cluster C03 L. 488/929).
Proper citation: MitoRes (RRID:SCR_008208) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 29, 2016. An algorithm that finds articles most relevant to a genetic sequence. In the genomic era, researchers often want to know more information about a biological sequence by retrieving its related articles. However, there is no available tool yet to achieve conveniently this goal. Here, a new literature-mining tool MedBlast is developed, which uses natural language processing techniques, to retrieve the related articles of a given sequence. An online server of this program is also provided. The genome sequencing projects generate such a large amount of data every day that many molecular biologists often encounter some sequences that they know nothing about. Literature is usually the principal resource of such information. It is relatively easy to mine the articles cited by the sequence annotation; however, it is a difficult task to retrieve those relevant articles without direct citation relationship. The related articles are those described in the given sequence (gene/protein), or its redundant sequences, or the close homologs in various species. They can be divided into two classes: direct references, which include those either cited by the sequence annotation or citing the sequence in its text; indirect references, those which contain gene symbols of the given sequence. A few additional issues make the task even more complicated: (1) symbols may have aliases; and (2) one sequence may have a couple of relatives that we want to take into account too, which include redundant (e.g. protein and gene sequences) and close homologs. Here the issues are addressed by the development of the software MedBlast, which can retrieve the related articles of the given sequence automatically. MedBlast uses BLAST to extend homology relationships, precompiled species-specific thesauruses, a useful semantics technique in natural language processing (NLP), to extend alias relationship, and EUtilities toolset to search and retrieve corresponding articles of each sequence from PubMed. MedBlast take a sequence in FASTA format as input. The program first uses BLAST to search the GenBank nucleic acid and protein non-redundant (nr) databases, to extend to those homologous and corresponding nucleic acid and protein sequences. Users can input the BLAST results directly, but it is recommended to input the result of both protein and nucleic acid nr databases. The hits with low e-values are chosen as the relatives because the low similarity hits often do not contain specific information. Very long sequences, e.g. 100k, which are usually genomic sequences, are discarded too, for they do not contain specific direct references. User can adjust these parameters to meet their own needs.
Proper citation: MedBlast (RRID:SCR_008202) Copy
http://www.ebi.ac.uk/parasites/parasite-genome.html
This website contains information about the genomic sequence of parasites. It also contains multiple search engines to search six frame translations of parasite nucleotide databases for motifs, parasite protein databases for motifs, and parasite protein databases for keywords and text terms. * Guide to Internet Access to Parasite Genome Information * Guide to web-based analysis tools * Parasite Genome BLAST Server: Search a range of parasite specific nucleotide sequence databases with your own sequence. * Parasite Proteome Keyword Search Facility: Search parasite protein databases for keywords and text terms * Parasite Proteome Motif Search Facility: Search parasite protein databases for motifs * Parasite Six Frame Translation Motif Search Facility: Search six frame translations of parasite nucleotide databases for motifs * Genome computing resources: A list of ftp and gopher sites where genome computing applications and other resources can be found.
Proper citation: Parasite genome databases and genome research resources (RRID:SCR_008150) Copy
http://www.hgsc.bcm.tmc.edu/content/bovine-genome-project
Downloadable files of the bos taurus genome. Draft assemblies available for download as contigs or linearized scaffolds of the genomic sequence of cow, Bos taurus, including the final draft assembly (7.1 coverage) and the two previous assemblies. The genome is sequenced to 6- to 8-fold sequence depth, with high-quality finished sequence in some areas. Accompanying EST and SNP analyses is also included. The bovine genome assembly and analysis and the study of cattle genetic history were published in April 24, 2009 issue of Science. The Human Genome Sequencing Center provides BLAST searches of the genome assemblies, either as contigs or as linearized chromosome sequences. The WGS sequence enriched BAC assemblies and the unassembled reads (sequencing reads that did not end up in the genome assembly) can also be searched by BLAST. Traces are available from the NCBI Trace Archive by using the link in the sidebar or by using NCBI MegaBLAST with a same species or cross species query.
Proper citation: Bovine Genome Project (RRID:SCR_008370) Copy
Non profit, private research and education institution that performs molecular and genetic research used to generate methods for better diagnostics and treatments for cancer and neurological diseases. Research of cancer causing genes and their respective signaling pathways, mutations and structural variations of the human genome that could cause neurodevelopmental and neurodegenerative illnesses such as autism, schizophrenia, and Alzheimer's and Parkinson's diseases and also research in plant genetics and quantitative biology.
Proper citation: Cold Spring Harbor Laboratory (RRID:SCR_008326) Copy
http://www.broad.mit.edu/mammals/dog
The genome of the domesticated dog, a close evolutionary relation to human, is a powerful new tool for understanding the human genome. Comparison of the dog with human and other mammals reveals key information about the structure and evolution of genes and genomes. The unique breeding history of dogs, with their extraordinary behavioral and physical diversity, offers the opportunity to find important genes underlying diseases shared between dogs and humans, such as cancer, diabetes, and epilepsy. The Canine Genome Sequencing Project produced a high-quality draft sequence of a female boxer named Tasha. By comparing Tasha with many other breeds, the project also compiled a comprehensive set of SNPs (single nucleotide polymorphisms) useful in all dog breeds. These closely spaced genomic landmarks are critical for disease mapping. By comparing the dog, rodent, and human lineages, researchers at the Broad Institute uncovered exciting new information about human genes, their evolution, and the regulatory mechanisms governing their expression. Using SNPs, researchers describe the strikingly different haplotype structure in dog breeds compared with the entire dog population. In addition, they show that by understanding the patterns of variation in dog breeds, scientists can design powerful gene mapping experiments for complex diseases that are difficult to map in human populations. Contribute Although the astounding generosity of Eli and Edythe L. Broad and several other venture philanthropists empowers our scientists to tackle many of the most important problems at the cutting edge of genomic medicine, there are many other critical challenges that they cannot yet pursue because of limited resources. We need additional visionary partners to join the Broads and the Broad Institute in transforming medicine with the power of genomics.
Proper citation: Dog Genome Project (RRID:SCR_008486) Copy
http://rgd.mcw.edu/rgdCuration/?module=portal&func=show&name=nuro
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on May 12,2023. Portal that provides researchers with easy access to data on rat genes, QTLs, strain models, biological processes and pathways related to neurological diseases. This resource also includes dynamic data analysis tools.
Proper citation: Rat Genome Database: Neurological Disease Portal (RRID:SCR_008685) Copy
A robust, secure, medical-grade, web application that lives in the cloud and has the ability to analyze and annotate entire human genomes in a rapid and cost-effective way.
Proper citation: Tute Genomics (RRID:SCR_008672) Copy
http://www.uni-koeln.de/med-fak/cgars/
Software package to dissect random from non-random patterns in copy number data and thereby to assess significantly enriched somatic copy number aberrations (SCNA) across a set of tumor specimens or cell lines.
Proper citation: CGARS (RRID:SCR_006404) Copy
http://www.clipz.unibas.ch/downloads/TSSer/index.php
A computational pipeline to analyze differential RNA sequencing (dRNA-seq) data to determine transcription start sites genome-wide.
Proper citation: TSSer (RRID:SCR_006419) Copy
Model organism database that provides organization of and access to scientific data for the fission yeast Schizosaccharomyces pombe. PomBase supports genomic sequence and features, genome-wide datasets and manual literature curation. PomBase also provides a community hub for researchers, providing genome statistics, a community curation interface, news, events, documentation, mailing lists, and welcomes data submissions.
Proper citation: PomBase (RRID:SCR_006586) Copy
http://bioinformatics.ubc.ca/ermineJ/
Data analysis software for gene sets in expression microarray data or other genome-wide data that results in rankings of genes. A typical goal is to determine whether particular biological pathways are doing something interesting in the data. The software is designed to be used by biologists with little or no informatics background. A command-line interface is available for users who wish to script the use of ermineJ. Major features include: * Implementation of multiple methods for gene set analysis: ** Over-representation analysis ** A resampling-based method that uses gene scores ** A rank-based method that uses gene scores ** A resampling-based method that uses correlation between gene expression profiles (a type of cluster-enrichment analysis). * Gene sets receive statistical scores (p-values), and multiple test correction is supported. * Support of the Gene Ontology terminology; users can choose which aspects to analyze. * User files use simple text formats. * Users can modify gene sets or create new ones. * The results can be visualized within the software. * It is simple to compare multiple analyses of the same data set with different settings. * User-definable hyperlinks are provided to external sites to allow more efficient browsing of the results. * For programmers, there is a command line interface as well as a simple application programming interface that can be used to plug ermineJ functionality into your own code Platform: Online tool, Windows compatible, Mac OS X compatible, Linux compatible, Unix compatible
Proper citation: ErmineJ (RRID:SCR_006450) Copy
Public archive providing a comprehensive record of the world''''s nucleotide sequencing information, covering raw sequencing data, sequence assembly information and functional annotation. All submitted data, once public, will be exchanged with the NCBI and DDBJ as part of the INSDC data exchange agreement. The European Nucleotide Archive (ENA) captures and presents information relating to experimental workflows that are based around nucleotide sequencing. A typical workflow includes the isolation and preparation of material for sequencing, a run of a sequencing machine in which sequencing data are produced and a subsequent bioinformatic analysis pipeline. ENA records this information in a data model that covers input information (sample, experimental setup, machine configuration), output machine data (sequence traces, reads and quality scores) and interpreted information (assembly, mapping, functional annotation). Data arrive at ENA from a variety of sources including submissions of raw data, assembled sequences and annotation from small-scale sequencing efforts, data provision from the major European sequencing centers and routine and comprehensive exchange with their partners in the International Nucleotide Sequence Database Collaboration (INSDC). Provision of nucleotide sequence data to ENA or its INSDC partners has become a central and mandatory step in the dissemination of research findings to the scientific community. ENA works with publishers of scientific literature and funding bodies to ensure compliance with these principles and to provide optimal submission systems and data access tools that work seamlessly with the published literature. ENA is made up of a number of distinct databases that includes the EMBL Nucleotide Sequence Database (Embl-Bank), the newly established Sequence Read Archive (SRA) and the Trace Archive. The main tool for downloading ENA data is the ENA Browser, which is available through REST URLs for easy programmatic use. All ENA data are available through the ENA Browser. Note: EMBL Nucleotide Sequence Database (EMBL-Bank) is entirely included within this resource.
Proper citation: European Nucleotide Archive (ENA) (RRID:SCR_006515) Copy
Set of measures intended for use in large-scale genomic studies. Facilitate replication and validation across studies. Includes links to standards and resources in effort to facilitate data harmonization to legacy data. Measurement protocols that address wide range of research domains. Information about each protocol to ensure consistent data collection.Collections of protocols that add depth to Toolkit in specific areas.Tools to help investigators implement measurement protocols.
Proper citation: Phenotypes and eXposures Toolkit (RRID:SCR_006532) Copy
Database for genetic, genomic, phenotype, and disease data generated from rat research. Centralized database that collects, manages, and distributes data generated from rat genetic and genomic research and makes these data available to scientific community. Curation of mapped positions for quantitative trait loci, known mutations and other phenotypic data is provided. Facilitates investigators research efforts by providing tools to search, mine, and analyze this data. Strain reports include description of strain origin, disease, phenotype, genetics, immunology, behavior with links to related genes, QTLs, sub-strains, and strain sources.
Proper citation: Rat Genome Database (RGD) (RRID:SCR_006444) Copy
http://www.gigasciencejournal.com/
An online open-access open-data journal, publishing ''big-data'' studies from the entire spectrum of life and biomedical sciences whose publication format links standard manuscript publication with its affiliated database, GigaDB, that hosts all associated data, provides data analysis tools, cloud-computing resources, and a DOI assignment to every dataset. GigaScience covers not just ''omic'' type data and the fields of high-throughput biology currently serviced by large public repositories, but also the growing range of more difficult-to-access data, such as imaging, neuroscience, ecology, cohort data, systems biology and other new types of large-scale sharable data. Supporting the open-data movement, they require that all supporting data and source code be publicly available in a suitable public repository and/or under a public domain CC0 license in the BGI GigaScience database. Using the BGI cloud as a test environment, they also consider open-source software tools / methods for the analysis or handling of large-scale data. When submitting a manuscript, please contact them if you have datasets or cloud applications you would like them to host. To maximize data usability submitters are encouraged to follow best practice for metadata reporting and are given the opportunity to submit in ISA-Tab format.
Proper citation: GigaScience (RRID:SCR_006565) Copy
Database of Drosophila genetic and genomic information with information about stock collections and fly genetic tools. Gene Ontology (GO) terms are used to describe three attributes of wild-type gene products: their molecular function, the biological processes in which they play a role, and their subcellular location. Additionally, FlyBase accepts data submissions. FlyBase can be searched for genes, alleles, aberrations and other genetic objects, phenotypes, sequences, stocks, images and movies, controlled terms, and Drosophila researchers using the tools available from the "Tools" drop-down menu in the Navigation bar.
Proper citation: FlyBase (RRID:SCR_006549) Copy
Service providing functional analysis of proteins by classifying them into families and predicting domains and important sites. They combine protein signatures from a number of member databases into a single searchable resource, capitalizing on their individual strengths to produce a powerful integrated database and diagnostic tool. This integrated database of predictive protein signatures is used for the classification and automatic annotation of proteins and genomes. InterPro classifies sequences at superfamily, family and subfamily levels, predicting the occurrence of functional domains, repeats and important sites. InterPro adds in-depth annotation, including GO terms, to the protein signatures. You can access the data programmatically, via Web Services. The member databases use a number of approaches: # ProDom: provider of sequence-clusters built from UniProtKB using PSI-BLAST. # PROSITE patterns: provider of simple regular expressions. # PROSITE and HAMAP profiles: provide sequence matrices. # PRINTS provider of fingerprints, which are groups of aligned, un-weighted Position Specific Sequence Matrices (PSSMs). # PANTHER, PIRSF, Pfam, SMART, TIGRFAMs, Gene3D and SUPERFAMILY: are providers of hidden Markov models (HMMs). Your contributions are welcome. You are encouraged to use the ''''Add your annotation'''' button on InterPro entry pages to suggest updated or improved annotation for individual InterPro entries.
Proper citation: InterPro (RRID:SCR_006695) Copy
Database of peer-reviewed, continually updated annotation for the Pseudomonas aeruginosa PAO1 reference strain genome expanded to include all Pseudomonas species to facilitate cross-strain and cross-species genome comparisons with high quality comparative genomics. The database contains robust assessment of orthologs, a novel ortholog clustering method, and incorporates five views of the data at the sequence and annotation levels (Gbrowse, Mauve and custom views) to facilitate genome comparisons. Other features include more accurate protein subcellular localization predictions and a user-friendly, Boolean searchable log file of updates for the reference strain PAO1. The current annotation is updated using recent research literature and peer-reviewed submissions by a worldwide community of PseudoCAP (Pseudomonas aeruginosa Community Annotation Project) participating researchers. If you are interested in participating, you are invited to get involved. Many annotations, DNA sequences, Orthologs, Intergenic DNA, and Protein sequences are available for download.
Proper citation: Pseudomonas Genome Database (RRID:SCR_006590) Copy
Collection of data related to crop plant and model organism Zea mays. Used to synthesize, display, and provide access to maize genomics and genetics data, prioritizing mutant and phenotype data and tools, structural and genetic map sets, and gene models and to provide support services to the community of maize researchers. Data stored at MaizeGDB was inherited from the MaizeDB and ZmDB projects. Sequence data are from GenBank. Data are searchable by phenotype, traits, Pests, Gel Pattern, and Mutant Images.
Proper citation: MaizeGDB (RRID:SCR_006600) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.