Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A collaborative ontology for the definition of sequence features used in biological sequence annotation. SO was initially developed by the Gene Ontology Consortium. Contributors to SO include the GMOD community, model organism database groups such as WormBase, FlyBase, Mouse Genome Informatics group, and institutes such as the Sanger Institute and the EBI. Input to SO is welcomed from the sequence annotation community. The OBO revision is available here: http://sourceforge.net/p/song/svn/HEAD/tree/ SO includes different kinds of features which can be located on the sequence. Biological features are those which are defined by their disposition to be involved in a biological process. Biomaterial features are those which are intended for use in an experiment such as aptamer and PCR_product. There are also experimental features which are the result of an experiment. SO also provides a rich set of attributes to describe these features such as polycistronic and maternally imprinted. The Sequence Ontologies use the OBO flat file format specification version 1.2, developed by the Gene Ontology Consortium. The ontology is also available in OWL from Open Biomedical Ontologies. This is updated nightly and may be slightly out of sync with the current obo file. An OWL version of the ontology is also available. The resolvable URI for the current version of SO is http://purl.obolibrary.org/obo/so.owl.
Proper citation: SO (RRID:SCR_004374) Copy
http://www.alliancegenome.org/
Organization that aims to develop and maintain sustainable genome information resources to promote understanding of the genetic and genomic basis of human biology, health, and disease. The Alliance is composed of FlyBase, Mouse Genome Database (MGD), the Gene Ontology Consortium (GOC), Saccharomyces Genome Database (SGD), Rat Genome Database (RGD), WormBase, and the Zebrafish Information Network (ZFIN).
Proper citation: Alliance of Genome Resources (RRID:SCR_015850) Copy
http://gmod.org/wiki/Flash_GViewer
Flash GViewer is a customizable Flash movie that can be easily inserted into a web page to display each chromosome in a genome along with the locations of individual features on the chromosomes. It is intended to provide an overview of the genomic locations of a specific set of features - eg. genes and QTLs associated with a specific phenotype, etc. rather than as a way to view all features on the genome. The features can hyperlink out to a detail page to enable to GViewer to be used as a navigation tool. In addition the bands on the chromosomes can link to defineable URL and new region selection sliders can be used to select a specific chromosome region and then link out to a genome browser for higher resolution information. Genome maps for Rat, Mouse, Human and C. elegans are provided but other genome maps can be easily created. Annotation data can be provided as static text files or produced as XML via server scripts. This tool is not GO-specific, but was built for the purpose of viewing GO annotation data. Platform: Online tool
Proper citation: Flash Gviewer (RRID:SCR_012870) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 14,2026. Integrated database of genomic, expression and protein data for Drosophila, Anopheles, C. elegans and other organisms. You can run flexible queries, export results and analyze lists of data. FlyMine presents data in categories, with each providing information on a particular type of data (for example Gene Expression or Protein Interactions). Template queries, as well as the QueryBuilder itself, allow you to perform searches that span data from more than one category. Advanced users can use a flexible query interface to construct their own data mining queries across the multiple integrated data sources, to modify existing template queries or to create your own template queries. Access our FlyMine data via our Application Programming Interface (API). We provide client libraries in the following languages: Perl, Python, Ruby and & Java API
Proper citation: FlyMine (RRID:SCR_002694) Copy
http://www.informatics.jax.org
International database for laboratory mouse. Data offered by The Jackson Laboratory includes information on integrated genetic, genomic, and biological data. MGI creates and maintains integrated representation of mouse genetic, genomic, expression, and phenotype data and develops reference data set and consensus data views, synthesizes comparative genomic data between mouse and other mammals, maintains set of links and collaborations with other bioinformatics resources, develops and supports analysis and data submission tools, and provides technical support for database users. Projects contributing to this resource are: Mouse Genome Database (MGD) Project, Gene Expression Database (GXD) Project, Mouse Tumor Biology (MTB) Database Project, Gene Ontology (GO) Project at MGI, and MouseCyc Project at MGI.
Proper citation: Mouse Genome Informatics (MGI) (RRID:SCR_006460) Copy
http://smd.stanford.edu/cgi-bin/source/sourceSearch
SOURCE compiles information from several publicly accessible databases, including UniGene, dbEST, UniProt Knowledgebase, GeneMap99, RHdb, GeneCards and LocusLink. GO terms associated with LocusLink entries appear in SOURCE. The mission of SOURCE is to provide a unique scientific resource that pools publicly available data commonly sought after for any clone, GenBank accession number, or gene. SOURCE is specifically designed to facilitate the analysis of large sets of data that biologists can now produce using genome-scale experimental approaches Platform: Online tool
Proper citation: SOURCE (RRID:SCR_005799) Copy
http://plantgrn.noble.org/LegumeIP/
LegumeIP is an integrative database and bioinformatics platform for comparative genomics and transcriptomics to facilitate the study of gene function and genome evolution in legumes, and ultimately to generate molecular based breeding tools to improve quality of crop legumes. LegumeIP currently hosts large-scale genomics and transcriptomics data, including: * Genomic sequences of three model legumes, i.e. Medicago truncatula, Glycine max (soybean) and Lotus japonicus, including two reference plant species, Arabidopsis thaliana and Poplar trichocarpa, with the annotation based on UniProt TrEMBL, InterProScan, Gene Ontology and KEGG databases. LegumeIP covers a total 222,217 protein-coding gene sequences. * Large-scale gene expression data compiled from 104 array hybridizations from L. japonicas, 156 array hybridizations from M. truncatula gene atlas database, and 14 RNA-Seq-based gene expression profiles from G. max on different tissues including four common tissues: Nodule, Flower, Root and Leaf. * Systematic synteny analysis among M. truncatula, G. max, L. japonicus and A. thaliana. * Reconstruction of gene family and gene family-wide phylogenetic analysis across the five hosted species. LegumeIP features comprehensive search and visualization tools to enable the flexible query on gene annotation, gene family, synteny, relative abundance of gene expression.
Proper citation: LegumeIP (RRID:SCR_008906) Copy
http://rgd.mcw.edu/rgdCuration/?module=portal&func=show&name=renal
An integrated resource for information on genes, QTLs and strains associated with a variety of kidney and renal system conditions such as Renal Hypertension, Polycystic Kidney Disease and Renal Insufficiency, as well as Kidney Neoplasms.
Proper citation: Renal Disease Portal (RRID:SCR_009030) Copy
http://www.ebi.ac.uk/Tools/pfa/iprscan/
Software package for functional analysis of sequences by classifying them into families and predicting presence of domains and sites. Scans sequences against InterPro's signatures. Characterizes nucleotide or protein function by matching it with models from several different databases. Used in large scale analysis of whole proteomes, genomes and metagenomes. Available as Web based version and standalone Perl version and SOAP Web Service.
Proper citation: InterProScan (RRID:SCR_005829) Copy
Suite of motif-based sequence analysis tools to discover motifs using MEME, DREME (DNA only) or GLAM2 on groups of related DNA or protein sequences; search sequence databases with motifs using MAST, FIMO, MCAST or GLAM2SCAN; compare a motif to all motifs in a database of motifs; associate motifs with Gene Ontology terms via their putative target genes, and analyze motif enrichment using SpaMo or CentriMo. Source code, binaries and a web server are freely available for noncommercial use.
Proper citation: MEME Suite - Motif-based sequence analysis tools (RRID:SCR_001783) Copy
Database for ESTs (Expressed Sequence Tags), consensus sequences, bacterial artificial chromosome (BAC) clones, BES (BAC End Sequences). They have generated 69,545 ESTs from 6 full-length cDNA libraries (Porcine Abdominal Fat, Porcine Fat Cell, Porcine Loin Muscle, Liver and Pituitary gland). They have also identified a total of 182 BAC contigs from chromosome 6. It is very valuable resources to study porcine quantitative trait loci (QTL) mapping and genome study. Users can explore genomic alignment of various data types, including expressed sequence tags (ESTs), consensus sequences, singletons, QTL, Marker, UniGene and BAC clones by several options. To estimate the genomic location of sequence dataset, their data aligned BES (BAC End Sequences) instead of genomic sequence because Pig Genome has low-coverage sequencing data. Sus scrofa Genome Database mainly provide comparative map of four species (pig, cattle, dog and mouse) in chromosome 6.
Proper citation: PiGenome (RRID:SCR_013394) Copy
http://genespeed.ccf.org/home/
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 16, 2013. Database and customized tools to study the PFAM protein domain content of the transcriptome for all expressed genes of Homo sapiens, Mus musculus, Drosophila melanogaster, and Caenorhabditis elegans tethered to both a genomics array repository database and a range of external information resources. GeneSpeed has merged information from several existing data sets including the Gene Ontology Consortium, InterPro, Pfam, Unigene, as well as micro-array datasets. GeneSpeed is a database of PFAM domain homology contained within Unigene. Because Unigene is a non-redundant dbEST database, this provides a wide encompassing overview of the domain content of the expressed transcriptome. We have structured the GeneSpeed Database to include a rich toolset allowing the investigator to study all domain homology, no matter how remote. As a result, homology cutoff score decisions are determined by the scientist, not by a computer algorithm. This quality is one of the novel defining features of the GeneSpeed database giving the user complete control of database content. In addition to a domain content toolset, GeneSpeed provides an assortment of links to external databases, a unique and manually curated Transcription Factor Classification list, as well as links to our newly evolving GeneSpeed BetaCell Database. GeneSpeed BetaCell is a micro-array depository combined with custom array analysis tools created with an emphasis around the meta analysis of developmental time series micro-array datasets and their significance in pancreatic beta cells.
Proper citation: GeneSpeed- A Database of Unigene Domain Organization (RRID:SCR_002779) Copy
A database designed for plant comparative and functional genomics based on complete genomes. It comprises complete proteome sequences from the major phylum of plant evolution. The clustering of these proteomes was performed to define a consistent and extensive set of homeomorphic plant families. Based on this, lists of gene families such as plant or species specific families and several tools are provided to facilitate comparative genomics within plant genomes. The analyses follow two main steps: gene family clustering and phylogenomic analysis of the generated families. Once a group of sequences (cluster) is validated, phylogenetic analyses are performed to predict homolog relationships such as orthologs and ultraparalogs.
Proper citation: GreenPhylDB (RRID:SCR_002834) Copy
http://www.compbio.dundee.ac.uk/gotcha/gotcha.php
GOtcha provides a prediction of a set of GO terms that can be associated with a given query sequence. Each term is scored independently and the scores calibrated against reference searches to give an accurate percentage likelihood of correctness. These results can be displayed graphically. Why is GOtcha different to what is already out there and why should you be using it? * GOtcha uses a method where it combines information from many search hits, up to and including E-values that are normally discarded. This gives much better sensitivity than other methods. * GOtcha provides a score for each individual term, not just the leaf term or branch. This allows the discrimination between confident assignments that one would find at a more general level and the more specific terms that one would have lower confidence in. * The scores GOtcha provides are calibrated to give a real estimate of correctness. This is expressed as a percentage, giving a result that non-experts are comfortable in interpreting. * GOtcha provides graphical output that gives an overview of the confidence in, or potential alternatives for, particular GO term assignments. The tool is currently web-based; contact David Martin for details of the standalone version. Platform: Online tool
Proper citation: GOtcha (RRID:SCR_005790) Copy
http://www.cdtdb.brain.riken.jp/CDT/Top.jsp
Transcriptomic information (spatiotemporal gene expression profile data) on the postnatal cerebellar development of mice (C57B/6J & ICR). It is a tool for mining cerebellar genes and gene expression, and provides a portal to relevant bioinformatics links. The mouse cerebellar circuit develops through a series of cellular and morphological events, including neuronal proliferation and migration, axonogenesis, dendritogenesis, and synaptogenesis, all within three weeks after birth, and each event is controlled by a specific gene group whose expression profile must be encoded in the genome. To elucidate the genetic basis of cerebellar circuit development, CDT-DB analyzes spatiotemporal gene expression by using in situ hybridization (ISH) for cellular resolution and by using fluorescence differential display and microarrays (GeneChip) for developmental time series resolution. The CDT-DB not only provides a cross-search function for large amounts of experimental data (ISH brain images, GeneChip graph, RT-PCR gel images), but also includes a portal function by which all registered genes have been provided with hyperlinks to websites of many relevant bioinformatics regarding gene ontology, genome, proteins, pathways, cell functions, and publications. Thus, the CDT-DB is a useful tool for mining potentially important genes based on characteristic expression profiles in particular cell types or during a particular time window in developing mouse brains.
Proper citation: Cerebellar Development Transcriptome Database (RRID:SCR_013096) Copy
http://manatee.sourceforge.net/
Manatee is a web-based gene evaluation and genome annotation tool; Manatee can store and view annotation for prokaryotic and eukaryotic genomes. The Manatee interface allows biologists to quickly identify genes and make high quality functional assignments, such as GO classifications, using search data, paralogous families, and annotation suggestions generated from automated analysis. Manatee can be downloaded and installed to run under the CGI area of a web server, such as Apache. Platform: Online tool, Linux compatible, Solaris
Proper citation: Manatee (RRID:SCR_005685) Copy
Curated, open-source, integrated data resource for comparative functional genomics in crops and model plant species to facilitate the study of cross-species comparisons using information generated from projects supported by public funds. It currently hosts annotated whole genomes in over two dozen plant species and partial assemblies for almost a dozen wild rice species in the Ensembl browser, genetic and physical maps with genes, ESTs and QTLs locations, genetic diversity data sets, structure-function analysis of proteins, plant pathways databases (BioCyc and Plant Reactome platforms), and descriptions of phenotypic traits and mutations. The web-based displays for phenotypes include the Genes and Quantitative Trait Loci (QTL) modules. Sequence based relationships are displayed in the Genomes module using the genome browser adapted from Ensembl, in the Maps module using the comparative map viewer (CMap) from GMOD, and in the Proteins module displays. BLAST is used to search for similar sequences. Literature supporting all the above data is organized in the Literature database. In addition, Gramene now hosts a variety of web services including a Distributed Annotation Server (DAS), BLAST and a public MySQL database. Twice a year, Gramene releases a major build of the database and makes interim releases to correct errors or to make important updates to software and/or data. Additionally you can access Gramene through an FTP site.
Proper citation: Gramene (RRID:SCR_002829) Copy
https://omictools.com/ecgene-tool
Database of functional annotation for alternatively spliced genes. It uses a gene-modeling algorithm that combines the genome-based expressed sequence tag (EST) clustering and graph-theoretic transcript assembly procedures. It contains genome, mRNA, and EST sequence data, as well as a genome browser application. Organisms included in the database are human, dog, chicken, fruit fly, mouse, rhesus, rat, worm, and zebrafish. Annotation is provided for the whole transcriptome, not just the alternatively spliced genes. Several viewers and applications are provided that are useful for the analysis of the transcript structure and gene expression. The summary viewer shows the gene summary and the essence of other annotation programs. The genome browser and the transcript viewer are available for comparing the gene structure of splice variants. Changes in the functional domains by alternative splicing can be seen at a glance in the transcript viewer. Two unique ways of analyzing gene expression is also provided. The SAGE tags deduced from the assembled transcripts are used to delineate quantitative expression patterns from SAGE libraries available publicly. The cDNA libraries of EST sequences in each cluster are used to infer qualitative expression patterns.
Proper citation: ECgene: Gene Modeling with Alternative Splicing (RRID:SCR_007634) Copy
http://david.abcc.ncifcrf.gov/content.jsp?file=/ease/ease1.htm&type=1
Windows(c) desktop software application, customizable and standalone, that facilitates the biological interpretation of gene lists derived from the results of microarray, proteomic, and SAGE experiments. Provides statistical methods for discovering enriched biological themes within gene lists, generates gene annotation tables, and enables automated linking to online analysis tools. Offers statistical models to deal with multi-test comparison problem. Platform: Windows compatible
Proper citation: EASE: the Expression Analysis Systematic Explorer (RRID:SCR_013361) Copy
BioPerl is a community effort to produce Perl code which is useful in biology. This toolkit of perl modules is useful in building bioinformatics solutions in Perl. It is built in an object-oriented manner so that many modules depend on each other to achieve a task. The collection of modules in the bioperl-live repository consist of the core of the functionality of bioperl. Additionally auxiliary modules for creating graphical interfaces (bioperl-gui), persistent storage in RDMBS (bioperl-db), running and parsing the results from hundreds of bioinformatics applications (Run package), software to automate bioinformatic analyses (bioperl-pipeline) are all available as Git modules in our repository. The BioPerl toolkit provides a library of hundreds of routines for processing sequence, annotation, alignment, and sequence analysis reports. It often serves as a bridge between different computational biology applications assisting the user to construct analysis pipelines. This chapter illustrates how BioPerl facilitates tasks such as writing scripts summarizing information from BLAST reports or extracting key annotation details from a GenBank sequence record. BioPerl includes modules written by Sohel Merchant of the GO Consortium for parsing and manipulating OBO ontologies. Platform: Windows compatible, Mac OS X compatible, Linux compatible, Unix compatible
Proper citation: BioPerl (RRID:SCR_002989) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.